Comparison of Large Language Models for Deployment Requirements
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yaman, Alper, Schwab, Jannik, Nitsche, Christof, Sinha, Abhirup, Huber, Marco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Study of Privacy-preserving Language Modeling Approaches
von: Saha, Pritilata, et al.
Veröffentlicht: (2025)
von: Saha, Pritilata, et al.
Veröffentlicht: (2025)
Current State in Privacy-Preserving Text Preprocessing for Domain-Agnostic NLP
von: Sinha, Abhirup, et al.
Veröffentlicht: (2025)
von: Sinha, Abhirup, et al.
Veröffentlicht: (2025)
Empirical Analysis of the Effect of Context in the Task of Automated Essay Scoring in Transformer-Based Models
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
The Effect of Language Diversity When Fine-Tuning Large Language Models for Translation
von: Stap, David, et al.
Veröffentlicht: (2025)
von: Stap, David, et al.
Veröffentlicht: (2025)
Measuring and Improving Persuasiveness of Large Language Models
von: Singh, Somesh, et al.
Veröffentlicht: (2024)
von: Singh, Somesh, et al.
Veröffentlicht: (2024)
QA-prompting: Improving Summarization with Large Language Models using Question-Answering
von: Sinha, Neelabh
Veröffentlicht: (2025)
von: Sinha, Neelabh
Veröffentlicht: (2025)
Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages
von: Brinkmann, Jannik, et al.
Veröffentlicht: (2025)
von: Brinkmann, Jannik, et al.
Veröffentlicht: (2025)
Towards Understanding Counseling Conversations: Domain Knowledge and Large Language Models
von: Lee, Younghun, et al.
Veröffentlicht: (2024)
von: Lee, Younghun, et al.
Veröffentlicht: (2024)
Are Large Language Models Economically Viable for Industry Deployment?
von: Mohammad, Abdullah, et al.
Veröffentlicht: (2026)
von: Mohammad, Abdullah, et al.
Veröffentlicht: (2026)
ALN-P3: Unified Language Alignment for Perception, Prediction, and Planning in Autonomous Driving
von: Ma, Yunsheng, et al.
Veröffentlicht: (2025)
von: Ma, Yunsheng, et al.
Veröffentlicht: (2025)
Does Model Size Matter? A Comparison of Small and Large Language Models for Requirements Classification
von: Zadenoori, Mohammad Amin, et al.
Veröffentlicht: (2025)
von: Zadenoori, Mohammad Amin, et al.
Veröffentlicht: (2025)
LSAQ: Layer-Specific Adaptive Quantization for Large Language Model Deployment
von: Zeng, Binrui, et al.
Veröffentlicht: (2024)
von: Zeng, Binrui, et al.
Veröffentlicht: (2024)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
von: Hida, Rem, et al.
Veröffentlicht: (2024)
von: Hida, Rem, et al.
Veröffentlicht: (2024)
LLMCBench: Benchmarking Large Language Model Compression for Efficient Deployment
von: Yang, Ge, et al.
Veröffentlicht: (2024)
von: Yang, Ge, et al.
Veröffentlicht: (2024)
Deploying Multi-task Online Server with Large Language Model
von: Qu, Yincen, et al.
Veröffentlicht: (2024)
von: Qu, Yincen, et al.
Veröffentlicht: (2024)
CLEAR: A Comprehensive Linguistic Evaluation of Argument Rewriting by Large Language Models
von: Huber, Thomas, et al.
Veröffentlicht: (2025)
von: Huber, Thomas, et al.
Veröffentlicht: (2025)
Fine-Tuning Large Language Models to Appropriately Abstain with Semantic Entropy
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024)
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024)
Entropy in Large Language Models
von: Scharringhausen, Marco
Veröffentlicht: (2026)
von: Scharringhausen, Marco
Veröffentlicht: (2026)
Leveraging Large Language Models for Risk Assessment in Hyperconnected Logistic Hub Network Deployment
von: Quan, Yinzhu, et al.
Veröffentlicht: (2025)
von: Quan, Yinzhu, et al.
Veröffentlicht: (2025)
Transformer-Lite: High-efficiency Deployment of Large Language Models on Mobile Phone GPUs
von: Li, Luchang, et al.
Veröffentlicht: (2024)
von: Li, Luchang, et al.
Veröffentlicht: (2024)
Evaluating Nuanced Bias in Large Language Model Free Response Answers
von: Healey, Jennifer, et al.
Veröffentlicht: (2024)
von: Healey, Jennifer, et al.
Veröffentlicht: (2024)
TIM: Teaching Large Language Models to Translate with Comparison
von: Zeng, Jiali, et al.
Veröffentlicht: (2023)
von: Zeng, Jiali, et al.
Veröffentlicht: (2023)
Code Comparison Tuning for Code Large Language Models
von: Jiang, Yufan, et al.
Veröffentlicht: (2024)
von: Jiang, Yufan, et al.
Veröffentlicht: (2024)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
von: Sinha, Neelabh, et al.
Veröffentlicht: (2024)
von: Sinha, Neelabh, et al.
Veröffentlicht: (2024)
TrackList: Tracing Back Query Linguistic Diversity for Head and Tail Knowledge in Open Large Language Models
von: Buhnila, Ioana, et al.
Veröffentlicht: (2025)
von: Buhnila, Ioana, et al.
Veröffentlicht: (2025)
Eka-Eval: An Evaluation Framework for Low-Resource Multilingual Large Language Models
von: Sinha, Samridhi Raj, et al.
Veröffentlicht: (2025)
von: Sinha, Samridhi Raj, et al.
Veröffentlicht: (2025)
LAVA: Language Model Assisted Verbal Autopsy for Cause-of-Death Determination
von: Chen, Yiqun T., et al.
Veröffentlicht: (2025)
von: Chen, Yiqun T., et al.
Veröffentlicht: (2025)
S2D: Sorted Speculative Decoding For More Efficient Deployment of Nested Large Language Models
von: Kavehzadeh, Parsa, et al.
Veröffentlicht: (2024)
von: Kavehzadeh, Parsa, et al.
Veröffentlicht: (2024)
Kiki or Bouba? Sound Symbolism in Vision-and-Language Models
von: Alper, Morris, et al.
Veröffentlicht: (2023)
von: Alper, Morris, et al.
Veröffentlicht: (2023)
ApiQ: Finetuning of 2-Bit Quantized Large Language Model
von: Liao, Baohao, et al.
Veröffentlicht: (2024)
von: Liao, Baohao, et al.
Veröffentlicht: (2024)
Performance Comparison of Large Language Models on Advanced Calculus Problems
von: Moon, In Hak
Veröffentlicht: (2025)
von: Moon, In Hak
Veröffentlicht: (2025)
Analyzing Narrative Processing in Large Language Models (LLMs): Using GPT4 to test BERT
von: Krauss, Patrick, et al.
Veröffentlicht: (2024)
von: Krauss, Patrick, et al.
Veröffentlicht: (2024)
Language Mixing in Reasoning Language Models: Patterns, Impact, and Internal Causes
von: Wang, Mingyang, et al.
Veröffentlicht: (2025)
von: Wang, Mingyang, et al.
Veröffentlicht: (2025)
Measuring Representation Robustness in Large Language Models for Geometry
von: Jawandhia, Vedant, et al.
Veröffentlicht: (2026)
von: Jawandhia, Vedant, et al.
Veröffentlicht: (2026)
Efficient Deployment of Large Language Models on Resource-constrained Devices
von: Yao, Zhiwei, et al.
Veröffentlicht: (2025)
von: Yao, Zhiwei, et al.
Veröffentlicht: (2025)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
Risks, Causes, and Mitigations of Widespread Deployments of Large Language Models (LLMs): A Survey
von: Sakib, Md Nazmus, et al.
Veröffentlicht: (2024)
von: Sakib, Md Nazmus, et al.
Veröffentlicht: (2024)
Do Language Models Reason Across Languages?
von: Meng, Yan, et al.
Veröffentlicht: (2026)
von: Meng, Yan, et al.
Veröffentlicht: (2026)
The SIFo Benchmark: Investigating the Sequential Instruction Following Ability of Large Language Models
von: Chen, Xinyi, et al.
Veröffentlicht: (2024)
von: Chen, Xinyi, et al.
Veröffentlicht: (2024)
Collaborative Distillation Strategies for Parameter-Efficient Language Model Deployment
von: Meng, Xiandong, et al.
Veröffentlicht: (2025)
von: Meng, Xiandong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Study of Privacy-preserving Language Modeling Approaches
von: Saha, Pritilata, et al.
Veröffentlicht: (2025) -
Current State in Privacy-Preserving Text Preprocessing for Domain-Agnostic NLP
von: Sinha, Abhirup, et al.
Veröffentlicht: (2025) -
Empirical Analysis of the Effect of Context in the Task of Automated Essay Scoring in Transformer-Based Models
von: Chakravarty, Abhirup
Veröffentlicht: (2025) -
The Effect of Language Diversity When Fine-Tuning Large Language Models for Translation
von: Stap, David, et al.
Veröffentlicht: (2025) -
Measuring and Improving Persuasiveness of Large Language Models
von: Singh, Somesh, et al.
Veröffentlicht: (2024)