The Impact of Hyperparameters on Large Language Model Inference Performance: An Evaluation of vLLM and HuggingFace Pipelines
Fuente:
arXiv
Salvato in:
| Autore principale: | Martinez, Matias |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
di: Taraghi, Mina, et al.
Pubblicazione: (2024)
di: Taraghi, Mina, et al.
Pubblicazione: (2024)
Comparative Analysis of Large Language Model Inference Serving Systems: A Performance Study of vLLM and HuggingFace TGI
di: Kolluru, Saicharan
Pubblicazione: (2025)
di: Kolluru, Saicharan
Pubblicazione: (2025)
Lessons Learned from Mining the Hugging Face Repository
di: Castaño, Joel, et al.
Pubblicazione: (2024)
di: Castaño, Joel, et al.
Pubblicazione: (2024)
Cataloguing Hugging Face Models to Software Engineering Activities: Automation and Findings
di: González, Alexandra, et al.
Pubblicazione: (2025)
di: González, Alexandra, et al.
Pubblicazione: (2025)
OptLLM: Optimal Assignment of Queries to Large Language Models
di: Liu, Yueyue, et al.
Pubblicazione: (2024)
di: Liu, Yueyue, et al.
Pubblicazione: (2024)
Evaluation and Improvement of Fault Detection for Large Language Models
di: Hu, Qiang, et al.
Pubblicazione: (2024)
di: Hu, Qiang, et al.
Pubblicazione: (2024)
Analyzing the Evolution and Maintenance of ML Models on Hugging Face
di: Castaño, Joel, et al.
Pubblicazione: (2023)
di: Castaño, Joel, et al.
Pubblicazione: (2023)
CodeJudge: Evaluating Code Generation with Large Language Models
di: Tong, Weixi, et al.
Pubblicazione: (2024)
di: Tong, Weixi, et al.
Pubblicazione: (2024)
Evaluating the Generalization Capabilities of Large Language Models on Code Reasoning
di: Yang, Rem, et al.
Pubblicazione: (2025)
di: Yang, Rem, et al.
Pubblicazione: (2025)
Evaluating the Process Modeling Abilities of Large Language Models -- Preliminary Foundations and Results
di: Fettke, Peter, et al.
Pubblicazione: (2025)
di: Fettke, Peter, et al.
Pubblicazione: (2025)
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
di: Jain, Naman, et al.
Pubblicazione: (2024)
di: Jain, Naman, et al.
Pubblicazione: (2024)
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences
di: Weyssow, Martin, et al.
Pubblicazione: (2024)
di: Weyssow, Martin, et al.
Pubblicazione: (2024)
Evaluating Language Models for Efficient Code Generation
di: Liu, Jiawei, et al.
Pubblicazione: (2024)
di: Liu, Jiawei, et al.
Pubblicazione: (2024)
Pimp My LLM: Leveraging Variability Modeling to Tune Inference Hyperparameters
di: Zine, Nada, et al.
Pubblicazione: (2026)
di: Zine, Nada, et al.
Pubblicazione: (2026)
vLLM Hook v0: A Plug-in for Programming Model Internals on vLLM
di: Ko, Ching-Yun, et al.
Pubblicazione: (2026)
di: Ko, Ching-Yun, et al.
Pubblicazione: (2026)
Quantifying Contamination in Evaluating Code Generation Capabilities of Language Models
di: Riddell, Martin, et al.
Pubblicazione: (2024)
di: Riddell, Martin, et al.
Pubblicazione: (2024)
LLMSYS-HPOBench: Hyperparameter Optimization Benchmark Suite for Real-World LLM Systems
di: Wu, Siyu, et al.
Pubblicazione: (2026)
di: Wu, Siyu, et al.
Pubblicazione: (2026)
SODBench: A Large Language Model Approach to Documenting Spreadsheet Operations
di: Indika, Amila, et al.
Pubblicazione: (2025)
di: Indika, Amila, et al.
Pubblicazione: (2025)
Exploring Large Language Models for Translating Romanian Computational Problems into English
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
di: Petrukha, Ivan, et al.
Pubblicazione: (2025)
di: Petrukha, Ivan, et al.
Pubblicazione: (2025)
The Impact of Fine-tuning Large Language Models on Automated Program Repair
di: Macháček, Roman, et al.
Pubblicazione: (2025)
di: Macháček, Roman, et al.
Pubblicazione: (2025)
An Empirical Analysis of Machine Learning Model and Dataset Documentation, Supply Chain, and Licensing Challenges on Hugging Face
di: Stalnaker, Trevor, et al.
Pubblicazione: (2025)
di: Stalnaker, Trevor, et al.
Pubblicazione: (2025)
Lessons from the Use of Natural Language Inference (NLI) in Requirements Engineering Tasks
di: Fazelnia, Mohamad, et al.
Pubblicazione: (2024)
di: Fazelnia, Mohamad, et al.
Pubblicazione: (2024)
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation
di: Haider, Md. Asif, et al.
Pubblicazione: (2024)
di: Haider, Md. Asif, et al.
Pubblicazione: (2024)
Suggesting Code Edits in Interactive Machine Learning Notebooks Using Large Language Models
di: Jin, Bihui, et al.
Pubblicazione: (2025)
di: Jin, Bihui, et al.
Pubblicazione: (2025)
Exploring Parameter-Efficient Fine-Tuning Techniques for Code Generation with Large Language Models
di: Weyssow, Martin, et al.
Pubblicazione: (2023)
di: Weyssow, Martin, et al.
Pubblicazione: (2023)
CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation
di: Peng, Jinjun, et al.
Pubblicazione: (2025)
di: Peng, Jinjun, et al.
Pubblicazione: (2025)
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
di: Guo, Daya, et al.
Pubblicazione: (2024)
di: Guo, Daya, et al.
Pubblicazione: (2024)
Uncertainty Awareness of Large Language Models Under Code Distribution Shifts: A Benchmark Study
di: Li, Yufei, et al.
Pubblicazione: (2024)
di: Li, Yufei, et al.
Pubblicazione: (2024)
EVALOOOP: A Self-Consistency-Centered Framework for Assessing Large Language Model Robustness in Programming
di: Fang, Sen, et al.
Pubblicazione: (2025)
di: Fang, Sen, et al.
Pubblicazione: (2025)
GLLM: Self-Corrective G-Code Generation using Large Language Models with User Feedback
di: Abdelaal, Mohamed, et al.
Pubblicazione: (2025)
di: Abdelaal, Mohamed, et al.
Pubblicazione: (2025)
Optimizing Case-Based Reasoning System for Functional Test Script Generation with Large Language Models
di: Guo, Siyuan, et al.
Pubblicazione: (2025)
di: Guo, Siyuan, et al.
Pubblicazione: (2025)
Assessing the Latent Automated Program Repair Capabilities of Large Language Models using Round-Trip Translation
di: Ruiz, Fernando Vallecillos, et al.
Pubblicazione: (2024)
di: Ruiz, Fernando Vallecillos, et al.
Pubblicazione: (2024)
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
di: Guo, Jiawei, et al.
Pubblicazione: (2024)
di: Guo, Jiawei, et al.
Pubblicazione: (2024)
Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy
di: Dong, Yihong, et al.
Pubblicazione: (2026)
di: Dong, Yihong, et al.
Pubblicazione: (2026)
TAROT: Test-driven and Capability-adaptive Curriculum Reinforcement Fine-tuning for Code Generation with Large Language Models
di: Park, Chansung, et al.
Pubblicazione: (2026)
di: Park, Chansung, et al.
Pubblicazione: (2026)
WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning
di: Jiang, Juyong, et al.
Pubblicazione: (2026)
di: Jiang, Juyong, et al.
Pubblicazione: (2026)
NExT: Teaching Large Language Models to Reason about Code Execution
di: Ni, Ansong, et al.
Pubblicazione: (2024)
di: Ni, Ansong, et al.
Pubblicazione: (2024)
GrowthHacker: Automated Off-Policy Evaluation Optimization Using Code-Modifying LLM Agents
di: Wu, Jie JW, et al.
Pubblicazione: (2025)
di: Wu, Jie JW, et al.
Pubblicazione: (2025)
Can OpenSource beat ChatGPT? -- A Comparative Study of Large Language Models for Text-to-Code Generation
di: Mayer, Luis, et al.
Pubblicazione: (2024)
di: Mayer, Luis, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
di: Taraghi, Mina, et al.
Pubblicazione: (2024) -
Comparative Analysis of Large Language Model Inference Serving Systems: A Performance Study of vLLM and HuggingFace TGI
di: Kolluru, Saicharan
Pubblicazione: (2025) -
Lessons Learned from Mining the Hugging Face Repository
di: Castaño, Joel, et al.
Pubblicazione: (2024) -
Cataloguing Hugging Face Models to Software Engineering Activities: Automation and Findings
di: González, Alexandra, et al.
Pubblicazione: (2025) -
OptLLM: Optimal Assignment of Queries to Large Language Models
di: Liu, Yueyue, et al.
Pubblicazione: (2024)