Performance Law of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Chuhan, Tang, Ruiming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Entropy Law: The Story Behind Data Compression and LLM Performance
von: Yin, Mingjia, et al.
Veröffentlicht: (2024)
von: Yin, Mingjia, et al.
Veröffentlicht: (2024)
Scaling Laws for Downstream Task Performance of Large Language Models
von: Isik, Berivan, et al.
Veröffentlicht: (2024)
von: Isik, Berivan, et al.
Veröffentlicht: (2024)
Exploring Scaling Laws for Local SGD in Large Language Model Training
von: He, Qiaozhi, et al.
Veröffentlicht: (2024)
von: He, Qiaozhi, et al.
Veröffentlicht: (2024)
Scaling Laws for Discriminative Classification in Large Language Models
von: Wyatte, Dean, et al.
Veröffentlicht: (2024)
von: Wyatte, Dean, et al.
Veröffentlicht: (2024)
Scaling Laws for Post Training Quantized Large Language Models
von: Xu, Zifei, et al.
Veröffentlicht: (2024)
von: Xu, Zifei, et al.
Veröffentlicht: (2024)
LawLLM: Law Large Language Model for the US Legal System
von: Shu, Dong, et al.
Veröffentlicht: (2024)
von: Shu, Dong, et al.
Veröffentlicht: (2024)
Observational Scaling Laws and the Predictability of Language Model Performance
von: Ruan, Yangjun, et al.
Veröffentlicht: (2024)
von: Ruan, Yangjun, et al.
Veröffentlicht: (2024)
Latent Performance Profiling of Large Language Models
von: Chakraborty, Tanmoy, et al.
Veröffentlicht: (2026)
von: Chakraborty, Tanmoy, et al.
Veröffentlicht: (2026)
Scaling Law for Language Models Training Considering Batch Size
von: Shuai, Xian, et al.
Veröffentlicht: (2024)
von: Shuai, Xian, et al.
Veröffentlicht: (2024)
Power-Law Decay Loss for Large Language Model Finetuning: A Theory Perspective
von: Shao, Jintian
Veröffentlicht: (2025)
von: Shao, Jintian
Veröffentlicht: (2025)
Scaling Laws for Multilingual Language Models
von: He, Yifei, et al.
Veröffentlicht: (2024)
von: He, Yifei, et al.
Veröffentlicht: (2024)
Parallel Scaling Law for Language Models
von: Chen, Mouxiang, et al.
Veröffentlicht: (2025)
von: Chen, Mouxiang, et al.
Veröffentlicht: (2025)
Long Context RAG Performance of Large Language Models
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
Uncovering Scaling Laws for Large Language Models via Inverse Problems
von: Verma, Arun, et al.
Veröffentlicht: (2025)
von: Verma, Arun, et al.
Veröffentlicht: (2025)
Towards Modeling Learner Performance with Large Language Models
von: Neshaei, Seyed Parsa, et al.
Veröffentlicht: (2024)
von: Neshaei, Seyed Parsa, et al.
Veröffentlicht: (2024)
A Law of Next-Token Prediction in Large Language Models
von: He, Hangfeng, et al.
Veröffentlicht: (2024)
von: He, Hangfeng, et al.
Veröffentlicht: (2024)
HUKUKBERT: Domain-Specific Language Model for Turkish Law
von: Öztürk, Mehmet Utku, et al.
Veröffentlicht: (2026)
von: Öztürk, Mehmet Utku, et al.
Veröffentlicht: (2026)
Scaling Laws for Upcycling Mixture-of-Experts Language Models
von: Liew, Seng Pei, et al.
Veröffentlicht: (2025)
von: Liew, Seng Pei, et al.
Veröffentlicht: (2025)
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
Evaluating the Performance of Large Language Models for SDG Mapping (Technical Report)
von: Yin, Hui, et al.
Veröffentlicht: (2024)
von: Yin, Hui, et al.
Veröffentlicht: (2024)
Provable Scaling Laws for the Test-Time Compute of Large Language Models
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance
von: Ye, Jiasheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiasheng, et al.
Veröffentlicht: (2024)
GreedLlama: Performance of Financial Value-Aligned Large Language Models in Moral Reasoning
von: Yu, Jeffy, et al.
Veröffentlicht: (2024)
von: Yu, Jeffy, et al.
Veröffentlicht: (2024)
A Survey on Mixture of Experts in Large Language Models
von: Cai, Weilin, et al.
Veröffentlicht: (2024)
von: Cai, Weilin, et al.
Veröffentlicht: (2024)
DarwinLM: Evolutionary Structured Pruning of Large Language Models
von: Tang, Shengkun, et al.
Veröffentlicht: (2025)
von: Tang, Shengkun, et al.
Veröffentlicht: (2025)
Multimodal Large Language Models for Medicine: A Comprehensive Survey
von: Ye, Jiarui, et al.
Veröffentlicht: (2025)
von: Ye, Jiarui, et al.
Veröffentlicht: (2025)
Selecting Large Language Model to Fine-tune via Rectified Scaling Law
von: Lin, Haowei, et al.
Veröffentlicht: (2024)
von: Lin, Haowei, et al.
Veröffentlicht: (2024)
CEQuest: Benchmarking Large Language Models for Construction Estimation
von: Wu, Yanzhao, et al.
Veröffentlicht: (2025)
von: Wu, Yanzhao, et al.
Veröffentlicht: (2025)
FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models
von: Mateega, Spencer, et al.
Veröffentlicht: (2025)
von: Mateega, Spencer, et al.
Veröffentlicht: (2025)
Collaborative Performance Prediction for Large Language Models
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2024)
Comparative Performance Evaluation of Large Language Models for Extracting Molecular Interactions and Pathway Knowledge
von: Park, Gilchan, et al.
Veröffentlicht: (2023)
von: Park, Gilchan, et al.
Veröffentlicht: (2023)
Exploring Large Language Models for Financial Applications: Techniques, Performance, and Challenges with FinMA
von: Djagba, Prudence, et al.
Veröffentlicht: (2025)
von: Djagba, Prudence, et al.
Veröffentlicht: (2025)
Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws
von: Sardana, Nikhil, et al.
Veröffentlicht: (2023)
von: Sardana, Nikhil, et al.
Veröffentlicht: (2023)
Large Language Models to Diffusion Finetuning
von: Cetin, Edoardo, et al.
Veröffentlicht: (2025)
von: Cetin, Edoardo, et al.
Veröffentlicht: (2025)
Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models
von: Zhou, Chenxi, et al.
Veröffentlicht: (2025)
von: Zhou, Chenxi, et al.
Veröffentlicht: (2025)
Scaling Laws for Forgetting When Fine-Tuning Large Language Models
von: Kalajdzievski, Damjan
Veröffentlicht: (2024)
von: Kalajdzievski, Damjan
Veröffentlicht: (2024)
Sparsing Law: Towards Large Language Models with Greater Activation Sparsity
von: Luo, Yuqi, et al.
Veröffentlicht: (2024)
von: Luo, Yuqi, et al.
Veröffentlicht: (2024)
PocketLLM: Ultimate Compression of Large Language Models via Meta Networks
von: Tian, Ye, et al.
Veröffentlicht: (2025)
von: Tian, Ye, et al.
Veröffentlicht: (2025)
Zero-Shot Performance Prediction for Probabilistic Scaling Laws
von: Schram, Viktoria, et al.
Veröffentlicht: (2025)
von: Schram, Viktoria, et al.
Veröffentlicht: (2025)
Query Performance Explanation through Large Language Model for HTAP Systems
von: Xiu, Haibo, et al.
Veröffentlicht: (2024)
von: Xiu, Haibo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Entropy Law: The Story Behind Data Compression and LLM Performance
von: Yin, Mingjia, et al.
Veröffentlicht: (2024) -
Scaling Laws for Downstream Task Performance of Large Language Models
von: Isik, Berivan, et al.
Veröffentlicht: (2024) -
Exploring Scaling Laws for Local SGD in Large Language Model Training
von: He, Qiaozhi, et al.
Veröffentlicht: (2024) -
Scaling Laws for Discriminative Classification in Large Language Models
von: Wyatte, Dean, et al.
Veröffentlicht: (2024) -
Scaling Laws for Post Training Quantized Large Language Models
von: Xu, Zifei, et al.
Veröffentlicht: (2024)