(Mis)Fitting: A Survey of Scaling Laws
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Margaret, Kudugunta, Sneha, Zettlemoyer, Luke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Slicing and Dicing: Configuring Optimal Mixtures of Experts
di: Li, Margaret, et al.
Pubblicazione: (2026)
di: Li, Margaret, et al.
Pubblicazione: (2026)
A Causal Lens for Evaluating Faithfulness Metrics
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
di: Chew, Robert, et al.
Pubblicazione: (2026)
di: Chew, Robert, et al.
Pubblicazione: (2026)
Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning
di: Cai, Hengrui, et al.
Pubblicazione: (2023)
di: Cai, Hengrui, et al.
Pubblicazione: (2023)
Text Rationalization for Robust Causal Effect Estimation
di: Zhang, Lijinghua, et al.
Pubblicazione: (2025)
di: Zhang, Lijinghua, et al.
Pubblicazione: (2025)
The Leaderboard Illusion
di: Singh, Shivalika, et al.
Pubblicazione: (2025)
di: Singh, Shivalika, et al.
Pubblicazione: (2025)
Majority of the Bests: Improving Best-of-N via Bootstrapping
di: Rakhsha, Amin, et al.
Pubblicazione: (2025)
di: Rakhsha, Amin, et al.
Pubblicazione: (2025)
Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation
di: Bhardwaj, Dhrupad, et al.
Pubblicazione: (2025)
di: Bhardwaj, Dhrupad, et al.
Pubblicazione: (2025)
RCT Rejection Sampling for Causal Estimation Evaluation
di: Keith, Katherine A., et al.
Pubblicazione: (2023)
di: Keith, Katherine A., et al.
Pubblicazione: (2023)
Evaluating Interventional Reasoning Capabilities of Large Language Models
di: Kasetty, Tejas, et al.
Pubblicazione: (2024)
di: Kasetty, Tejas, et al.
Pubblicazione: (2024)
AutoEval Done Right: Using Synthetic Data for Model Evaluation
di: Boyeau, Pierre, et al.
Pubblicazione: (2024)
di: Boyeau, Pierre, et al.
Pubblicazione: (2024)
Efficient Exploration for LLMs
di: Dwaracherla, Vikranth, et al.
Pubblicazione: (2024)
di: Dwaracherla, Vikranth, et al.
Pubblicazione: (2024)
ALCM: Autonomous LLM-Augmented Causal Discovery Framework
di: Khatibi, Elahe, et al.
Pubblicazione: (2024)
di: Khatibi, Elahe, et al.
Pubblicazione: (2024)
Adaptive Uncertainty Quantification for Generative AI
di: Kim, Jungeum, et al.
Pubblicazione: (2024)
di: Kim, Jungeum, et al.
Pubblicazione: (2024)
Industrial-Grade Smart Troubleshooting through Causal Technical Language Processing: a Proof of Concept
di: Trilla, Alexandre, et al.
Pubblicazione: (2024)
di: Trilla, Alexandre, et al.
Pubblicazione: (2024)
Propagation and Pitfalls: Reasoning-based Assessment of Knowledge Editing through Counterfactual Tasks
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
CLEAR: Can Language Models Really Understand Causal Graphs?
di: Chen, Sirui, et al.
Pubblicazione: (2024)
di: Chen, Sirui, et al.
Pubblicazione: (2024)
Uncertainty Quantification for Prior-Data Fitted Networks using Martingale Posteriors
di: Nagler, Thomas, et al.
Pubblicazione: (2025)
di: Nagler, Thomas, et al.
Pubblicazione: (2025)
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2024)
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2024)
Recycling the Web: A Method to Enhance Pre-training Data Quality and Quantity for Language Models
di: Nguyen, Thao, et al.
Pubblicazione: (2025)
di: Nguyen, Thao, et al.
Pubblicazione: (2025)
Removing Spurious Correlation from Neural Network Interpretations
di: Fotouhi, Milad, et al.
Pubblicazione: (2024)
di: Fotouhi, Milad, et al.
Pubblicazione: (2024)
Language Models as Causal Effect Generators
di: Bynum, Lucius E. J., et al.
Pubblicazione: (2024)
di: Bynum, Lucius E. J., et al.
Pubblicazione: (2024)
Discovering influential text using convolutional neural networks
di: Ayers, Megan, et al.
Pubblicazione: (2024)
di: Ayers, Megan, et al.
Pubblicazione: (2024)
Better Alignment with Instruction Back-and-Forth Translation
di: Nguyen, Thao, et al.
Pubblicazione: (2024)
di: Nguyen, Thao, et al.
Pubblicazione: (2024)
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation
di: Sarmah, Bhaskarjit, et al.
Pubblicazione: (2024)
di: Sarmah, Bhaskarjit, et al.
Pubblicazione: (2024)
Few-shot Personalization of LLMs with Mis-aligned Responses
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
Distillation Scaling Laws
di: Busbridge, Dan, et al.
Pubblicazione: (2025)
di: Busbridge, Dan, et al.
Pubblicazione: (2025)
P$^2$ Law: Scaling Law for Post-Training After Model Pruning
di: Chen, Xiaodong, et al.
Pubblicazione: (2024)
di: Chen, Xiaodong, et al.
Pubblicazione: (2024)
What Scales in Cross-Entropy Scaling Law?
di: Yan, Junxi, et al.
Pubblicazione: (2025)
di: Yan, Junxi, et al.
Pubblicazione: (2025)
Can Language Models Discover Scaling Laws?
di: Lin, Haowei, et al.
Pubblicazione: (2025)
di: Lin, Haowei, et al.
Pubblicazione: (2025)
Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy (short paper)
di: Dobariya, Om, et al.
Pubblicazione: (2025)
di: Dobariya, Om, et al.
Pubblicazione: (2025)
Scaling Retrieval-Based Language Models with a Trillion-Token Datastore
di: Shao, Rulin, et al.
Pubblicazione: (2024)
di: Shao, Rulin, et al.
Pubblicazione: (2024)
SILO Language Models: Isolating Legal Risk In a Nonparametric Datastore
di: Min, Sewon, et al.
Pubblicazione: (2023)
di: Min, Sewon, et al.
Pubblicazione: (2023)
MisSynth: Improving MISSCI Logical Fallacies Classification with Synthetic Data
di: Poliakov, Mykhailo, et al.
Pubblicazione: (2025)
di: Poliakov, Mykhailo, et al.
Pubblicazione: (2025)
Alignment Pretraining: AI Discourse Causes Self-Fulfilling (Mis)alignment
di: Tice, Cameron, et al.
Pubblicazione: (2026)
di: Tice, Cameron, et al.
Pubblicazione: (2026)
Theoretical Foundations of Scaling Law in Familial Models
di: Song, Huan, et al.
Pubblicazione: (2025)
di: Song, Huan, et al.
Pubblicazione: (2025)
Generative Adapter: Contextualizing Language Models in Parameters with A Single Forward Pass
di: Chen, Tong, et al.
Pubblicazione: (2024)
di: Chen, Tong, et al.
Pubblicazione: (2024)
A Hitchhiker's Guide to Scaling Law Estimation
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
Evaluation of Stress Detection as Time Series Events -- A Novel Window-Based F1-Metric
di: Skat-Rørdam, Harald Vilhelm, et al.
Pubblicazione: (2025)
di: Skat-Rørdam, Harald Vilhelm, et al.
Pubblicazione: (2025)
Fast Byte Latent Transformer
di: Kallini, Julie, et al.
Pubblicazione: (2026)
di: Kallini, Julie, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Slicing and Dicing: Configuring Optimal Mixtures of Experts
di: Li, Margaret, et al.
Pubblicazione: (2026) -
A Causal Lens for Evaluating Faithfulness Metrics
di: Zaman, Kerem, et al.
Pubblicazione: (2025) -
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
di: Chew, Robert, et al.
Pubblicazione: (2026) -
Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning
di: Cai, Hengrui, et al.
Pubblicazione: (2023) -
Text Rationalization for Robust Causal Effect Estimation
di: Zhang, Lijinghua, et al.
Pubblicazione: (2025)