(Mis)Fitting: A Survey of Scaling Laws
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Margaret, Kudugunta, Sneha, Zettlemoyer, Luke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Slicing and Dicing: Configuring Optimal Mixtures of Experts
von: Li, Margaret, et al.
Veröffentlicht: (2026)
von: Li, Margaret, et al.
Veröffentlicht: (2026)
A Causal Lens for Evaluating Faithfulness Metrics
von: Zaman, Kerem, et al.
Veröffentlicht: (2025)
von: Zaman, Kerem, et al.
Veröffentlicht: (2025)
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
von: Chew, Robert, et al.
Veröffentlicht: (2026)
von: Chew, Robert, et al.
Veröffentlicht: (2026)
Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning
von: Cai, Hengrui, et al.
Veröffentlicht: (2023)
von: Cai, Hengrui, et al.
Veröffentlicht: (2023)
Text Rationalization for Robust Causal Effect Estimation
von: Zhang, Lijinghua, et al.
Veröffentlicht: (2025)
von: Zhang, Lijinghua, et al.
Veröffentlicht: (2025)
The Leaderboard Illusion
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)
Majority of the Bests: Improving Best-of-N via Bootstrapping
von: Rakhsha, Amin, et al.
Veröffentlicht: (2025)
von: Rakhsha, Amin, et al.
Veröffentlicht: (2025)
Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation
von: Bhardwaj, Dhrupad, et al.
Veröffentlicht: (2025)
von: Bhardwaj, Dhrupad, et al.
Veröffentlicht: (2025)
RCT Rejection Sampling for Causal Estimation Evaluation
von: Keith, Katherine A., et al.
Veröffentlicht: (2023)
von: Keith, Katherine A., et al.
Veröffentlicht: (2023)
Evaluating Interventional Reasoning Capabilities of Large Language Models
von: Kasetty, Tejas, et al.
Veröffentlicht: (2024)
von: Kasetty, Tejas, et al.
Veröffentlicht: (2024)
AutoEval Done Right: Using Synthetic Data for Model Evaluation
von: Boyeau, Pierre, et al.
Veröffentlicht: (2024)
von: Boyeau, Pierre, et al.
Veröffentlicht: (2024)
Efficient Exploration for LLMs
von: Dwaracherla, Vikranth, et al.
Veröffentlicht: (2024)
von: Dwaracherla, Vikranth, et al.
Veröffentlicht: (2024)
ALCM: Autonomous LLM-Augmented Causal Discovery Framework
von: Khatibi, Elahe, et al.
Veröffentlicht: (2024)
von: Khatibi, Elahe, et al.
Veröffentlicht: (2024)
Adaptive Uncertainty Quantification for Generative AI
von: Kim, Jungeum, et al.
Veröffentlicht: (2024)
von: Kim, Jungeum, et al.
Veröffentlicht: (2024)
Industrial-Grade Smart Troubleshooting through Causal Technical Language Processing: a Proof of Concept
von: Trilla, Alexandre, et al.
Veröffentlicht: (2024)
von: Trilla, Alexandre, et al.
Veröffentlicht: (2024)
Propagation and Pitfalls: Reasoning-based Assessment of Knowledge Editing through Counterfactual Tasks
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
CLEAR: Can Language Models Really Understand Causal Graphs?
von: Chen, Sirui, et al.
Veröffentlicht: (2024)
von: Chen, Sirui, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for Prior-Data Fitted Networks using Martingale Posteriors
von: Nagler, Thomas, et al.
Veröffentlicht: (2025)
von: Nagler, Thomas, et al.
Veröffentlicht: (2025)
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
von: Limisiewicz, Tomasz, et al.
Veröffentlicht: (2024)
von: Limisiewicz, Tomasz, et al.
Veröffentlicht: (2024)
Recycling the Web: A Method to Enhance Pre-training Data Quality and Quantity for Language Models
von: Nguyen, Thao, et al.
Veröffentlicht: (2025)
von: Nguyen, Thao, et al.
Veröffentlicht: (2025)
Removing Spurious Correlation from Neural Network Interpretations
von: Fotouhi, Milad, et al.
Veröffentlicht: (2024)
von: Fotouhi, Milad, et al.
Veröffentlicht: (2024)
Language Models as Causal Effect Generators
von: Bynum, Lucius E. J., et al.
Veröffentlicht: (2024)
von: Bynum, Lucius E. J., et al.
Veröffentlicht: (2024)
Discovering influential text using convolutional neural networks
von: Ayers, Megan, et al.
Veröffentlicht: (2024)
von: Ayers, Megan, et al.
Veröffentlicht: (2024)
Better Alignment with Instruction Back-and-Forth Translation
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation
von: Sarmah, Bhaskarjit, et al.
Veröffentlicht: (2024)
von: Sarmah, Bhaskarjit, et al.
Veröffentlicht: (2024)
Few-shot Personalization of LLMs with Mis-aligned Responses
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
Distillation Scaling Laws
von: Busbridge, Dan, et al.
Veröffentlicht: (2025)
von: Busbridge, Dan, et al.
Veröffentlicht: (2025)
P$^2$ Law: Scaling Law for Post-Training After Model Pruning
von: Chen, Xiaodong, et al.
Veröffentlicht: (2024)
von: Chen, Xiaodong, et al.
Veröffentlicht: (2024)
What Scales in Cross-Entropy Scaling Law?
von: Yan, Junxi, et al.
Veröffentlicht: (2025)
von: Yan, Junxi, et al.
Veröffentlicht: (2025)
Can Language Models Discover Scaling Laws?
von: Lin, Haowei, et al.
Veröffentlicht: (2025)
von: Lin, Haowei, et al.
Veröffentlicht: (2025)
Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy (short paper)
von: Dobariya, Om, et al.
Veröffentlicht: (2025)
von: Dobariya, Om, et al.
Veröffentlicht: (2025)
Scaling Retrieval-Based Language Models with a Trillion-Token Datastore
von: Shao, Rulin, et al.
Veröffentlicht: (2024)
von: Shao, Rulin, et al.
Veröffentlicht: (2024)
SILO Language Models: Isolating Legal Risk In a Nonparametric Datastore
von: Min, Sewon, et al.
Veröffentlicht: (2023)
von: Min, Sewon, et al.
Veröffentlicht: (2023)
MisSynth: Improving MISSCI Logical Fallacies Classification with Synthetic Data
von: Poliakov, Mykhailo, et al.
Veröffentlicht: (2025)
von: Poliakov, Mykhailo, et al.
Veröffentlicht: (2025)
Alignment Pretraining: AI Discourse Causes Self-Fulfilling (Mis)alignment
von: Tice, Cameron, et al.
Veröffentlicht: (2026)
von: Tice, Cameron, et al.
Veröffentlicht: (2026)
Theoretical Foundations of Scaling Law in Familial Models
von: Song, Huan, et al.
Veröffentlicht: (2025)
von: Song, Huan, et al.
Veröffentlicht: (2025)
Generative Adapter: Contextualizing Language Models in Parameters with A Single Forward Pass
von: Chen, Tong, et al.
Veröffentlicht: (2024)
von: Chen, Tong, et al.
Veröffentlicht: (2024)
A Hitchhiker's Guide to Scaling Law Estimation
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
Evaluation of Stress Detection as Time Series Events -- A Novel Window-Based F1-Metric
von: Skat-Rørdam, Harald Vilhelm, et al.
Veröffentlicht: (2025)
von: Skat-Rørdam, Harald Vilhelm, et al.
Veröffentlicht: (2025)
Fast Byte Latent Transformer
von: Kallini, Julie, et al.
Veröffentlicht: (2026)
von: Kallini, Julie, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Slicing and Dicing: Configuring Optimal Mixtures of Experts
von: Li, Margaret, et al.
Veröffentlicht: (2026) -
A Causal Lens for Evaluating Faithfulness Metrics
von: Zaman, Kerem, et al.
Veröffentlicht: (2025) -
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
von: Chew, Robert, et al.
Veröffentlicht: (2026) -
Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning
von: Cai, Hengrui, et al.
Veröffentlicht: (2023) -
Text Rationalization for Robust Causal Effect Estimation
von: Zhang, Lijinghua, et al.
Veröffentlicht: (2025)