Chinchilla Scaling: A replication attempt
Fuente:
arXiv
Saved in:
| Main Authors: | Besiroglu, Tamay, Erdil, Ege, Barnett, Matthew, You, Josh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Explosive growth from AI automation: A review of the arguments
by: Erdil, Ege, et al.
Published: (2023)
by: Erdil, Ege, et al.
Published: (2023)
Algorithmic progress in language models
by: Ho, Anson, et al.
Published: (2024)
by: Ho, Anson, et al.
Published: (2024)
Estimating Idea Production: A Methodological Survey
by: Erdil, Ege, et al.
Published: (2024)
by: Erdil, Ege, et al.
Published: (2024)
Will we run out of data? Limits of LLM scaling based on human-generated data
by: Villalobos, Pablo, et al.
Published: (2022)
by: Villalobos, Pablo, et al.
Published: (2022)
Data movement limits to frontier model training
by: Erdil, Ege, et al.
Published: (2024)
by: Erdil, Ege, et al.
Published: (2024)
The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?
by: Besiroglu, Tamay, et al.
Published: (2024)
by: Besiroglu, Tamay, et al.
Published: (2024)
Simulating Policy Impacts: Developing a Generative Scenario Writing Method to Evaluate the Perceived Effects of Regulation
by: Barnett, Julia, et al.
Published: (2024)
by: Barnett, Julia, et al.
Published: (2024)
Large-Scale Constraint Generation -- Can LLMs Parse Hundreds of Constraints?
by: Boffa, Matteo, et al.
Published: (2025)
by: Boffa, Matteo, et al.
Published: (2025)
CiteBART: Learning to Generate Citations for Local Citation Recommendation
by: Çelik, Ege Yiğit, et al.
Published: (2024)
by: Çelik, Ege Yiğit, et al.
Published: (2024)
GATE: An Integrated Assessment Model for AI Automation
by: Erdil, Ege, et al.
Published: (2025)
by: Erdil, Ege, et al.
Published: (2025)
Can LLM-Generated Textual Explanations Enhance Model Classification Performance? An Empirical Study
by: Dhaini, Mahdi, et al.
Published: (2025)
by: Dhaini, Mahdi, et al.
Published: (2025)
Overcoming Data Scarcity in Generative Language Modelling for Low-Resource Languages: A Systematic Review
by: McGiff, Josh, et al.
Published: (2025)
by: McGiff, Josh, et al.
Published: (2025)
Fine-Tuning or Fine-Failing? Debunking Performance Myths in Large Language Models
by: Barnett, Scott, et al.
Published: (2024)
by: Barnett, Scott, et al.
Published: (2024)
Large Scale Transfer Learning for Tabular Data via Language Modeling
by: Gardner, Josh, et al.
Published: (2024)
by: Gardner, Josh, et al.
Published: (2024)
Rewarding Doubt: A Reinforcement Learning Approach to Calibrated Confidence Expression of Large Language Models
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Scaling, Simplification, and Adaptation: Lessons from Pretraining on Machine-Translated Text
by: Velasco, Dan John, et al.
Published: (2025)
by: Velasco, Dan John, et al.
Published: (2025)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
by: Rozner, Josh, et al.
Published: (2021)
by: Rozner, Josh, et al.
Published: (2021)
Gender Bias in Explainability: Investigating Performance Disparity in Post-hoc Methods
by: Dhaini, Mahdi, et al.
Published: (2025)
by: Dhaini, Mahdi, et al.
Published: (2025)
Guided Persona-based AI Surveys: Can we replicate personal mobility preferences at scale using LLMs?
by: Tzachristas, Ioannis, et al.
Published: (2025)
by: Tzachristas, Ioannis, et al.
Published: (2025)
Digestion Algorithm in Hierarchical Symbolic Forests: A Fast Text Normalization Algorithm and Semantic Parsing Framework for Specific Scenarios and Lightweight Deployment
by: You, Kevin
Published: (2024)
by: You, Kevin
Published: (2024)
Frontier AI systems have surpassed the self-replicating red line
by: Pan, Xudong, et al.
Published: (2024)
by: Pan, Xudong, et al.
Published: (2024)
Inference economics of language models
by: Erdil, Ege
Published: (2025)
by: Erdil, Ege
Published: (2025)
DermaBench: A Clinician-Annotated Benchmark Dataset for Dermatology Visual Question Answering and Reasoning
by: Yilmaz, Abdurrahim, et al.
Published: (2026)
by: Yilmaz, Abdurrahim, et al.
Published: (2026)
Using Language Models to Disambiguate Lexical Choices in Translation
by: Barua, Josh, et al.
Published: (2024)
by: Barua, Josh, et al.
Published: (2024)
Learning to Think Like a Cartoon Captionist: Incongruity-Resolution Supervision for Multimodal Humor Understanding
by: Vural, Hatice Merve, et al.
Published: (2026)
by: Vural, Hatice Merve, et al.
Published: (2026)
Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement Learning
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Evaluating Deduplication Techniques for Economic Research Paper Titles with a Focus on Semantic Similarity using NLP and LLMs
by: You, Doohee, et al.
Published: (2024)
by: You, Doohee, et al.
Published: (2024)
Evolving Prompts In-Context: An Open-ended, Self-replicating Perspective
by: Wang, Jianyu, et al.
Published: (2025)
by: Wang, Jianyu, et al.
Published: (2025)
PhonemeFake: Redefining Deepfake Realism with Language-Driven Segmental Manipulation and Adaptive Bilevel Detection
by: Baser, Oguzhan, et al.
Published: (2025)
by: Baser, Oguzhan, et al.
Published: (2025)
Transparent Reference-free Automated Evaluation of Open-Ended User Survey Responses
by: An, Subin, et al.
Published: (2025)
by: An, Subin, et al.
Published: (2025)
A Proactive EMR Assistant for Doctor-Patient Dialogue: Streaming ASR, Belief Stabilization, and Preliminary Controlled Evaluation
by: Pan, Zhenhai, et al.
Published: (2026)
by: Pan, Zhenhai, et al.
Published: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
by: You, Runyang, et al.
Published: (2025)
by: You, Runyang, et al.
Published: (2025)
Reconciling Kaplan and Chinchilla Scaling Laws
by: Pearce, Tim, et al.
Published: (2024)
by: Pearce, Tim, et al.
Published: (2024)
Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model
by: Ling Team, et al.
Published: (2025)
by: Ling Team, et al.
Published: (2025)
Scaling Up, Speeding Up: A Benchmark of Speculative Decoding for Efficient LLM Test-Time Scaling
by: Sun, Shengyin, et al.
Published: (2025)
by: Sun, Shengyin, et al.
Published: (2025)
FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI
by: Glazer, Elliot, et al.
Published: (2024)
by: Glazer, Elliot, et al.
Published: (2024)
Classification is a RAG problem: A case study on hate speech detection
by: Willats, Richard, et al.
Published: (2025)
by: Willats, Richard, et al.
Published: (2025)
UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling
by: Huang, Kaiyu, et al.
Published: (2026)
by: Huang, Kaiyu, et al.
Published: (2026)
Efficient Contextual LLM Cascades through Budget-Constrained Policy Learning
by: Zhang, Xuechen, et al.
Published: (2024)
by: Zhang, Xuechen, et al.
Published: (2024)
Codenames as a Benchmark for Large Language Models
by: Stephenson, Matthew, et al.
Published: (2024)
by: Stephenson, Matthew, et al.
Published: (2024)
Similar Items
-
Explosive growth from AI automation: A review of the arguments
by: Erdil, Ege, et al.
Published: (2023) -
Algorithmic progress in language models
by: Ho, Anson, et al.
Published: (2024) -
Estimating Idea Production: A Methodological Survey
by: Erdil, Ege, et al.
Published: (2024) -
Will we run out of data? Limits of LLM scaling based on human-generated data
by: Villalobos, Pablo, et al.
Published: (2022) -
Data movement limits to frontier model training
by: Erdil, Ege, et al.
Published: (2024)