Algorithmic progress in language models
Fuente:
arXiv
Saved in:
| Main Authors: | Ho, Anson, Besiroglu, Tamay, Erdil, Ege, Owen, David, Rahman, Robi, Guo, Zifan Carl, Atkinson, David, Thompson, Neil, Sevilla, Jaime |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Chinchilla Scaling: A replication attempt
by: Besiroglu, Tamay, et al.
Published: (2024)
by: Besiroglu, Tamay, et al.
Published: (2024)
Estimating Idea Production: A Methodological Survey
by: Erdil, Ege, et al.
Published: (2024)
by: Erdil, Ege, et al.
Published: (2024)
Explosive growth from AI automation: A review of the arguments
by: Erdil, Ege, et al.
Published: (2023)
by: Erdil, Ege, et al.
Published: (2023)
Will we run out of data? Limits of LLM scaling based on human-generated data
by: Villalobos, Pablo, et al.
Published: (2022)
by: Villalobos, Pablo, et al.
Published: (2022)
GATE: An Integrated Assessment Model for AI Automation
by: Erdil, Ege, et al.
Published: (2025)
by: Erdil, Ege, et al.
Published: (2025)
The rising costs of training frontier AI models
by: Cottier, Ben, et al.
Published: (2024)
by: Cottier, Ben, et al.
Published: (2024)
Data movement limits to frontier model training
by: Erdil, Ege, et al.
Published: (2024)
by: Erdil, Ege, et al.
Published: (2024)
The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?
by: Besiroglu, Tamay, et al.
Published: (2024)
by: Besiroglu, Tamay, et al.
Published: (2024)
Inference economics of language models
by: Erdil, Ege
Published: (2025)
by: Erdil, Ege
Published: (2025)
Does Distributed Training Undermine Compute Governance?
by: Rahman, Robi
Published: (2026)
by: Rahman, Robi
Published: (2026)
Fluent dreaming for language models
by: Thompson, T. Ben, et al.
Published: (2024)
by: Thompson, T. Ben, et al.
Published: (2024)
Humans overrely on overconfident language models, across languages
by: Rathi, Neil, et al.
Published: (2025)
by: Rathi, Neil, et al.
Published: (2025)
FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI
by: Glazer, Elliot, et al.
Published: (2024)
by: Glazer, Elliot, et al.
Published: (2024)
Natural language processing for African languages
by: Adelani, David Ifeoluwa
Published: (2025)
by: Adelani, David Ifeoluwa
Published: (2025)
(How) Do Language Models Track State?
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
How predictable is language model benchmark performance?
by: Owen, David
Published: (2024)
by: Owen, David
Published: (2024)
A survey on fairness of large language models in e-commerce: progress, application, and challenge
by: Ren, Qingyang, et al.
Published: (2024)
by: Ren, Qingyang, et al.
Published: (2024)
Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
by: Park, Jihyung, et al.
Published: (2025)
by: Park, Jihyung, et al.
Published: (2025)
Evidence of interrelated cognitive-like capabilities in large language models: Indications of artificial general intelligence or achievement?
by: Ilić, David, et al.
Published: (2023)
by: Ilić, David, et al.
Published: (2023)
Dissociating language and thought in large language models
by: Mahowald, Kyle, et al.
Published: (2023)
by: Mahowald, Kyle, et al.
Published: (2023)
Investigating the interaction of linguistic and mathematical reasoning in language models using multilingual number puzzles
by: Bhattacharya, Antara Raaghavi, et al.
Published: (2025)
by: Bhattacharya, Antara Raaghavi, et al.
Published: (2025)
CXR-LLAVA: a multimodal large language model for interpreting chest X-ray images
by: Lee, Seowoo, et al.
Published: (2023)
by: Lee, Seowoo, et al.
Published: (2023)
Do Chinese models speak Chinese languages?
by: Wen-Yi, Andrea W, et al.
Published: (2025)
by: Wen-Yi, Andrea W, et al.
Published: (2025)
Training Language Models to Explain Their Own Computations
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
Position: Avoid Overstretching LLMs for every Enterprise Task
by: Singh, Kuldeep, et al.
Published: (2026)
by: Singh, Kuldeep, et al.
Published: (2026)
Vibe-Eval: A hard evaluation suite for measuring progress of multimodal language models
by: Padlewski, Piotr, et al.
Published: (2024)
by: Padlewski, Piotr, et al.
Published: (2024)
Fresh in memory: Training-order recency is linearly encoded in language model activations
by: Krasheninnikov, Dmitrii, et al.
Published: (2025)
by: Krasheninnikov, Dmitrii, et al.
Published: (2025)
Block removal for large language models through constrained binary optimization
by: Jansen, David, et al.
Published: (2026)
by: Jansen, David, et al.
Published: (2026)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
by: Byers, Neil, et al.
Published: (2025)
by: Byers, Neil, et al.
Published: (2025)
On the attribution of confidence to large language models
by: Keeling, Geoff, et al.
Published: (2024)
by: Keeling, Geoff, et al.
Published: (2024)
Large language models and linguistic intentionality
by: Grindrod, Jumbly
Published: (2024)
by: Grindrod, Jumbly
Published: (2024)
A survey of textual cyber abuse detection using cutting-edge language models and large language models
by: Diaz-Garcia, Jose A., et al.
Published: (2025)
by: Diaz-Garcia, Jose A., et al.
Published: (2025)
Do language models practice what they preach? Examining language ideologies about gendered language reform encoded in LLMs
by: Watson, Julia, et al.
Published: (2024)
by: Watson, Julia, et al.
Published: (2024)
Rewarding Doubt: A Reinforcement Learning Approach to Calibrated Confidence Expression of Large Language Models
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement Learning
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Linguistic traces of stochastic empathy in language models
by: Kleinberg, Bennett, et al.
Published: (2024)
by: Kleinberg, Bennett, et al.
Published: (2024)
Infusing clinical knowledge into tokenisers for language models
by: Hasan, Abul, et al.
Published: (2024)
by: Hasan, Abul, et al.
Published: (2024)
CONFLARE: CONFormal LArge language model REtrieval
by: Rouzrokh, Pouria, et al.
Published: (2024)
by: Rouzrokh, Pouria, et al.
Published: (2024)
Experimental evidence of progressive ChatGPT models self-convergence
by: Xylogiannopoulos, Konstantinos F., et al.
Published: (2026)
by: Xylogiannopoulos, Konstantinos F., et al.
Published: (2026)
The use of large language models to enhance cancer clinical trial educational materials
by: Gao, Mingye, et al.
Published: (2024)
by: Gao, Mingye, et al.
Published: (2024)
Similar Items
-
Chinchilla Scaling: A replication attempt
by: Besiroglu, Tamay, et al.
Published: (2024) -
Estimating Idea Production: A Methodological Survey
by: Erdil, Ege, et al.
Published: (2024) -
Explosive growth from AI automation: A review of the arguments
by: Erdil, Ege, et al.
Published: (2023) -
Will we run out of data? Limits of LLM scaling based on human-generated data
by: Villalobos, Pablo, et al.
Published: (2022) -
GATE: An Integrated Assessment Model for AI Automation
by: Erdil, Ege, et al.
Published: (2025)