Procedural Knowledge in Pretraining Drives Reasoning in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ruis, Laura, Mozes, Maximilian, Bae, Juhan, Kamalakara, Siddhartha Rao, Talupuru, Dwarak, Locatelli, Acyr, Kirk, Robert, Rocktäschel, Tim, Grefenstette, Edward, Bartolo, Max |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rope to Nope and Back Again: A New Hybrid Attention Strategy
by: Yang, Bowen, et al.
Published: (2025)
by: Yang, Bowen, et al.
Published: (2025)
Investigating Non-Transitivity in LLM-as-a-Judge
by: Xu, Yi, et al.
Published: (2025)
by: Xu, Yi, et al.
Published: (2025)
Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions
by: Rosser, J, et al.
Published: (2026)
by: Rosser, J, et al.
Published: (2026)
minimax: Efficient Baselines for Autocurricula in JAX
by: Jiang, Minqi, et al.
Published: (2023)
by: Jiang, Minqi, et al.
Published: (2023)
Understanding Likelihood Over-optimisation in Direct Alignment Algorithms
by: Shi, Zhengyan, et al.
Published: (2024)
by: Shi, Zhengyan, et al.
Published: (2024)
Debating with More Persuasive LLMs Leads to More Truthful Answers
by: Khan, Akbir, et al.
Published: (2024)
by: Khan, Akbir, et al.
Published: (2024)
Efficient Benchmarking Is Just Feature Selection and Multiple Regression
by: Bowyer, Sam, et al.
Published: (2026)
by: Bowyer, Sam, et al.
Published: (2026)
Mechanistically analyzing the effects of fine-tuning on procedurally defined tasks
by: Jain, Samyak, et al.
Published: (2023)
by: Jain, Samyak, et al.
Published: (2023)
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
by: Cook, Jonathan, et al.
Published: (2025)
by: Cook, Jonathan, et al.
Published: (2025)
Scaling Opponent Shaping to High Dimensional Games
by: Khan, Akbir, et al.
Published: (2023)
by: Khan, Akbir, et al.
Published: (2023)
Aya 23: Open Weight Releases to Further Multilingual Progress
by: Aryabumi, Viraat, et al.
Published: (2024)
by: Aryabumi, Viraat, et al.
Published: (2024)
Reverse Engineering Human Preferences with Reinforcement Learning
by: Alazraki, Lisa, et al.
Published: (2025)
by: Alazraki, Lisa, et al.
Published: (2025)
No Need for Explanations: LLMs can implicitly learn from mistakes in-context
by: Alazraki, Lisa, et al.
Published: (2025)
by: Alazraki, Lisa, et al.
Published: (2025)
Nexus: Specialization meets Adaptability for Efficiently Training Mixture of Experts
by: Gritsch, Nikolas, et al.
Published: (2024)
by: Gritsch, Nikolas, et al.
Published: (2024)
Interaction Dynamics as a Reward Signal for LLMs
by: Gooding, Sian, et al.
Published: (2025)
by: Gooding, Sian, et al.
Published: (2025)
O que os analistas pensam sobre a homossexualidade?
by: Acyr Maya
Published: (2007)
by: Acyr Maya
Published: (2007)
Learning When to Plan: Efficiently Allocating Test-Time Compute for LLM Agents
by: Paglieri, Davide, et al.
Published: (2025)
by: Paglieri, Davide, et al.
Published: (2025)
LLM-First Search: Self-Guided Exploration of the Solution Space
by: Herr, Nathan, et al.
Published: (2025)
by: Herr, Nathan, et al.
Published: (2025)
Síndrome de Frey simulando eritema malar por alergia alimentar
by: Fabiana Mozes
Published: (2007)
by: Fabiana Mozes
Published: (2007)
Mis andares con los tiples
by: Jorge Velosa Ruis
Published: (2013)
by: Jorge Velosa Ruis
Published: (2013)
The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning
by: Xu, Yi, et al.
Published: (2026)
by: Xu, Yi, et al.
Published: (2026)
Preference-Based Alignment of Discrete Diffusion Models
by: Borso, Umberto, et al.
Published: (2025)
by: Borso, Umberto, et al.
Published: (2025)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
by: Pignatelli, Eduardo, et al.
Published: (2024)
by: Pignatelli, Eduardo, et al.
Published: (2024)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
by: Kirk, Robert, et al.
Published: (2023)
by: Kirk, Robert, et al.
Published: (2023)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
by: Yuan, Kun, et al.
Published: (2024)
by: Yuan, Kun, et al.
Published: (2024)
Transformers Pretrained on Procedural Data Contain Modular Structures for Algorithmic Reasoning
by: Shinnick, Zachary, et al.
Published: (2025)
by: Shinnick, Zachary, et al.
Published: (2025)
Writing as a testbed for open ended agents
by: Gooding, Sian, et al.
Published: (2025)
by: Gooding, Sian, et al.
Published: (2025)
Procedural Knowledge at Scale Improves Reasoning
by: Wu, Di, et al.
Published: (2026)
by: Wu, Di, et al.
Published: (2026)
Fishing for Magikarp: Automatically Detecting Under-trained Tokens in Large Language Models
by: Land, Sander, et al.
Published: (2024)
by: Land, Sander, et al.
Published: (2024)
El análisis cultural en los estudios de impacto ambiental. Dos estudios de caso: Proyecto Eólico Piloto Jepirachi y Proyecto de Conexión Vial entre los Valles de Aburrá y del Río Cauca
by: Aura Luz Ruis A.
Published: (2006)
by: Aura Luz Ruis A.
Published: (2006)
Outliers and Calibration Sets have Diminishing Effect on Quantization of Modern LLMs
by: Paglieri, Davide, et al.
Published: (2024)
by: Paglieri, Davide, et al.
Published: (2024)
Distances in Planar Graphs are Almost for Free!
by: Mozes, Shay, et al.
Published: (2026)
by: Mozes, Shay, et al.
Published: (2026)
A Conceptual Framework for Teaching Legal Research to Undergraduates.
by: Bartolo, Laura M.
Published: (1991)
by: Bartolo, Laura M.
Published: (1991)
Exploring Training Data Attribution under Limited Access Constraints
by: Zhang, Shiyuan, et al.
Published: (2025)
by: Zhang, Shiyuan, et al.
Published: (2025)
Training Data Attribution via Approximate Unrolled Differentiation
by: Bae, Juhan, et al.
Published: (2024)
by: Bae, Juhan, et al.
Published: (2024)
IF-GUIDE: Influence Function-Guided Detoxification of LLMs
by: Coalson, Zachary, et al.
Published: (2025)
by: Coalson, Zachary, et al.
Published: (2025)
Training Data Attribution (TDA): Examining Its Adoption & Use Cases
by: Cheng, Deric, et al.
Published: (2025)
by: Cheng, Deric, et al.
Published: (2025)
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
by: Cook, Jonathan, et al.
Published: (2024)
by: Cook, Jonathan, et al.
Published: (2024)
Desarrollo de un modelo de cultivo in vitro para Vallisneria americana Michx
by: Ruis Carrera, V - Sáchez, AJ
Published: (2008)
by: Ruis Carrera, V - Sáchez, AJ
Published: (2008)
Do Papers with Titles Ending in a Question Mark Usually Have the Answer "No"?
by: Stern, Daniel, et al.
Published: (2026)
by: Stern, Daniel, et al.
Published: (2026)
Similar Items
-
Rope to Nope and Back Again: A New Hybrid Attention Strategy
by: Yang, Bowen, et al.
Published: (2025) -
Investigating Non-Transitivity in LLM-as-a-Judge
by: Xu, Yi, et al.
Published: (2025) -
Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions
by: Rosser, J, et al.
Published: (2026) -
minimax: Efficient Baselines for Autocurricula in JAX
by: Jiang, Minqi, et al.
Published: (2023) -
Understanding Likelihood Over-optimisation in Direct Alignment Algorithms
by: Shi, Zhengyan, et al.
Published: (2024)