Escaping the Cognitive Well: Efficient Competition Math with Off-the-Shelf Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Dang, Xingyu, Agarwal, Rohit, Porto, Rodrigo, Goyal, Anirudh, Fowl, Liam H, Arora, Sanjeev |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Impossibility of Retrain Equivalence in Machine Unlearning
por: Yu, Jiatong, et al.
Publicado: (2025)
por: Yu, Jiatong, et al.
Publicado: (2025)
Instruct-SkillMix: A Powerful Pipeline for LLM Instruction Tuning
por: Kaur, Simran, et al.
Publicado: (2024)
por: Kaur, Simran, et al.
Publicado: (2024)
Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
por: Didolkar, Aniket, et al.
Publicado: (2025)
por: Didolkar, Aniket, et al.
Publicado: (2025)
Can Models Learn Skill Composition from Examples?
por: Zhao, Haoyu, et al.
Publicado: (2024)
por: Zhao, Haoyu, et al.
Publicado: (2024)
AI-Assisted Generation of Difficult Math Questions
por: Shah, Vedant, et al.
Publicado: (2024)
por: Shah, Vedant, et al.
Publicado: (2024)
On the Power of Context-Enhanced Learning in LLMs
por: Zhu, Xingyu, et al.
Publicado: (2025)
por: Zhu, Xingyu, et al.
Publicado: (2025)
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates
por: Lyu, Kaifeng, et al.
Publicado: (2024)
por: Lyu, Kaifeng, et al.
Publicado: (2024)
Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?
por: Park, Simon, et al.
Publicado: (2025)
por: Park, Simon, et al.
Publicado: (2025)
Unlearning via Sparse Representations
por: Shah, Vedant, et al.
Publicado: (2023)
por: Shah, Vedant, et al.
Publicado: (2023)
Contextual Drag: How Errors in the Context Affect LLM Reasoning
por: Cheng, Yun, et al.
Publicado: (2026)
por: Cheng, Yun, et al.
Publicado: (2026)
Rethinking Thinking Tokens: LLMs as Improvement Operators
por: Madaan, Lovish, et al.
Publicado: (2025)
por: Madaan, Lovish, et al.
Publicado: (2025)
How Does RL Post-training Induce Skill Composition? A Case Study on Countdown
por: Park, Simon, et al.
Publicado: (2025)
por: Park, Simon, et al.
Publicado: (2025)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
por: Didolkar, Aniket, et al.
Publicado: (2024)
por: Didolkar, Aniket, et al.
Publicado: (2024)
Approximate Attributions for Off-the-Shelf Siamese Transformers
por: Möller, Lucas, et al.
Publicado: (2024)
por: Möller, Lucas, et al.
Publicado: (2024)
Detecting Drunk Driving Using Off-the-Shelf Smartwatches
por: Deuber, Robin, et al.
Publicado: (2026)
por: Deuber, Robin, et al.
Publicado: (2026)
Why is Your Language Model a Poor Implicit Reward Model?
por: Razin, Noam, et al.
Publicado: (2025)
por: Razin, Noam, et al.
Publicado: (2025)
On the SDEs and Scaling Rules for Adaptive Gradient Algorithms
por: Malladi, Sadhika, et al.
Publicado: (2022)
por: Malladi, Sadhika, et al.
Publicado: (2022)
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions
por: Artiles, Alejandro H., et al.
Publicado: (2026)
por: Artiles, Alejandro H., et al.
Publicado: (2026)
QR-LoRA: QR-Based Low-Rank Adaptation for Efficient Fine-Tuning of Large Language Models
por: Liang, Jessica, et al.
Publicado: (2025)
por: Liang, Jessica, et al.
Publicado: (2025)
Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models
por: Noukhovitch, Michael, et al.
Publicado: (2024)
por: Noukhovitch, Michael, et al.
Publicado: (2024)
Improving Sparse Memory Finetuning
por: Goyal, Satyam, et al.
Publicado: (2026)
por: Goyal, Satyam, et al.
Publicado: (2026)
Sparse Memory Finetuning as a Low-Forgetting Alternative to LoRA and Full Finetuning
por: Gupta, Prakhar, et al.
Publicado: (2026)
por: Gupta, Prakhar, et al.
Publicado: (2026)
$α$-TCVAE: On the relationship between Disentanglement and Diversity
por: Meo, Cristian, et al.
Publicado: (2024)
por: Meo, Cristian, et al.
Publicado: (2024)
Narrowing the Focus: Learned Optimizers for Pretrained Models
por: Kristiansen, Gus, et al.
Publicado: (2024)
por: Kristiansen, Gus, et al.
Publicado: (2024)
Training Language Models to Reason Efficiently
por: Arora, Daman, et al.
Publicado: (2025)
por: Arora, Daman, et al.
Publicado: (2025)
Skill-Targeted Adaptive Training
por: He, Yinghui, et al.
Publicado: (2025)
por: He, Yinghui, et al.
Publicado: (2025)
Trainable Transformer in Transformer
por: Panigrahi, Abhishek, et al.
Publicado: (2023)
por: Panigrahi, Abhishek, et al.
Publicado: (2023)
Provable unlearning in topic modeling and downstream tasks
por: Wei, Stanley, et al.
Publicado: (2024)
por: Wei, Stanley, et al.
Publicado: (2024)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
por: Prabhudesai, Mihir, et al.
Publicado: (2023)
por: Prabhudesai, Mihir, et al.
Publicado: (2023)
A Quadratic Synchronization Rule for Distributed Deep Learning
por: Gu, Xinran, et al.
Publicado: (2023)
por: Gu, Xinran, et al.
Publicado: (2023)
Unrealized Expectations: Comparing AI Methods vs Classical Algorithms for Maximum Independent Set
por: Wu, Yikai, et al.
Publicado: (2025)
por: Wu, Yikai, et al.
Publicado: (2025)
Data-driven Multistage Distributionally Robust Linear Optimization with Nested Distance
por: Gao, Rui, et al.
Publicado: (2024)
por: Gao, Rui, et al.
Publicado: (2024)
MathAtlas: A Benchmark for Autoformalization in the Wild
por: Patel, Nilay, et al.
Publicado: (2026)
por: Patel, Nilay, et al.
Publicado: (2026)
ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation
por: Patni, Suraj, et al.
Publicado: (2024)
por: Patni, Suraj, et al.
Publicado: (2024)
Labels in Extremes: How Well Calibrated are Extreme Multi-label Classifiers?
por: Ullah, Nasib, et al.
Publicado: (2024)
por: Ullah, Nasib, et al.
Publicado: (2024)
ESSAM: A Novel Competitive Evolution Strategies Approach to Reinforcement Learning for Memory Efficient LLMs Fine-Tuning
por: Sun, Zhishen, et al.
Publicado: (2026)
por: Sun, Zhishen, et al.
Publicado: (2026)
Efficient Matrix Factorization Via Householder Reflections
por: Dash, Anirudh, et al.
Publicado: (2024)
por: Dash, Anirudh, et al.
Publicado: (2024)
On the Escaping Efficiency of Distributed Adversarial Training Algorithms
por: Cao, Ying, et al.
Publicado: (2025)
por: Cao, Ying, et al.
Publicado: (2025)
A Machine Learning Framework for Off Ball Defensive Role and Performance Evaluation in Football
por: Groom, Sean, et al.
Publicado: (2026)
por: Groom, Sean, et al.
Publicado: (2026)
Ejemplares similares
-
On the Impossibility of Retrain Equivalence in Machine Unlearning
por: Yu, Jiatong, et al.
Publicado: (2025) -
Instruct-SkillMix: A Powerful Pipeline for LLM Instruction Tuning
por: Kaur, Simran, et al.
Publicado: (2024) -
Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
por: Didolkar, Aniket, et al.
Publicado: (2025) -
Can Models Learn Skill Composition from Examples?
por: Zhao, Haoyu, et al.
Publicado: (2024) -
AI-Assisted Generation of Difficult Math Questions
por: Shah, Vedant, et al.
Publicado: (2024)