ReasonCACHE: Teaching LLMs To Reason Without Weight Updates
Fuente:
arXiv
Saved in:
| Main Authors: | Gupta, Sharut, Isola, Phillip, Jegelka, Stefanie, Lopez-Paz, David, Ahuja, Kartik, Ibrahim, Mark, Pezeshki, Mohammad |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Multi-Token Prediction: Pretraining LLMs with Future Summaries
by: Mahajan, Divyat, et al.
Published: (2025)
by: Mahajan, Divyat, et al.
Published: (2025)
G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning
by: Guo, Xiaojun, et al.
Published: (2025)
by: Guo, Xiaojun, et al.
Published: (2025)
Canonicalizing Multimodal Contrastive Representation Learning
by: Gupta, Sharut, et al.
Published: (2026)
by: Gupta, Sharut, et al.
Published: (2026)
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
by: Gupta, Sharut, et al.
Published: (2025)
by: Gupta, Sharut, et al.
Published: (2025)
Fairness Aware Reward Optimization
by: Choi, Ching Lam, et al.
Published: (2026)
by: Choi, Ching Lam, et al.
Published: (2026)
The Challenge of Teaching Reasoning to LLMs Without RL or Distillation
by: Du, Wei, et al.
Published: (2025)
by: Du, Wei, et al.
Published: (2025)
Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights
by: Gan, Yulu, et al.
Published: (2026)
by: Gan, Yulu, et al.
Published: (2026)
Understanding the Role of Equivariance in Self-supervised Learning
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
Compositional Risk Minimization
by: Mahajan, Divyat, et al.
Published: (2024)
by: Mahajan, Divyat, et al.
Published: (2024)
Discovering environments with XRM
by: Pezeshki, Mohammad, et al.
Published: (2023)
by: Pezeshki, Mohammad, et al.
Published: (2023)
An Information Criterion for Controlled Disentanglement of Multimodal Data
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
Learning Diffusion Models with Flexible Representation Guidance
by: Wang, Chenyu, et al.
Published: (2025)
by: Wang, Chenyu, et al.
Published: (2025)
When Less is Enough: Efficient Inference via Collaborative Reasoning
by: Chen, Yilei, et al.
Published: (2026)
by: Chen, Yilei, et al.
Published: (2026)
AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions
by: Kirichenko, Polina, et al.
Published: (2025)
by: Kirichenko, Polina, et al.
Published: (2025)
The Pitfalls of Memorization: When Memorization Hurts Generalization
by: Bayat, Reza, et al.
Published: (2024)
by: Bayat, Reza, et al.
Published: (2024)
Teaching and Evaluating LLMs to Reason About Polymer Design Related Tasks
by: Mohanty, Dikshya, et al.
Published: (2026)
by: Mohanty, Dikshya, et al.
Published: (2026)
Near-Optimal Algorithms for Group Distributionally Robust Optimization and Beyond
by: Soma, Tasuku, et al.
Published: (2022)
by: Soma, Tasuku, et al.
Published: (2022)
Teaching Transformers Causal Reasoning through Axiomatic Training
by: Vashishtha, Aniket, et al.
Published: (2024)
by: Vashishtha, Aniket, et al.
Published: (2024)
LANPO: Bootstrapping Language and Numerical Feedback for Reinforcement Learning in LLMs
by: Li, Ang, et al.
Published: (2025)
by: Li, Ang, et al.
Published: (2025)
Memorization vs. Reasoning: Updating LLMs with New Knowledge
by: Li, Aochong Oliver, et al.
Published: (2025)
by: Li, Aochong Oliver, et al.
Published: (2025)
Comprehension Without Competence: Architectural Limits of LLMs in Symbolic Computation and Reasoning
by: Zhang, Zheng
Published: (2025)
by: Zhang, Zheng
Published: (2025)
Distilling Reasoning Without Knowledge: A Framework for Reliable LLMs
by: Kietkajornrit, Auksarapak, et al.
Published: (2026)
by: Kietkajornrit, Auksarapak, et al.
Published: (2026)
Fine-Tuning Qwen 2.5 3B for Realistic Movie Dialogue Generation
by: Gupta, Kartik
Published: (2025)
by: Gupta, Kartik
Published: (2025)
Iterative Auto-Annotation for Scientific Named Entity Recognition Using BERT-Based Models
by: Gupta, Kartik
Published: (2025)
by: Gupta, Kartik
Published: (2025)
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs
by: Gjerde, Magnus F., et al.
Published: (2025)
by: Gjerde, Magnus F., et al.
Published: (2025)
Teaching LLMs to Ask: Self-Querying Category-Theoretic Planning for Under-Specified Reasoning
by: Qu, Shuhui
Published: (2026)
by: Qu, Shuhui
Published: (2026)
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
by: Wang, Tianle, et al.
Published: (2026)
by: Wang, Tianle, et al.
Published: (2026)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
by: Ning, Xuefei, et al.
Published: (2024)
by: Ning, Xuefei, et al.
Published: (2024)
Teaching LLMs According to Their Aptitude: Adaptive Reasoning for Mathematical Problem Solving
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Digital Red Queen: Adversarial Program Evolution in Core War with LLMs
by: Kumar, Akarsh, et al.
Published: (2026)
by: Kumar, Akarsh, et al.
Published: (2026)
Counting Without Running: Evaluating LLMs' Reasoning About Code Complexity
by: Bolet, Gregory, et al.
Published: (2025)
by: Bolet, Gregory, et al.
Published: (2025)
When More is Less: Understanding Chain-of-Thought Length in LLMs
by: Wu, Yuyang, et al.
Published: (2025)
by: Wu, Yuyang, et al.
Published: (2025)
Collective Reasoning Among LLMs: A Framework for Answer Validation Without Ground Truth
by: Davoudi, Seyed Pouyan Mousavi, et al.
Published: (2025)
by: Davoudi, Seyed Pouyan Mousavi, et al.
Published: (2025)
Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification
by: Donhauser, Konstantin, et al.
Published: (2025)
by: Donhauser, Konstantin, et al.
Published: (2025)
Towards efficient representation identification in supervised learning
by: Ahuja, Kartik, et al.
Published: (2022)
by: Ahuja, Kartik, et al.
Published: (2022)
Off-Trajectory Reasoning: Can LLMs Collaborate on Reasoning Trajectory?
by: Li, Aochong Oliver, et al.
Published: (2025)
by: Li, Aochong Oliver, et al.
Published: (2025)
Better Call CLAUSE: A Discrepancy Benchmark for Auditing LLMs Legal Reasoning Capabilities
by: Choudhury, Manan Roy, et al.
Published: (2025)
by: Choudhury, Manan Roy, et al.
Published: (2025)
Learning with Exact Invariances in Polynomial Time
by: Soleymani, Ashkan, et al.
Published: (2025)
by: Soleymani, Ashkan, et al.
Published: (2025)
Survey on Generalization Theory for Graph Neural Networks
by: Vasileiou, Antonis, et al.
Published: (2025)
by: Vasileiou, Antonis, et al.
Published: (2025)
CORE: Collaborative Reasoning via Cross Teaching
by: Mishra, Kshitij, et al.
Published: (2026)
by: Mishra, Kshitij, et al.
Published: (2026)
Similar Items
-
Beyond Multi-Token Prediction: Pretraining LLMs with Future Summaries
by: Mahajan, Divyat, et al.
Published: (2025) -
G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning
by: Guo, Xiaojun, et al.
Published: (2025) -
Canonicalizing Multimodal Contrastive Representation Learning
by: Gupta, Sharut, et al.
Published: (2026) -
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
by: Gupta, Sharut, et al.
Published: (2025) -
Fairness Aware Reward Optimization
by: Choi, Ching Lam, et al.
Published: (2026)