Self-Evolving Curriculum for LLM Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xiaoyin, Lu, Jiarui, Kim, Minsu, Zhang, Dinghuai, Tang, Jian, Piché, Alexandre, Gontier, Nicolas, Bengio, Yoshua, Kamalloo, Ehsan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
PipelineRL: Faster On-policy Reinforcement Learning for Long Sequence Generation
by: Piché, Alexandre, et al.
Published: (2025)
by: Piché, Alexandre, et al.
Published: (2025)
Distributional GFlowNets with Quantile Flows
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
Efficient Causal Graph Discovery Using Large Language Models
by: Jiralerspong, Thomas, et al.
Published: (2024)
by: Jiralerspong, Thomas, et al.
Published: (2024)
Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
Active Attacks: Red-teaming LLMs via Adaptive Environments
by: Yun, Taeyoung, et al.
Published: (2025)
by: Yun, Taeyoung, et al.
Published: (2025)
Generative Recursive Reasoning
by: Baek, Junyeob, et al.
Published: (2026)
by: Baek, Junyeob, et al.
Published: (2026)
Efficient Regression-Based Training of Normalizing Flows for Boltzmann Generators
by: Rehman, Danyal, et al.
Published: (2025)
by: Rehman, Danyal, et al.
Published: (2025)
Expert-Guided LLM Reasoning for Battery Discovery: From AI-Driven Hypothesis to Synthesis and Characterization
by: Liu, Shengchao, et al.
Published: (2025)
by: Liu, Shengchao, et al.
Published: (2025)
Preventing Curriculum Collapse in Self-Evolving Reasoning Systems
by: Mishra, Vaibhav
Published: (2026)
by: Mishra, Vaibhav
Published: (2026)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
Local Search GFlowNets
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Baking Symmetry into GFlowNets
by: Ma, George, et al.
Published: (2024)
by: Ma, George, et al.
Published: (2024)
LongRecall: A Structured Approach for Robust Recall Evaluation in Long-Form Text
by: Ardestani, MohamamdJavad, et al.
Published: (2025)
by: Ardestani, MohamamdJavad, et al.
Published: (2025)
In-Context Parametric Inference: Point or Distribution Estimators?
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
GFlowNet Foundations
by: Bengio, Yoshua, et al.
Published: (2021)
by: Bengio, Yoshua, et al.
Published: (2021)
Structure Language Models for Protein Conformation Generation
by: Lu, Jiarui, et al.
Published: (2024)
by: Lu, Jiarui, et al.
Published: (2024)
Machine learning and information theory concepts towards an AI Mathematician
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
TapeAgents: a Holistic Framework for Agent Development and Optimization
by: Bahdanau, Dzmitry, et al.
Published: (2024)
by: Bahdanau, Dzmitry, et al.
Published: (2024)
TTCS: Test-Time Curriculum Synthesis for Self-Evolving
by: Yang, Chengyi, et al.
Published: (2026)
by: Yang, Chengyi, et al.
Published: (2026)
How to Train Your LLM Web Agent: A Statistical Diagnosis
by: Vattikonda, Dheeraj, et al.
Published: (2025)
by: Vattikonda, Dheeraj, et al.
Published: (2025)
A Complexity-Based Theory of Compositionality
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Monte Carlo Tree Diffusion for System 2 Planning
by: Yoon, Jaesik, et al.
Published: (2025)
by: Yoon, Jaesik, et al.
Published: (2025)
Shaping Inductive Bias in Diffusion Models through Frequency-Based Noise Control
by: Jiralerspong, Thomas, et al.
Published: (2025)
by: Jiralerspong, Thomas, et al.
Published: (2025)
In-Context Reinforcement Learning through Bayesian Fusion of Context and Value Prior
by: Berkes, Anaïs, et al.
Published: (2026)
by: Berkes, Anaïs, et al.
Published: (2026)
Learning to Scale Logits for Temperature-Conditional GFlowNets
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Were RNNs All We Needed?
by: Feng, Leo, et al.
Published: (2024)
by: Feng, Leo, et al.
Published: (2024)
FALCON: Few-step Accurate Likelihoods for Continuous Flows
by: Rehman, Danyal, et al.
Published: (2025)
by: Rehman, Danyal, et al.
Published: (2025)
Solving Bayesian inverse problems with diffusion priors and off-policy RL
by: Scimeca, Luca, et al.
Published: (2025)
by: Scimeca, Luca, et al.
Published: (2025)
Privileged Information Distillation for Language Models
by: Penaloza, Emiliano, et al.
Published: (2026)
by: Penaloza, Emiliano, et al.
Published: (2026)
SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning
by: Chen, Jiaqi, et al.
Published: (2025)
by: Chen, Jiaqi, et al.
Published: (2025)
R-Zero: Self-Evolving Reasoning LLM from Zero Data
by: Huang, Chengsong, et al.
Published: (2025)
by: Huang, Chengsong, et al.
Published: (2025)
Discrete, compositional, and symbolic representations through attractor dynamics
by: Nam, Andrew, et al.
Published: (2023)
by: Nam, Andrew, et al.
Published: (2023)
Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning
by: Zhao, Mingde, et al.
Published: (2023)
by: Zhao, Mingde, et al.
Published: (2023)
Action abstractions for amortized sampling
by: Boussif, Oussama, et al.
Published: (2024)
by: Boussif, Oussama, et al.
Published: (2024)
Structure-Aligned Protein Language Model
by: Chen, Can, et al.
Published: (2025)
by: Chen, Can, et al.
Published: (2025)
Reinforcing Chain-of-Thought Reasoning with Self-Evolving Rubrics
by: Sheng, Leheng, et al.
Published: (2026)
by: Sheng, Leheng, et al.
Published: (2026)
Apriel-1.5-OpenReasoner: RL Post-Training for General-Purpose and Efficient Reasoning
by: Pardinas, Rafael, et al.
Published: (2026)
by: Pardinas, Rafael, et al.
Published: (2026)
MathSE: Improving Multimodal Mathematical Reasoning via Self-Evolving Iterative Reflection and Reward-Guided Fine-Tuning
by: Chen, Jinhao, et al.
Published: (2025)
by: Chen, Jinhao, et al.
Published: (2025)
Learning to Self-Evolve
by: Chen, Xiaoyin, et al.
Published: (2026)
by: Chen, Xiaoyin, et al.
Published: (2026)
Similar Items
-
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025) -
PipelineRL: Faster On-policy Reinforcement Learning for Long Sequence Generation
by: Piché, Alexandre, et al.
Published: (2025) -
Distributional GFlowNets with Quantile Flows
by: Zhang, Dinghuai, et al.
Published: (2023) -
Efficient Causal Graph Discovery Using Large Language Models
by: Jiralerspong, Thomas, et al.
Published: (2024) -
Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization
by: Zhang, Dinghuai, et al.
Published: (2023)