Deep Self-Evolving Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zihan, Zheng, Shun, Wen, Xumeng, Wang, Yang, Bian, Jiang, Yang, Mao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
von: Wen, Xumeng, et al.
Veröffentlicht: (2025)
von: Wen, Xumeng, et al.
Veröffentlicht: (2025)
Scalable In-Context Learning on Tabular Data via Retrieval-Augmented Large Language Models
von: Wen, Xumeng, et al.
Veröffentlicht: (2025)
von: Wen, Xumeng, et al.
Veröffentlicht: (2025)
Large Language Model as a Universal Clinical Multi-task Decoder
von: Wu, Yujiang, et al.
Veröffentlicht: (2024)
von: Wu, Yujiang, et al.
Veröffentlicht: (2024)
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
von: Guan, Xinyu, et al.
Veröffentlicht: (2025)
von: Guan, Xinyu, et al.
Veröffentlicht: (2025)
EvolvR: Self-Evolving Pairwise Reasoning for Story Evaluation to Enhance Generation
von: Wang, Xinda, et al.
Veröffentlicht: (2025)
von: Wang, Xinda, et al.
Veröffentlicht: (2025)
EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle
von: Wu, Rong, et al.
Veröffentlicht: (2025)
von: Wu, Rong, et al.
Veröffentlicht: (2025)
ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory
von: Ouyang, Siru, et al.
Veröffentlicht: (2025)
von: Ouyang, Siru, et al.
Veröffentlicht: (2025)
Emulating Clinician Cognition via Self-Evolving Deep Clinical Research
von: Ren, Ruiyang, et al.
Veröffentlicht: (2026)
von: Ren, Ruiyang, et al.
Veröffentlicht: (2026)
On the Generalization Gap in Self-Evolving Language Model Reasoning
von: Qi, Zhenting, et al.
Veröffentlicht: (2026)
von: Qi, Zhenting, et al.
Veröffentlicht: (2026)
NOTAM-Evolve: A Knowledge-Guided Self-Evolving Optimization Framework with LLMs for NOTAM Interpretation
von: Liu, Maoqi, et al.
Veröffentlicht: (2025)
von: Liu, Maoqi, et al.
Veröffentlicht: (2025)
MetaGen: Self-Evolving Roles and Topologies for Multi-Agent LLM Reasoning
von: Wang, Yimeng, et al.
Veröffentlicht: (2026)
von: Wang, Yimeng, et al.
Veröffentlicht: (2026)
Self-Evolving LLM Memory Extraction Across Heterogeneous Tasks
von: Yang, Yuqing, et al.
Veröffentlicht: (2026)
von: Yang, Yuqing, et al.
Veröffentlicht: (2026)
EvolveSearch: An Iterative Self-Evolving Search Agent
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025)
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025)
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition
von: Chen, Yanyu, et al.
Veröffentlicht: (2026)
von: Chen, Yanyu, et al.
Veröffentlicht: (2026)
Diving into Self-Evolving Training for Multimodal Reasoning
von: Liu, Wei, et al.
Veröffentlicht: (2024)
von: Liu, Wei, et al.
Veröffentlicht: (2024)
Rethinking Experience Utilization in Self-Evolving Language Model Agents
von: Zhao, Weixiang, et al.
Veröffentlicht: (2026)
von: Zhao, Weixiang, et al.
Veröffentlicht: (2026)
On Safety Risks in Experience-Driven Self-Evolving Agents
von: Zhao, Weixiang, et al.
Veröffentlicht: (2026)
von: Zhao, Weixiang, et al.
Veröffentlicht: (2026)
Learning to Self-Evolve
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2026)
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2026)
MiMoTable: A Multi-scale Spreadsheet Benchmark with Meta Operations for Table Reasoning
von: Li, Zheng, et al.
Veröffentlicht: (2024)
von: Li, Zheng, et al.
Veröffentlicht: (2024)
From Supervised to Generative: A Novel Paradigm for Tabular Deep Learning with Large Language Models
von: Wen, Xumeng, et al.
Veröffentlicht: (2023)
von: Wen, Xumeng, et al.
Veröffentlicht: (2023)
iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models
von: Sunil, Meghana, et al.
Veröffentlicht: (2026)
von: Sunil, Meghana, et al.
Veröffentlicht: (2026)
AgenticGEO: A Self-Evolving Agentic System for Generative Engine Optimization
von: Yuan, Jiaqi, et al.
Veröffentlicht: (2026)
von: Yuan, Jiaqi, et al.
Veröffentlicht: (2026)
CreativeBench: Benchmarking and Enhancing Machine Creativity via Self-Evolving Challenges
von: Wang, Zi-Han, et al.
Veröffentlicht: (2026)
von: Wang, Zi-Han, et al.
Veröffentlicht: (2026)
SkillOpt: Executive Strategy for Self-Evolving Agent Skills
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
Language Models as Continuous Self-Evolving Data Engineers
von: Wang, Peidong, et al.
Veröffentlicht: (2024)
von: Wang, Peidong, et al.
Veröffentlicht: (2024)
Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki
von: Ming, Haoliang, et al.
Veröffentlicht: (2026)
von: Ming, Haoliang, et al.
Veröffentlicht: (2026)
MemRL: Self-Evolving Agents via Runtime Reinforcement Learning on Episodic Memory
von: Zhang, Shengtao, et al.
Veröffentlicht: (2026)
von: Zhang, Shengtao, et al.
Veröffentlicht: (2026)
Self-Evolved Reward Learning for LLMs
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking
von: Li, Jindong, et al.
Veröffentlicht: (2026)
von: Li, Jindong, et al.
Veröffentlicht: (2026)
LatentEvolve: Self-Evolving Test-Time Scaling in Latent Space
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
FABSVer: Faster Training and Better Self-Verification for LLM Mathematical Reasoning
von: Pan, Haihui, et al.
Veröffentlicht: (2026)
von: Pan, Haihui, et al.
Veröffentlicht: (2026)
$G^2$-Reader: Dual Evolving Graphs for Multimodal Document QA
von: Du, Yaxin, et al.
Veröffentlicht: (2026)
von: Du, Yaxin, et al.
Veröffentlicht: (2026)
Search-E1: Self-Distillation Drives Self-Evolution in Search-Augmented Reasoning
von: Liang, Zihan, et al.
Veröffentlicht: (2026)
von: Liang, Zihan, et al.
Veröffentlicht: (2026)
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization
von: Xin, Haidong, et al.
Veröffentlicht: (2026)
von: Xin, Haidong, et al.
Veröffentlicht: (2026)
Self-Contradictory Reasoning Evaluation and Detection
von: Liu, Ziyi, et al.
Veröffentlicht: (2023)
von: Liu, Ziyi, et al.
Veröffentlicht: (2023)
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
EvoWiki: Evaluating LLMs on Evolving Knowledge
von: Tang, Wei, et al.
Veröffentlicht: (2024)
von: Tang, Wei, et al.
Veröffentlicht: (2024)
R-Zero: Self-Evolving Reasoning LLM from Zero Data
von: Huang, Chengsong, et al.
Veröffentlicht: (2025)
von: Huang, Chengsong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
von: Wen, Xumeng, et al.
Veröffentlicht: (2025) -
Scalable In-Context Learning on Tabular Data via Retrieval-Augmented Large Language Models
von: Wen, Xumeng, et al.
Veröffentlicht: (2025) -
Large Language Model as a Universal Clinical Multi-task Decoder
von: Wu, Yujiang, et al.
Veröffentlicht: (2024) -
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
von: Guan, Xinyu, et al.
Veröffentlicht: (2025) -
EvolvR: Self-Evolving Pairwise Reasoning for Story Evaluation to Enhance Generation
von: Wang, Xinda, et al.
Veröffentlicht: (2025)