Saved in:
| Main Authors: | Gungordu, Oguzhan, Xiong, Siheng, Fekri, Faramarz |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.20539 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scaling Search-Augmented LLM Reasoning via Adaptive Information Control
by: Xiong, Siheng, et al.
Published: (2026)
by: Xiong, Siheng, et al.
Published: (2026)
Enhancing Long Chain-of-Thought Reasoning through Multi-Path Plan Aggregation
by: Xiong, Siheng, et al.
Published: (2025)
by: Xiong, Siheng, et al.
Published: (2025)
Deliberate Reasoning in Language Models as Structure-Aware Planning with an Accurate World Model
by: Xiong, Siheng, et al.
Published: (2024)
by: Xiong, Siheng, et al.
Published: (2024)
The Compressor-Retriever Architecture for Language Model OS
by: Yang, Yuan, et al.
Published: (2024)
by: Yang, Yuan, et al.
Published: (2024)
Large Language Models Can Learn Temporal Reasoning
by: Xiong, Siheng, et al.
Published: (2024)
by: Xiong, Siheng, et al.
Published: (2024)
Can LLMs Reason in the Wild with Programs?
by: Yang, Yuan, et al.
Published: (2024)
by: Yang, Yuan, et al.
Published: (2024)
Generalization of Compositional Tasks with Logical Specification via Implicit Planning
by: Xu, Duo, et al.
Published: (2024)
by: Xu, Duo, et al.
Published: (2024)
Temporal Inductive Logic Reasoning over Hypergraphs
by: Yang, Yuan, et al.
Published: (2022)
by: Yang, Yuan, et al.
Published: (2022)
TILP: Differentiable Learning of Temporal Logical Rules on Knowledge Graphs
by: Xiong, Siheng, et al.
Published: (2024)
by: Xiong, Siheng, et al.
Published: (2024)
TEILP: Time Prediction over Knowledge Graphs via Logical Reasoning
by: Xiong, Siheng, et al.
Published: (2023)
by: Xiong, Siheng, et al.
Published: (2023)
MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation
by: He, Junlin, et al.
Published: (2026)
by: He, Junlin, et al.
Published: (2026)
LLM-A*: Large Language Model Enhanced Incremental Heuristic Search on Path Planning
by: Meng, Silin, et al.
Published: (2024)
by: Meng, Silin, et al.
Published: (2024)
Self-Alignment of Large Language Models via Monopolylogue-based Social Scene Simulation
by: Pang, Xianghe, et al.
Published: (2024)
by: Pang, Xianghe, et al.
Published: (2024)
Self-Evolved Reward Learning for LLMs
by: Huang, Chenghua, et al.
Published: (2024)
by: Huang, Chenghua, et al.
Published: (2024)
Conversation for Non-verifiable Learning: Self-Evolving LLMs through Meta-Evaluation
by: Sui, Yuan, et al.
Published: (2026)
by: Sui, Yuan, et al.
Published: (2026)
GLIDR: Graph-Like Inductive Logic Programming with Differentiable Reasoning
by: Johnson, Blair, et al.
Published: (2025)
by: Johnson, Blair, et al.
Published: (2025)
Learning Cyclic Causal Models from Incomplete Data
by: Sethuraman, Muralikrishnna G., et al.
Published: (2024)
by: Sethuraman, Muralikrishnna G., et al.
Published: (2024)
Robust Heuristic Algorithm Design with LLMs
by: Karimi, Pantea, et al.
Published: (2025)
by: Karimi, Pantea, et al.
Published: (2025)
EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle
by: Wu, Rong, et al.
Published: (2025)
by: Wu, Rong, et al.
Published: (2025)
Learning Hidden Subgoals under Temporal Ordering Constraints in Reinforcement Learning
by: Xu, Duo, et al.
Published: (2024)
by: Xu, Duo, et al.
Published: (2024)
NOTAM-Evolve: A Knowledge-Guided Self-Evolving Optimization Framework with LLMs for NOTAM Interpretation
by: Liu, Maoqi, et al.
Published: (2025)
by: Liu, Maoqi, et al.
Published: (2025)
Lossy Semantic Communication for the Logical Deduction of the State of the World
by: Saz, Ahmet Faruk, et al.
Published: (2024)
by: Saz, Ahmet Faruk, et al.
Published: (2024)
How to Alleviate Catastrophic Forgetting in LLMs Finetuning? Hierarchical Layer-Wise and Element-Wise Regularization
by: Song, Shezheng, et al.
Published: (2025)
by: Song, Shezheng, et al.
Published: (2025)
Synthesizing Post-Training Data for LLMs through Multi-Agent Simulation
by: Tang, Shuo, et al.
Published: (2024)
by: Tang, Shuo, et al.
Published: (2024)
Guided Self-Evolving LLMs with Minimal Human Supervision
by: Yu, Wenhao, et al.
Published: (2025)
by: Yu, Wenhao, et al.
Published: (2025)
SEW: Self-Evolving Agentic Workflows for Automated Code Generation
by: Liu, Siwei, et al.
Published: (2025)
by: Liu, Siwei, et al.
Published: (2025)
Learning to Self-Evolve
by: Chen, Xiaoyin, et al.
Published: (2026)
by: Chen, Xiaoyin, et al.
Published: (2026)
Generative Floor Plan Design with LLMs via Reinforcement Learning with Verifiable Rewards
by: Lara, Luis, et al.
Published: (2026)
by: Lara, Luis, et al.
Published: (2026)
Large Language Models Can Self-Improve in Long-context Reasoning
by: Li, Siheng, et al.
Published: (2024)
by: Li, Siheng, et al.
Published: (2024)
Systematic Analysis of LLM Contributions to Planning: Solver, Verifier, Heuristic
by: Li, Haoming, et al.
Published: (2024)
by: Li, Haoming, et al.
Published: (2024)
Language Models as Continuous Self-Evolving Data Engineers
by: Wang, Peidong, et al.
Published: (2024)
by: Wang, Peidong, et al.
Published: (2024)
Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforcement Planning
by: Liu, Dong, et al.
Published: (2026)
by: Liu, Dong, et al.
Published: (2026)
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
by: Luo, Xiaoyu, et al.
Published: (2025)
by: Luo, Xiaoyu, et al.
Published: (2025)
Frontier-Eng: Benchmarking Self-Evolving Agents on Real-World Engineering Tasks with Generative Optimization
by: Chi, Yizhe, et al.
Published: (2026)
by: Chi, Yizhe, et al.
Published: (2026)
Evolving and Executing Research Plans via Double-Loop Multi-Agent Collaboration
by: Zhang, Zhi, et al.
Published: (2025)
by: Zhang, Zhi, et al.
Published: (2025)
Mem$^2$Evolve: Towards Self-Evolving Agents via Co-Evolutionary Capability Expansion and Experience Distillation
by: Cheng, Zihao, et al.
Published: (2026)
by: Cheng, Zihao, et al.
Published: (2026)
LLM-Augmented Symbolic Reinforcement Learning with Landmark-Based Task Decomposition
by: Kheirandish, Alireza, et al.
Published: (2024)
by: Kheirandish, Alireza, et al.
Published: (2024)
Achieving Tokenizer Flexibility in Language Models through Heuristic Adaptation and Supertoken Learning
by: Sharthak, Shaurya, et al.
Published: (2025)
by: Sharthak, Shaurya, et al.
Published: (2025)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
by: Liu, Youwei, et al.
Published: (2026)
by: Liu, Youwei, et al.
Published: (2026)
Emulating Clinician Cognition via Self-Evolving Deep Clinical Research
by: Ren, Ruiyang, et al.
Published: (2026)
by: Ren, Ruiyang, et al.
Published: (2026)
Similar Items
-
Scaling Search-Augmented LLM Reasoning via Adaptive Information Control
by: Xiong, Siheng, et al.
Published: (2026) -
Enhancing Long Chain-of-Thought Reasoning through Multi-Path Plan Aggregation
by: Xiong, Siheng, et al.
Published: (2025) -
Deliberate Reasoning in Language Models as Structure-Aware Planning with an Accurate World Model
by: Xiong, Siheng, et al.
Published: (2024) -
The Compressor-Retriever Architecture for Language Model OS
by: Yang, Yuan, et al.
Published: (2024) -
Large Language Models Can Learn Temporal Reasoning
by: Xiong, Siheng, et al.
Published: (2024)