Unlocking Recursive Thinking of LLMs: Alignment via Refinement
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Haoke, Liang, Xiaobo, Wang, Cunxiang, Li, Juntao, Zhang, Min |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging
par: Liang, Xiaobo, et autres
Publié: (2024)
par: Liang, Xiaobo, et autres
Publié: (2024)
LongRM: Revealing and Unlocking the Context Boundary of Reward Modeling
par: Tang, Zecheng, et autres
Publié: (2025)
par: Tang, Zecheng, et autres
Publié: (2025)
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement
par: Li, Guanghao, et autres
Publié: (2025)
par: Li, Guanghao, et autres
Publié: (2025)
Unleashing LLM Reasoning Capability via Scalable Question Synthesis from Scratch
par: Ding, Yuyang, et autres
Publié: (2024)
par: Ding, Yuyang, et autres
Publié: (2024)
Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space
par: Zhang, Zhen, et autres
Publié: (2025)
par: Zhang, Zhen, et autres
Publié: (2025)
RTQA : Recursive Thinking for Complex Temporal Knowledge Graph Question Answering with Large Language Models
par: Gong, Zhaoyan, et autres
Publié: (2025)
par: Gong, Zhaoyan, et autres
Publié: (2025)
How Likely Do LLMs with CoT Mimic Human Reasoning?
par: Bao, Guangsheng, et autres
Publié: (2024)
par: Bao, Guangsheng, et autres
Publié: (2024)
Search and Refine During Think: Facilitating Knowledge Refinement for Improved Retrieval-Augmented Reasoning
par: Shi, Yaorui, et autres
Publié: (2025)
par: Shi, Yaorui, et autres
Publié: (2025)
Direct Alignment of Language Models via Quality-Aware Self-Refinement
par: Yu, Runsheng, et autres
Publié: (2024)
par: Yu, Runsheng, et autres
Publié: (2024)
Think Carefully and Check Again! Meta-Generation Unlocking LLMs for Low-Resource Cross-Lingual Summarization
par: Li, Zhecheng, et autres
Publié: (2024)
par: Li, Zhecheng, et autres
Publié: (2024)
MatryoshkaThinking: Recursive Test-Time Scaling Enables Efficient Reasoning
par: Chen, Hongwei, et autres
Publié: (2025)
par: Chen, Hongwei, et autres
Publié: (2025)
LOGO -- Long cOntext aliGnment via efficient preference Optimization
par: Tang, Zecheng, et autres
Publié: (2024)
par: Tang, Zecheng, et autres
Publié: (2024)
Condor: Enhance LLM Alignment with Knowledge-Driven Data Synthesis and Refinement
par: Cao, Maosong, et autres
Publié: (2025)
par: Cao, Maosong, et autres
Publié: (2025)
Thinker: Training LLMs in Hierarchical Thinking for Deep Search via Multi-Turn Interaction
par: Xu, Jun, et autres
Publié: (2025)
par: Xu, Jun, et autres
Publié: (2025)
Thinking in Latents: Adaptive Anchor Refinement for Implicit Reasoning in LLMs
par: Sheshanarayana, Disha, et autres
Publié: (2026)
par: Sheshanarayana, Disha, et autres
Publié: (2026)
PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference
par: Ji, Jiaming, et autres
Publié: (2024)
par: Ji, Jiaming, et autres
Publié: (2024)
Unlocking the Power of Large Language Models for Entity Alignment
par: Jiang, Xuhui, et autres
Publié: (2024)
par: Jiang, Xuhui, et autres
Publié: (2024)
KAG-Thinker: Interactive Thinking and Deep Reasoning in LLMs via Knowledge-Augmented Generation
par: Zhang, Dalong, et autres
Publié: (2025)
par: Zhang, Dalong, et autres
Publié: (2025)
Semantic Refinement with LLMs for Graph Representations
par: Thapaliya, Safal, et autres
Publié: (2025)
par: Thapaliya, Safal, et autres
Publié: (2025)
Step-On-Feet Tuning: Scaling Self-Alignment of LLMs via Bootstrapping
par: Wang, Haoyu, et autres
Publié: (2024)
par: Wang, Haoyu, et autres
Publié: (2024)
Alignment-Enhanced Decoding:Defending via Token-Level Adaptive Refining of Probability Distributions
par: Liu, Quan, et autres
Publié: (2024)
par: Liu, Quan, et autres
Publié: (2024)
Refining Positive and Toxic Samples for Dual Safety Self-Alignment of LLMs with Minimal Human Interventions
par: Xu, Jingxin, et autres
Publié: (2025)
par: Xu, Jingxin, et autres
Publié: (2025)
TraceSIR: A Multi-Agent Framework for Structured Analysis and Reporting of Agentic Execution Traces
par: Yang, Shu-Xun, et autres
Publié: (2026)
par: Yang, Shu-Xun, et autres
Publié: (2026)
MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading
par: Ji, Baibei, et autres
Publié: (2026)
par: Ji, Baibei, et autres
Publié: (2026)
REAL: Response Embedding-based Alignment for LLMs
par: Zhang, Honggen, et autres
Publié: (2024)
par: Zhang, Honggen, et autres
Publié: (2024)
Direct Value Optimization: Improving Chain-of-Thought Reasoning in LLMs with Refined Values
par: Zhang, Hongbo, et autres
Publié: (2025)
par: Zhang, Hongbo, et autres
Publié: (2025)
Fake Alignment: Are LLMs Really Aligned Well?
par: Wang, Yixu, et autres
Publié: (2023)
par: Wang, Yixu, et autres
Publié: (2023)
Soundwave: Less is More for Speech-Text Alignment in LLMs
par: Zhang, Yuhao, et autres
Publié: (2025)
par: Zhang, Yuhao, et autres
Publié: (2025)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
par: Wang, Qibin, et autres
Publié: (2025)
par: Wang, Qibin, et autres
Publié: (2025)
Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment
par: Sun, Yanru, et autres
Publié: (2025)
par: Sun, Yanru, et autres
Publié: (2025)
ALI-Agent: Assessing LLMs' Alignment with Human Values via Agent-based Evaluation
par: Zheng, Jingnan, et autres
Publié: (2024)
par: Zheng, Jingnan, et autres
Publié: (2024)
Knowledge Conflicts for LLMs: A Survey
par: Xu, Rongwu, et autres
Publié: (2024)
par: Xu, Rongwu, et autres
Publié: (2024)
Think Thrice Before You Act: Progressive Thought Refinement in Large Language Models
par: Du, Chengyu, et autres
Publié: (2024)
par: Du, Chengyu, et autres
Publié: (2024)
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation
par: Li, Ziniu, et autres
Publié: (2025)
par: Li, Ziniu, et autres
Publié: (2025)
Knowledgeable Preference Alignment for LLMs in Domain-specific Question Answering
par: Zhang, Yichi, et autres
Publié: (2023)
par: Zhang, Yichi, et autres
Publié: (2023)
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
par: Cheng, Jiale, et autres
Publié: (2024)
par: Cheng, Jiale, et autres
Publié: (2024)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
par: Zhang, Xiaoying, et autres
Publié: (2024)
par: Zhang, Xiaoying, et autres
Publié: (2024)
Data Analysis and Performance Evaluation of Simulation Deduction Based on LLMs
par: Zhang, Shansi, et autres
Publié: (2025)
par: Zhang, Shansi, et autres
Publié: (2025)
MemLong: Memory-Augmented Retrieval for Long Text Modeling
par: Liu, Weijie, et autres
Publié: (2024)
par: Liu, Weijie, et autres
Publié: (2024)
Revealing and Mitigating the Local Pattern Shortcuts of Mamba
par: You, Wangjie, et autres
Publié: (2024)
par: You, Wangjie, et autres
Publié: (2024)
Documents similaires
-
Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging
par: Liang, Xiaobo, et autres
Publié: (2024) -
LongRM: Revealing and Unlocking the Context Boundary of Reward Modeling
par: Tang, Zecheng, et autres
Publié: (2025) -
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement
par: Li, Guanghao, et autres
Publié: (2025) -
Unleashing LLM Reasoning Capability via Scalable Question Synthesis from Scratch
par: Ding, Yuyang, et autres
Publié: (2024) -
Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space
par: Zhang, Zhen, et autres
Publié: (2025)