Unleashing the True Potential of LLMs: A Feedback-Triggered Self-Correction with Long-Term Multipath Decoding
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Jipeng, Gao, Zeyu, Qi, Yubin, Dong, Hande, Chen, Weijian, Lin, Qiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GAPO: Robust Advantage Estimation for Real-World Code LLMs
di: Zhang, Jianqing, et al.
Pubblicazione: (2025)
di: Zhang, Jianqing, et al.
Pubblicazione: (2025)
Learning to Check: Unleashing Potentials for Self-Correction in Large Language Models
di: Zhang, Che, et al.
Pubblicazione: (2024)
di: Zhang, Che, et al.
Pubblicazione: (2024)
Self-Correction as Feedback Control: Error Dynamics, Stability Thresholds, and Prompt Interventions in LLMs
di: Liu, Aofan, et al.
Pubblicazione: (2026)
di: Liu, Aofan, et al.
Pubblicazione: (2026)
GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback
di: Xu, Ruiyao, et al.
Pubblicazione: (2026)
di: Xu, Ruiyao, et al.
Pubblicazione: (2026)
Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)
Dynamic Knowledge Exchange and Dual-diversity Review: Concisely Unleashing the Potential of a Multi-Agent Research Team
di: Yu, Weilun, et al.
Pubblicazione: (2025)
di: Yu, Weilun, et al.
Pubblicazione: (2025)
Temporal User Profiling with LLMs: Balancing Short-Term and Long-Term Preferences for Recommendations
di: Sabouri, Milad, et al.
Pubblicazione: (2025)
di: Sabouri, Milad, et al.
Pubblicazione: (2025)
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
di: Zhang, Chaowei, et al.
Pubblicazione: (2026)
di: Zhang, Chaowei, et al.
Pubblicazione: (2026)
Re-Triggering Safeguards within LLMs for Jailbreak Detection
di: Lin, Zheng, et al.
Pubblicazione: (2026)
di: Lin, Zheng, et al.
Pubblicazione: (2026)
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
di: Ohashi, Atsumoto, et al.
Pubblicazione: (2024)
di: Ohashi, Atsumoto, et al.
Pubblicazione: (2024)
RED: Unleashing Token-Level Rewards from Holistic Feedback via Reward Redistribution
di: Li, Jiahui, et al.
Pubblicazione: (2024)
di: Li, Jiahui, et al.
Pubblicazione: (2024)
True Zero-Shot Inference of Dynamical Systems Preserving Long-Term Statistics
di: Hemmer, Christoph Jürgen, et al.
Pubblicazione: (2025)
di: Hemmer, Christoph Jürgen, et al.
Pubblicazione: (2025)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
SATURN: SAT-based Reinforcement Learning to Unleash LLMs Reasoning
di: Liu, Huanyu, et al.
Pubblicazione: (2025)
di: Liu, Huanyu, et al.
Pubblicazione: (2025)
Unleash LLMs Potential for Recommendation by Coordinating Twin-Tower Dynamic Semantic Token Generator
di: Yin, Jun, et al.
Pubblicazione: (2024)
di: Yin, Jun, et al.
Pubblicazione: (2024)
Learning to Decode in Parallel: Self-Coordinating Neural Network for Real-Time Quantum Error Correction
di: Zhang, Kai, et al.
Pubblicazione: (2026)
di: Zhang, Kai, et al.
Pubblicazione: (2026)
Long Term Memory: The Foundation of AI Self-Evolution
di: Jiang, Xun, et al.
Pubblicazione: (2024)
di: Jiang, Xun, et al.
Pubblicazione: (2024)
ISMRNN: An Implicitly Segmented RNN Method with Mamba for Long-Term Time Series Forecasting
di: Zhao, GaoXiang, et al.
Pubblicazione: (2024)
di: Zhao, GaoXiang, et al.
Pubblicazione: (2024)
Source-Free Cross-Modal Knowledge Transfer by Unleashing the Potential of Task-Irrelevant Data
di: Zhu, Jinjing, et al.
Pubblicazione: (2024)
di: Zhu, Jinjing, et al.
Pubblicazione: (2024)
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection
di: Zhang, Chaowei, et al.
Pubblicazione: (2025)
di: Zhang, Chaowei, et al.
Pubblicazione: (2025)
DeepInnovator: Triggering the Innovative Capabilities of LLMs
di: Fan, Tianyu, et al.
Pubblicazione: (2026)
di: Fan, Tianyu, et al.
Pubblicazione: (2026)
ChiseLLM: Unleashing the Power of Reasoning LLMs for Chisel Agile Hardware Development
di: Wang, Bowei, et al.
Pubblicazione: (2025)
di: Wang, Bowei, et al.
Pubblicazione: (2025)
MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO
di: Xiao, Yicheng, et al.
Pubblicazione: (2025)
di: Xiao, Yicheng, et al.
Pubblicazione: (2025)
Price of Fairness in Short-Term and Long-Term Algorithmic Selections
di: Jabbari, Shahin, et al.
Pubblicazione: (2026)
di: Jabbari, Shahin, et al.
Pubblicazione: (2026)
Scheduling Your LLM Reinforcement Learning with Reasoning Trees
di: Wang, Hong, et al.
Pubblicazione: (2025)
di: Wang, Hong, et al.
Pubblicazione: (2025)
ReCreate: Reasoning and Creating Domain Agents Driven by Experience
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)
Reconstructing Deep Neural Networks: Unleashing the Optimization Potential of Natural Gradient Descent
di: Liu, Weihua, et al.
Pubblicazione: (2024)
di: Liu, Weihua, et al.
Pubblicazione: (2024)
SCANet: Correcting LEGO Assembly Errors with Self-Correct Assembly Network
di: Wan, Yuxuan, et al.
Pubblicazione: (2024)
di: Wan, Yuxuan, et al.
Pubblicazione: (2024)
Personalizing LLMs with Binary Feedback: A Preference-Corrected Optimization Framework
di: Ma, Xilai, et al.
Pubblicazione: (2026)
di: Ma, Xilai, et al.
Pubblicazione: (2026)
Enhancing the Medical Context-Awareness Ability of LLMs via Multifaceted Self-Refinement Learning
di: Zhou, Yuxuan, et al.
Pubblicazione: (2025)
di: Zhou, Yuxuan, et al.
Pubblicazione: (2025)
LEPO: Latent Reasoning Policy Optimization for Large Language Models
di: Zhou, Yuyan, et al.
Pubblicazione: (2026)
di: Zhou, Yuyan, et al.
Pubblicazione: (2026)
Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving
di: Zheng, Yinan, et al.
Pubblicazione: (2026)
di: Zheng, Yinan, et al.
Pubblicazione: (2026)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
di: Banerjee, Tanushree, et al.
Pubblicazione: (2024)
di: Banerjee, Tanushree, et al.
Pubblicazione: (2024)
Unleashing the Potential of Two-Tower Models: Diffusion-Based Cross-Interaction for Large-Scale Matching
di: Wang, Yihan, et al.
Pubblicazione: (2025)
di: Wang, Yihan, et al.
Pubblicazione: (2025)
PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
Decoding Cortical Microcircuits: A Generative Model for Latent Space Exploration and Controlled Synthesis
di: Liu, Xingyu, et al.
Pubblicazione: (2025)
di: Liu, Xingyu, et al.
Pubblicazione: (2025)
SynGR: Unleashing the Potential of Cross-Modal Synergy for Generative Recommendation
di: Chen, Wei, et al.
Pubblicazione: (2026)
di: Chen, Wei, et al.
Pubblicazione: (2026)
EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control
di: Yang, Kai, et al.
Pubblicazione: (2025)
di: Yang, Kai, et al.
Pubblicazione: (2025)
BadMoE: Backdooring Mixture-of-Experts LLMs via Optimizing Routing Triggers and Infecting Dormant Experts
di: Wang, Qingyue, et al.
Pubblicazione: (2025)
di: Wang, Qingyue, et al.
Pubblicazione: (2025)
LycheeDecode: Accelerating Long-Context LLM Inference via Hybrid-Head Sparse Decoding
di: Lin, Gang, et al.
Pubblicazione: (2026)
di: Lin, Gang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
GAPO: Robust Advantage Estimation for Real-World Code LLMs
di: Zhang, Jianqing, et al.
Pubblicazione: (2025) -
Learning to Check: Unleashing Potentials for Self-Correction in Large Language Models
di: Zhang, Che, et al.
Pubblicazione: (2024) -
Self-Correction as Feedback Control: Error Dynamics, Stability Thresholds, and Prompt Interventions in LLMs
di: Liu, Aofan, et al.
Pubblicazione: (2026) -
GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback
di: Xu, Ruiyao, et al.
Pubblicazione: (2026) -
Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)