Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Bowen, Wang, Nan, Zhou, Yuqing, Pan, Jinhao, Zhu, Ziwei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KnowBias: Mitigating Social Bias in LLMs via Know-Bias Neuron Enhancement
by: Pan, Jinhao, et al.
Published: (2026)
by: Pan, Jinhao, et al.
Published: (2026)
ClawSafety: "Safe" LLMs, Unsafe Agents
by: Wei, Bowen, et al.
Published: (2026)
by: Wei, Bowen, et al.
Published: (2026)
Interpretable Classification via a Rule Network with Selective Logical Operators
by: Wei, Bowen, et al.
Published: (2024)
by: Wei, Bowen, et al.
Published: (2024)
Advancing Interpretability in Text Classification through Prototype Learning
by: Wei, Bowen, et al.
Published: (2024)
by: Wei, Bowen, et al.
Published: (2024)
Uncertain Knowledge Graph Completion via Semi-Supervised Confidence Distribution Learning
by: Wu, Tianxing, et al.
Published: (2025)
by: Wu, Tianxing, et al.
Published: (2025)
Prompt Orchestration Markup Language
by: Zhang, Yuge, et al.
Published: (2025)
by: Zhang, Yuge, et al.
Published: (2025)
SELF: Self-Evolution with Language Feedback
by: Lu, Jianqiao, et al.
Published: (2023)
by: Lu, Jianqiao, et al.
Published: (2023)
SAGE: Multi-Agent Self-Evolution for LLM Reasoning
by: Peng, Yulin, et al.
Published: (2026)
by: Peng, Yulin, et al.
Published: (2026)
iGRPO: Self-Feedback-Driven LLM Reasoning
by: Hatamizadeh, Ali, et al.
Published: (2026)
by: Hatamizadeh, Ali, et al.
Published: (2026)
Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution
by: Maheswaran, Monishwaran, et al.
Published: (2026)
by: Maheswaran, Monishwaran, et al.
Published: (2026)
Agentic Meta-Orchestrator for Multi-task Copilots
by: Zhu, Xiaofeng, et al.
Published: (2025)
by: Zhu, Xiaofeng, et al.
Published: (2025)
Orchestrating Intelligence: Confidence-Aware Routing for Efficient Multi-Agent Collaboration across Multi-Scale Models
by: Wang, Jingbo, et al.
Published: (2026)
by: Wang, Jingbo, et al.
Published: (2026)
Towards Scalable Lightweight GUI Agents via Multi-role Orchestration
by: Wang, Ziwei, et al.
Published: (2026)
by: Wang, Ziwei, et al.
Published: (2026)
Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents
by: Lin, Minhua, et al.
Published: (2026)
by: Lin, Minhua, et al.
Published: (2026)
OMS: On-the-fly, Multi-Objective, Self-Reflective Ad Keyword Generation via LLM Agent
by: Chen, Bowen, et al.
Published: (2025)
by: Chen, Bowen, et al.
Published: (2025)
COGNITION: From Evaluation to Defense against Multimodal LLM CAPTCHA Solvers
by: Wang, Junyu, et al.
Published: (2025)
by: Wang, Junyu, et al.
Published: (2025)
MineEvolve: Self-Evolution with Accumulated Knowledge for Long-Horizon Embodied Minecraft Agents
by: Xie, Zhengwei, et al.
Published: (2026)
by: Xie, Zhengwei, et al.
Published: (2026)
AI Annotation Orchestration: Evaluating LLM verifiers to Improve the Quality of LLM Annotations in Learning Analytics
by: Ahtisham, Bakhtawar, et al.
Published: (2025)
by: Ahtisham, Bakhtawar, et al.
Published: (2025)
Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems
by: Zhu, Junze, et al.
Published: (2026)
by: Zhu, Junze, et al.
Published: (2026)
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
by: Luo, Beier, et al.
Published: (2025)
by: Luo, Beier, et al.
Published: (2025)
A Logical-Rule Autoencoder for Interpretable Recommendations
by: Pan, Jinhao, et al.
Published: (2026)
by: Pan, Jinhao, et al.
Published: (2026)
Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback
by: Wang, Ning, et al.
Published: (2025)
by: Wang, Ning, et al.
Published: (2025)
Rationale-Augmented Retrieval with Constrained LLM Re-Ranking for Task Discovery
by: Wei, Bowen
Published: (2025)
by: Wei, Bowen
Published: (2025)
Z-Space: A Multi-Agent Tool Orchestration Framework for Enterprise-Grade LLM Automation
by: He, Qingsong, et al.
Published: (2025)
by: He, Qingsong, et al.
Published: (2025)
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
by: Xu, Wenda, et al.
Published: (2024)
by: Xu, Wenda, et al.
Published: (2024)
Cost-Aware Model Orchestration for LLM-based Systems
by: Smirnova, Daria, et al.
Published: (2025)
by: Smirnova, Daria, et al.
Published: (2025)
Who is a Better Player: LLM against LLM
by: Zhou, Yingjie, et al.
Published: (2025)
by: Zhou, Yingjie, et al.
Published: (2025)
SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration
by: Zhang, Mingda, et al.
Published: (2026)
by: Zhang, Mingda, et al.
Published: (2026)
DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management
by: Zhou, Zhongchun, et al.
Published: (2025)
by: Zhou, Zhongchun, et al.
Published: (2025)
Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models
by: Dong, Guanting, et al.
Published: (2024)
by: Dong, Guanting, et al.
Published: (2024)
Fact-Level Confidence Calibration and Self-Correction
by: Yuan, Yige, et al.
Published: (2024)
by: Yuan, Yige, et al.
Published: (2024)
Mission Impossible: Feedback-Guided Dynamic Interactive Planning for Improving Reasoning on LLMs
by: Yan, Dong, et al.
Published: (2025)
by: Yan, Dong, et al.
Published: (2025)
Calibrating Verbalized Confidence with Self-Generated Distractors
by: Wang, Victor, et al.
Published: (2025)
by: Wang, Victor, et al.
Published: (2025)
Utility-Guided Agent Orchestration for Efficient LLM Tool Use
by: Liu, Boyan, et al.
Published: (2026)
by: Liu, Boyan, et al.
Published: (2026)
Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution
by: Tian, Zailong, et al.
Published: (2025)
by: Tian, Zailong, et al.
Published: (2025)
Confidence as a Reward: Transforming LLMs into Reward Models
by: Du, He, et al.
Published: (2025)
by: Du, He, et al.
Published: (2025)
The Internet of Large Language Models: An Orchestration Framework for LLM Training and Knowledge Exchange Toward Artificial General Intelligence
by: Wei, Wilson, et al.
Published: (2025)
by: Wei, Wilson, et al.
Published: (2025)
Enhanced Sample Selection with Confidence Tracking: Identifying Correctly Labeled yet Hard-to-Learn Samples in Noisy Data
by: Pan, Weiran, et al.
Published: (2025)
by: Pan, Weiran, et al.
Published: (2025)
Reflective Confidence: Correcting Reasoning Flaws via Online Self-Correction
by: Zeng, Qinglin, et al.
Published: (2025)
by: Zeng, Qinglin, et al.
Published: (2025)
ReVEL: Multi-Turn Reflective LLM-Guided Heuristic Evolution via Structured Performance Feedback
by: Van Duc, Cuong, et al.
Published: (2026)
by: Van Duc, Cuong, et al.
Published: (2026)
Similar Items
-
KnowBias: Mitigating Social Bias in LLMs via Know-Bias Neuron Enhancement
by: Pan, Jinhao, et al.
Published: (2026) -
ClawSafety: "Safe" LLMs, Unsafe Agents
by: Wei, Bowen, et al.
Published: (2026) -
Interpretable Classification via a Rule Network with Selective Logical Operators
by: Wei, Bowen, et al.
Published: (2024) -
Advancing Interpretability in Text Classification through Prototype Learning
by: Wei, Bowen, et al.
Published: (2024) -
Uncertain Knowledge Graph Completion via Semi-Supervised Confidence Distribution Learning
by: Wu, Tianxing, et al.
Published: (2025)