Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Wenhao, Lin, Chenchen, Chen, Jian, Xu, Jinfeng, Wang, Xuehe, Ngai, Edith Cheuk Han |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Belief-Guided Inference Control for Large Language Model Services via Verifiable Observations
von: Yuan, Wenhao, et al.
Veröffentlicht: (2026)
von: Yuan, Wenhao, et al.
Veröffentlicht: (2026)
LATTE: Forecasting Peer Anchored Preference Trajectories for Personalized LLM Generation
von: Li, Jinze, et al.
Veröffentlicht: (2026)
von: Li, Jinze, et al.
Veröffentlicht: (2026)
OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory
von: Li, Jinze, et al.
Veröffentlicht: (2026)
von: Li, Jinze, et al.
Veröffentlicht: (2026)
Walk Before You Run! Concise LLM Reasoning via Reinforcement Learning
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
von: Lin, Junda, et al.
Veröffentlicht: (2026)
von: Lin, Junda, et al.
Veröffentlicht: (2026)
CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era
von: Shi, Kaiwen, et al.
Veröffentlicht: (2026)
von: Shi, Kaiwen, et al.
Veröffentlicht: (2026)
Look Before You Leap: Autonomous Exploration for LLM Agents
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
von: Lin, Wei-Hsiang, et al.
Veröffentlicht: (2025)
von: Lin, Wei-Hsiang, et al.
Veröffentlicht: (2025)
s3: You Don't Need That Much Data to Train a Search Agent via RL
von: Jiang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Jiang, Pengcheng, et al.
Veröffentlicht: (2025)
Reason and Verify: A Framework for Faithful Retrieval-Augmented Generation
von: Khan, Eeham, et al.
Veröffentlicht: (2026)
von: Khan, Eeham, et al.
Veröffentlicht: (2026)
Look Before You Leap: Towards Decision-Aware and Generalizable Tool-Usage for Large Language Models
von: Gui, Anchun, et al.
Veröffentlicht: (2024)
von: Gui, Anchun, et al.
Veröffentlicht: (2024)
LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
Multi-Hop Privacy Propagation for Differentially Private Federated Learning in Social Networks
von: Lin, Chenchen, et al.
Veröffentlicht: (2025)
von: Lin, Chenchen, et al.
Veröffentlicht: (2025)
Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
NeuroFaith: Evaluating LLM Self-Explanation Faithfulness via Internal Representation Alignment
von: Bhan, Milan, et al.
Veröffentlicht: (2025)
von: Bhan, Milan, et al.
Veröffentlicht: (2025)
Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
Think Before You Prune: Selective Self-Generated Calibration for Pruning Large Reasoning Models
von: Xiang, Yang, et al.
Veröffentlicht: (2025)
von: Xiang, Yang, et al.
Veröffentlicht: (2025)
Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2025)
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2025)
Think Before You Lie: How Reasoning Leads to Honesty
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces
von: Zhang, Chenchen
Veröffentlicht: (2026)
von: Zhang, Chenchen
Veröffentlicht: (2026)
Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
von: Luo, Jiani, et al.
Veröffentlicht: (2026)
von: Luo, Jiani, et al.
Veröffentlicht: (2026)
Self-Train Before You Transcribe
von: Flynn, Robert, et al.
Veröffentlicht: (2024)
von: Flynn, Robert, et al.
Veröffentlicht: (2024)
Think Twice Before You Write -- an Entropy-based Decoding Strategy to Enhance LLM Reasoning
von: He, Jiashu, et al.
Veröffentlicht: (2026)
von: He, Jiashu, et al.
Veröffentlicht: (2026)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
DSVD: Dynamic Self-Verify Decoding for Faithful Generation in Large Language Models
von: Guo, YiQiu, et al.
Veröffentlicht: (2025)
von: Guo, YiQiu, et al.
Veröffentlicht: (2025)
Commitment Checklist: Auditing Author Commitments in Peer Review
von: Chen, Chung-Chi, et al.
Veröffentlicht: (2026)
von: Chen, Chung-Chi, et al.
Veröffentlicht: (2026)
RAMA: Retrieval-Augmented Multi-Agent Framework for Misinformation Detection in Multimodal Fact-Checking
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
VerifiAgent: a Unified Verification Agent in Language Model Reasoning
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
QI-DPFL: Quality-Aware and Incentive-Boosted Federated Learning with Differential Privacy
von: Yuan, Wenhao, et al.
Veröffentlicht: (2024)
von: Yuan, Wenhao, et al.
Veröffentlicht: (2024)
FedAgg: Adaptive Federated Learning with Aggregated Gradients
von: Yuan, Wenhao, et al.
Veröffentlicht: (2023)
von: Yuan, Wenhao, et al.
Veröffentlicht: (2023)
A Game-Theoretic Framework for Privacy-Aware Client Sampling in Federated Learning
von: Yuan, Wenhao, et al.
Veröffentlicht: (2024)
von: Yuan, Wenhao, et al.
Veröffentlicht: (2024)
Read Before You Think: Mitigating LLM Comprehension Failures with Step-by-Step Reading
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
Exploring Knowledge Conflicts for Faithful LLM Reasoning: Benchmark and Method
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2026)
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2026)
Cite Before You Speak: Enhancing Context-Response Grounding in E-commerce Conversational LLM-Agents
von: Zeng, Jingying, et al.
Veröffentlicht: (2025)
von: Zeng, Jingying, et al.
Veröffentlicht: (2025)
Faithful Logical Reasoning via Symbolic Chain-of-Thought
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
Could Thinking Multilingually Empower LLM Reasoning?
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
PAG: Multi-Turn Reinforced LLM Self-Correction with Policy as Generative Verifier
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
SEVerA: Verified Synthesis of Self-Evolving Agents
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2026)
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2026)
Compress the Context, Keep the Commitments: A Formal Framework for Verifiable LLM Context Compression
von: Trukhina, Natalia, et al.
Veröffentlicht: (2026)
von: Trukhina, Natalia, et al.
Veröffentlicht: (2026)
Can LLMs Produce Faithful Explanations For Fact-checking? Towards Faithful Explainable Fact-Checking via Multi-Agent Debate
von: Kim, Kyungha, et al.
Veröffentlicht: (2024)
von: Kim, Kyungha, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Belief-Guided Inference Control for Large Language Model Services via Verifiable Observations
von: Yuan, Wenhao, et al.
Veröffentlicht: (2026) -
LATTE: Forecasting Peer Anchored Preference Trajectories for Personalized LLM Generation
von: Li, Jinze, et al.
Veröffentlicht: (2026) -
OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory
von: Li, Jinze, et al.
Veröffentlicht: (2026) -
Walk Before You Run! Concise LLM Reasoning via Reinforcement Learning
von: Song, Mingyang, et al.
Veröffentlicht: (2025) -
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
von: Lin, Junda, et al.
Veröffentlicht: (2026)