Recon, Answer, Verify: Agents in Search of Truth
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shukla, Satyam, Dutta, Himanshu, Bhattacharyya, Pushpak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GRAFT: A Graph-based Flow-aware Agentic Framework for Document-level Machine Translation
von: Dutta, Himanshu, et al.
Veröffentlicht: (2025)
von: Dutta, Himanshu, et al.
Veröffentlicht: (2025)
"I understand why I got this grade": Automatic Short Answer Grading with Feedback
von: Aggarwal, Dishank, et al.
Veröffentlicht: (2024)
von: Aggarwal, Dishank, et al.
Veröffentlicht: (2024)
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
von: Ahmad, Arif, et al.
Veröffentlicht: (2024)
von: Ahmad, Arif, et al.
Veröffentlicht: (2024)
Debating with More Persuasive LLMs Leads to More Truthful Answers
von: Khan, Akbir, et al.
Veröffentlicht: (2024)
von: Khan, Akbir, et al.
Veröffentlicht: (2024)
Can We Verify Step by Step for Incorrect Answer Detection?
von: Xu, Xin, et al.
Veröffentlicht: (2024)
von: Xu, Xin, et al.
Veröffentlicht: (2024)
Yes, this is what I was looking for! Towards Multi-modal Medical Consultation Concern Summary Generation
von: Tiwari, Abhisek, et al.
Veröffentlicht: (2024)
von: Tiwari, Abhisek, et al.
Veröffentlicht: (2024)
Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding
von: Kotaprolu, Hemanth, et al.
Veröffentlicht: (2026)
von: Kotaprolu, Hemanth, et al.
Veröffentlicht: (2026)
How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?
von: Sil, Pritam, et al.
Veröffentlicht: (2026)
von: Sil, Pritam, et al.
Veröffentlicht: (2026)
Unveiling the Invisible: Captioning Videos with Metaphors
von: Kalarani, Abisek Rajakumar, et al.
Veröffentlicht: (2024)
von: Kalarani, Abisek Rajakumar, et al.
Veröffentlicht: (2024)
Understand the Implication: Learning to Think for Pragmatic Understanding
von: Sravanthi, Settaluri Lakshmi, et al.
Veröffentlicht: (2025)
von: Sravanthi, Settaluri Lakshmi, et al.
Veröffentlicht: (2025)
Recon: Reconstruction-Guided Reasoning Synthesis for User Modeling
von: Zhu, Alan, et al.
Veröffentlicht: (2026)
von: Zhu, Alan, et al.
Veröffentlicht: (2026)
Correct Answers from Sound Reasoning: Verifiable Process Supervision for Language Models
von: Kim, Kyuyoung, et al.
Veröffentlicht: (2026)
von: Kim, Kyuyoung, et al.
Veröffentlicht: (2026)
ToxVidLM: A Multimodal Framework for Toxicity Detection in Code-Mixed Videos
von: Maity, Krishanu, et al.
Veröffentlicht: (2024)
von: Maity, Krishanu, et al.
Veröffentlicht: (2024)
Collective Reasoning Among LLMs: A Framework for Answer Validation Without Ground Truth
von: Davoudi, Seyed Pouyan Mousavi, et al.
Veröffentlicht: (2025)
von: Davoudi, Seyed Pouyan Mousavi, et al.
Veröffentlicht: (2025)
Mental Disorder Classification via Temporal Representation of Text
von: Kumar, Raja, et al.
Veröffentlicht: (2024)
von: Kumar, Raja, et al.
Veröffentlicht: (2024)
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
von: Li, Kenneth, et al.
Veröffentlicht: (2023)
von: Li, Kenneth, et al.
Veröffentlicht: (2023)
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024)
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024)
EVE-Agent: Evidence-Verifiable Self-Evolving Agents
von: Arai, Yamato, et al.
Veröffentlicht: (2026)
von: Arai, Yamato, et al.
Veröffentlicht: (2026)
ORBIT: Scalable and Verifiable Data Generation for Search Agents on a Tight Budget
von: Thakur, Nandan, et al.
Veröffentlicht: (2026)
von: Thakur, Nandan, et al.
Veröffentlicht: (2026)
VerifiAgent: a Unified Verification Agent in Language Model Reasoning
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
TruthStance: An Annotated Dataset of Conversations on Truth Social
von: Ameen, Fathima, et al.
Veröffentlicht: (2026)
von: Ameen, Fathima, et al.
Veröffentlicht: (2026)
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)
StackRAG Agent: Improving Developer Answers with Retrieval-Augmented Generation
von: Abrahamyan, Davit, et al.
Veröffentlicht: (2024)
von: Abrahamyan, Davit, et al.
Veröffentlicht: (2024)
DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search
von: Wu, Fang, et al.
Veröffentlicht: (2025)
von: Wu, Fang, et al.
Veröffentlicht: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)
DocCGen: Document-based Controlled Code Generation
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2024)
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2024)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
AI-LieDar: Examine the Trade-off Between Utility and Truthfulness in LLM Agents
von: Su, Zhe, et al.
Veröffentlicht: (2024)
von: Su, Zhe, et al.
Veröffentlicht: (2024)
Dr. Bench: A Multidimensional Evaluation for Deep Research Agents, from Answers to Reports
von: Yao, Yang, et al.
Veröffentlicht: (2025)
von: Yao, Yang, et al.
Veröffentlicht: (2025)
Truth Knows No Language: Evaluating Truthfulness Beyond English
von: Figueras, Blanca Calvo, et al.
Veröffentlicht: (2025)
von: Figueras, Blanca Calvo, et al.
Veröffentlicht: (2025)
Re-Search for The Truth: Multi-round Retrieval-augmented Large Language Models are Strong Fake News Detectors
von: Li, Guanghua, et al.
Veröffentlicht: (2024)
von: Li, Guanghua, et al.
Veröffentlicht: (2024)
Towards Autonomous Agents: Adaptive-planning, Reasoning, and Acting in Language Models
von: Dutta, Abhishek, et al.
Veröffentlicht: (2024)
von: Dutta, Abhishek, et al.
Veröffentlicht: (2024)
How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs
von: Adarsh, Shivam, et al.
Veröffentlicht: (2026)
von: Adarsh, Shivam, et al.
Veröffentlicht: (2026)
Sandwich Reasoning: An Answer-Reasoning-Answer Approach for Low-Latency Query Correction
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
von: Wiegreffe, Sarah, et al.
Veröffentlicht: (2024)
von: Wiegreffe, Sarah, et al.
Veröffentlicht: (2024)
Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
Toward Verifiable Misinformation Detection: A Multi-Tool LLM Agent Framework
von: Cui, Zikun, et al.
Veröffentlicht: (2025)
von: Cui, Zikun, et al.
Veröffentlicht: (2025)
On Verifiable Legal Reasoning: A Multi-Agent Framework with Formalized Knowledge Representations
von: Sadowski, Albert, et al.
Veröffentlicht: (2025)
von: Sadowski, Albert, et al.
Veröffentlicht: (2025)
No Answer Needed: Predicting LLM Answer Accuracy from Question-Only Linear Probes
von: Cencerrado, Iván Vicente Moreno, et al.
Veröffentlicht: (2025)
von: Cencerrado, Iván Vicente Moreno, et al.
Veröffentlicht: (2025)
RLVER: Reinforcement Learning with Verifiable Emotion Rewards for Empathetic Agents
von: Wang, Peisong, et al.
Veröffentlicht: (2025)
von: Wang, Peisong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GRAFT: A Graph-based Flow-aware Agentic Framework for Document-level Machine Translation
von: Dutta, Himanshu, et al.
Veröffentlicht: (2025) -
"I understand why I got this grade": Automatic Short Answer Grading with Feedback
von: Aggarwal, Dishank, et al.
Veröffentlicht: (2024) -
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
von: Ahmad, Arif, et al.
Veröffentlicht: (2024) -
Debating with More Persuasive LLMs Leads to More Truthful Answers
von: Khan, Akbir, et al.
Veröffentlicht: (2024) -
Can We Verify Step by Step for Incorrect Answer Detection?
von: Xu, Xin, et al.
Veröffentlicht: (2024)