A Debate-Driven Experiment on LLM Hallucinations and Accuracy
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Ray, Bagade, Tanishka, Martinez, Kevin, Yasmin, Flora, Ayala, Grant, Lam, Michael, Zhu, Kevin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Town Hall Debate Prompting: Enhancing Logical Reasoning in LLMs through Multi-Persona Interaction
by: Sandwar, Vivaan, et al.
Published: (2025)
by: Sandwar, Vivaan, et al.
Published: (2025)
Removal of Hallucination on Hallucination: Debate-Augmented RAG
by: Hu, Wentao, et al.
Published: (2025)
by: Hu, Wentao, et al.
Published: (2025)
ProofSketch: Efficient Verified Reasoning for Large Language Models
by: Sheshanarayana, Disha, et al.
Published: (2025)
by: Sheshanarayana, Disha, et al.
Published: (2025)
CLAIM: An Intent-Driven Multi-Agent Framework for Analyzing Manipulation in Courtroom Dialogues
by: Sheshanarayana, Disha, et al.
Published: (2025)
by: Sheshanarayana, Disha, et al.
Published: (2025)
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
by: Mao, Nathan, et al.
Published: (2026)
by: Mao, Nathan, et al.
Published: (2026)
Counterfactual Debating with Preset Stances for Hallucination Elimination of LLMs
by: Fang, Yi, et al.
Published: (2024)
by: Fang, Yi, et al.
Published: (2024)
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
by: Chen, Xuhang, et al.
Published: (2025)
by: Chen, Xuhang, et al.
Published: (2025)
Evaluating LLM-Driven Summarisation of Parliamentary Debates with Computational Argumentation
by: Cunningham, Eoghan, et al.
Published: (2026)
by: Cunningham, Eoghan, et al.
Published: (2026)
Pragmatic Metacognitive Prompting Improves LLM Performance on Sarcasm Detection
by: Lee, Joshua, et al.
Published: (2024)
by: Lee, Joshua, et al.
Published: (2024)
LLM Hallucination Detection: HSAD
by: Li, JinXin, et al.
Published: (2025)
by: Li, JinXin, et al.
Published: (2025)
Hierarchical Text Classification Using Contrastive Learning Informed Path Guided Hierarchy
by: Agrawal, Neeraj, et al.
Published: (2025)
by: Agrawal, Neeraj, et al.
Published: (2025)
COMPASS: Context-Modulated PID Attention Steering System for Hallucination Mitigation
by: Sahay, Kenji, et al.
Published: (2025)
by: Sahay, Kenji, et al.
Published: (2025)
On Mitigating Code LLM Hallucinations with API Documentation
by: Jain, Nihal, et al.
Published: (2024)
by: Jain, Nihal, et al.
Published: (2024)
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks
by: Yugeswardeenoo, Dharunish, et al.
Published: (2024)
by: Yugeswardeenoo, Dharunish, et al.
Published: (2024)
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases
by: Li, Jiarui, et al.
Published: (2024)
by: Li, Jiarui, et al.
Published: (2024)
Enhancing Depression Diagnosis with Chain-of-Thought Prompting
by: Shi, Elysia, et al.
Published: (2024)
by: Shi, Elysia, et al.
Published: (2024)
Enhancing Knowledge Distillation for LLMs with Response-Priming Prompting
by: Goyal, Vijay, et al.
Published: (2024)
by: Goyal, Vijay, et al.
Published: (2024)
InterpDetect: Interpretable Signals for Detecting Hallucinations in Retrieval-Augmented Generation
by: Tan, Likun, et al.
Published: (2025)
by: Tan, Likun, et al.
Published: (2025)
A Comparative Study of Translation Bias and Accuracy in Multilingual Large Language Models for Cross-Language Claim Verification
by: Singhal, Aryan, et al.
Published: (2024)
by: Singhal, Aryan, et al.
Published: (2024)
Towards Detecting LLMs Hallucination via Markov Chain-based Multi-agent Debate Framework
by: Sun, Xiaoxi, et al.
Published: (2024)
by: Sun, Xiaoxi, et al.
Published: (2024)
Enhancing Short-Text Topic Modeling with LLM-Driven Context Expansion and Prefix-Tuned VAEs
by: Akash, Pritom Saha, et al.
Published: (2024)
by: Akash, Pritom Saha, et al.
Published: (2024)
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation
by: Sternlicht, Noy, et al.
Published: (2025)
by: Sternlicht, Noy, et al.
Published: (2025)
How to Detect and Defeat Molecular Mirage: A Metric-Driven Benchmark for Hallucination in LLM-based Molecular Comprehension
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
HalluCana: Fixing LLM Hallucination with A Canary Lookahead
by: Li, Tianyi, et al.
Published: (2024)
by: Li, Tianyi, et al.
Published: (2024)
Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate
by: Lu, Zhixiang, et al.
Published: (2026)
by: Lu, Zhixiang, et al.
Published: (2026)
Attribution Techniques for Mitigating Hallucinated Information in RAG Systems: A Survey
by: Zhao, Yuqing, et al.
Published: (2026)
by: Zhao, Yuqing, et al.
Published: (2026)
Hypothesis Generation for Materials Discovery and Design Using Goal-Driven and Constraint-Guided LLM Agents
by: Kumbhar, Shrinidhi, et al.
Published: (2025)
by: Kumbhar, Shrinidhi, et al.
Published: (2025)
Beyond Accuracy: Rethinking Hallucination and Regulatory Response in Generative AI
by: Li, Zihao, et al.
Published: (2025)
by: Li, Zihao, et al.
Published: (2025)
Scaling Natural-Language Graph-Based Test Time Compute for Automated Theorem Proving
by: Li, Vincent, et al.
Published: (2025)
by: Li, Vincent, et al.
Published: (2025)
InspireDebate: Multi-Dimensional Subjective-Objective Evaluation-Guided Reasoning and Optimization for Debating
by: Wang, Fuyu, et al.
Published: (2025)
by: Wang, Fuyu, et al.
Published: (2025)
Beyond Accuracy: Risk-Sensitive Evaluation of Hallucinated Medical Advice
by: Doshi, Savan
Published: (2026)
by: Doshi, Savan
Published: (2026)
Distributional Semantics Tracing: A Framework for Explaining Hallucinations in Large Language Models
by: Bhatia, Gagan, et al.
Published: (2025)
by: Bhatia, Gagan, et al.
Published: (2025)
FRED: Financial Retrieval-Enhanced Detection and Editing of Hallucinations in Language Models
by: Tan, Likun, et al.
Published: (2025)
by: Tan, Likun, et al.
Published: (2025)
NusaMT-7B: Machine Translation for Low-Resource Indonesian Languages with Large Language Models
by: Tan, William, et al.
Published: (2024)
by: Tan, William, et al.
Published: (2024)
Probing LLM Hallucination from Within: Perturbation-Driven Approach via Internal Knowledge
by: Lee, Seongmin, et al.
Published: (2024)
by: Lee, Seongmin, et al.
Published: (2024)
AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents
by: Liu, Xuannan, et al.
Published: (2026)
by: Liu, Xuannan, et al.
Published: (2026)
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination
by: Li, Zhuo, et al.
Published: (2026)
by: Li, Zhuo, et al.
Published: (2026)
Latent Debate: A Surrogate Framework for Interpreting LLM Thinking
by: Chen, Lihu, et al.
Published: (2025)
by: Chen, Lihu, et al.
Published: (2025)
Detecting Hallucinations in Authentic LLM-Human Interactions
by: Ren, Yujie, et al.
Published: (2025)
by: Ren, Yujie, et al.
Published: (2025)
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
by: Zhu, Runchuan, et al.
Published: (2025)
by: Zhu, Runchuan, et al.
Published: (2025)
Similar Items
-
Town Hall Debate Prompting: Enhancing Logical Reasoning in LLMs through Multi-Persona Interaction
by: Sandwar, Vivaan, et al.
Published: (2025) -
Removal of Hallucination on Hallucination: Debate-Augmented RAG
by: Hu, Wentao, et al.
Published: (2025) -
ProofSketch: Efficient Verified Reasoning for Large Language Models
by: Sheshanarayana, Disha, et al.
Published: (2025) -
CLAIM: An Intent-Driven Multi-Agent Framework for Analyzing Manipulation in Courtroom Dialogues
by: Sheshanarayana, Disha, et al.
Published: (2025) -
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
by: Mao, Nathan, et al.
Published: (2026)