HalluDetect: Detecting, Mitigating, and Benchmarking Hallucinations in Conversational Systems in the Legal Domain
Fuente:
arXiv
Saved in:
| Main Authors: | Anaokar, Spandan, Ganatra, Shrey, Kashid, Harshvivek, Bhattacharyya, Swapnil, Nair, Shruti, Sekhar, Reshma, Manohar, Siddharth, Hemrajani, Rahul, Bhattacharyya, Pushpak |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
$\textit{Grahak-Nyay:}$ Consumer Grievance Redressal through Large Language Models
by: Ganatra, Shrey, et al.
Published: (2025)
by: Ganatra, Shrey, et al.
Published: (2025)
Nyay-Darpan: Enhancing Decision Making Through Summarization and Case Retrieval for Consumer Law in India
by: Bhattacharyya, Swapnil, et al.
Published: (2025)
by: Bhattacharyya, Swapnil, et al.
Published: (2025)
Timing Matters: Enhancing User Experience through Temporal Prediction in Smart Homes
by: Ganatra, Shrey, et al.
Published: (2024)
by: Ganatra, Shrey, et al.
Published: (2024)
RoundTripOCR: A Data Generation Technique for Enhancing Post-OCR Error Correction in Low-Resource Devanagari Languages
by: Kashid, Harshvivek, et al.
Published: (2024)
by: Kashid, Harshvivek, et al.
Published: (2024)
Evaluating the Role of Large Language Models in Legal Practice in India
by: Hemrajani, Rahul
Published: (2025)
by: Hemrajani, Rahul
Published: (2025)
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
by: Shejole, Kaustubh Shivshankar, et al.
Published: (2025)
by: Shejole, Kaustubh Shivshankar, et al.
Published: (2025)
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation
by: Moon, Palash, et al.
Published: (2024)
by: Moon, Palash, et al.
Published: (2024)
Stereotype Detection as a Catalyst for Enhanced Bias Detection: A Multi-Task Learning Approach
by: Tomar, Aditya, et al.
Published: (2025)
by: Tomar, Aditya, et al.
Published: (2025)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
by: Pandit, Shrey, et al.
Published: (2025)
by: Pandit, Shrey, et al.
Published: (2025)
Enhancing Food-Domain Question Answering with a Multimodal Knowledge Graph: Hybrid QA Generation and Diversity Analysis
by: B, Srihari K, et al.
Published: (2025)
by: B, Srihari K, et al.
Published: (2025)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
by: Urlana, Ashok, et al.
Published: (2025)
by: Urlana, Ashok, et al.
Published: (2025)
Towards Emotion Consistency Analysis of Large Language Models in Emotional Conversational Contexts
by: Oram, Sneha, et al.
Published: (2026)
by: Oram, Sneha, et al.
Published: (2026)
Reconsidering SMT Over NMT for Closely Related Languages: A Case Study of Persian-Hindi Pair
by: Yousofi, Waisullah, et al.
Published: (2024)
by: Yousofi, Waisullah, et al.
Published: (2024)
Facts-and-Feelings: Capturing both Objectivity and Subjectivity in Table-to-Text Generation
by: Dey, Tathagata, et al.
Published: (2024)
by: Dey, Tathagata, et al.
Published: (2024)
P-ReMIS: Pragmatic Reasoning in Mental Health and a Social Implication
by: Oram, Sneha, et al.
Published: (2025)
by: Oram, Sneha, et al.
Published: (2025)
Main Predicate and Their Arguments as Explanation Signals For Intent Classification
by: Pimparkhede, Sameer, et al.
Published: (2025)
by: Pimparkhede, Sameer, et al.
Published: (2025)
Expect the unexpected: Harnessing Sentence Completion for Sarcasm Detection
by: Joshi, Aditya, et al.
Published: (2017)
by: Joshi, Aditya, et al.
Published: (2017)
HalluGraph: Auditable Hallucination Detection for Legal RAG Systems via Knowledge Graph Alignment
by: Noël, Valentin, et al.
Published: (2025)
by: Noël, Valentin, et al.
Published: (2025)
HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs
by: Cherif, Ahmed
Published: (2026)
by: Cherif, Ahmed
Published: (2026)
ETF: An Entity Tracing Framework for Hallucination Detection in Code Summaries
by: Maharaj, Kishan, et al.
Published: (2024)
by: Maharaj, Kishan, et al.
Published: (2024)
ToxVidLM: A Multimodal Framework for Toxicity Detection in Code-Mixed Videos
by: Maity, Krishanu, et al.
Published: (2024)
by: Maity, Krishanu, et al.
Published: (2024)
HalluZig: Hallucination Detection using Zigzag Persistence
by: Samaga, Shreyas N., et al.
Published: (2026)
by: Samaga, Shreyas N., et al.
Published: (2026)
Managers' Sustainability Mindset: The Role of Place Attachment and Absorptive Capacity
by: Asha K. S. Nair, et al.
Published: (2026)
by: Asha K. S. Nair, et al.
Published: (2026)
HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection
by: Emery, Deanna, et al.
Published: (2025)
by: Emery, Deanna, et al.
Published: (2025)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
by: Yeh, Min-Hsuan, et al.
Published: (2025)
by: Yeh, Min-Hsuan, et al.
Published: (2025)
Explain Thyself Bully: Sentiment Aided Cyberbullying Detection with Explanation
by: Maity, Krishanu, et al.
Published: (2024)
by: Maity, Krishanu, et al.
Published: (2024)
HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents
by: Jin, Chao, et al.
Published: (2026)
by: Jin, Chao, et al.
Published: (2026)
The HalluRAG Dataset: Detecting Closed-Domain Hallucinations in RAG Applications Using an LLM's Internal States
by: Ridder, Fabian, et al.
Published: (2024)
by: Ridder, Fabian, et al.
Published: (2024)
Pretraining Language Models Using Translationese
by: Doshi, Meet, et al.
Published: (2024)
by: Doshi, Meet, et al.
Published: (2024)
Who Will Top the Charts? Multimodal Music Popularity Prediction via Adaptive Fusion of Modality Experts and Temporal Engagement Modeling
by: Choudhary, Yash, et al.
Published: (2025)
by: Choudhary, Yash, et al.
Published: (2025)
Evaluating Extremely Low-Resource Machine Translation: A Comparative Study of ChrF++ and BLEU Metrics
by: Kumar, Sanjeev, et al.
Published: (2026)
by: Kumar, Sanjeev, et al.
Published: (2026)
Together We Can: Multilingual Automatic Post-Editing for Low-Resource Languages
by: Deoghare, Sourabh, et al.
Published: (2024)
by: Deoghare, Sourabh, et al.
Published: (2024)
Ta-G-T: Subjectivity Capture in Table to Text Generation via RDF Graphs
by: Upasham, Ronak, et al.
Published: (2025)
by: Upasham, Ronak, et al.
Published: (2025)
Giving the Old a Fresh Spin: Quality Estimation-Assisted Constrained Decoding for Automatic Post-Editing
by: Deoghare, Sourabh, et al.
Published: (2025)
by: Deoghare, Sourabh, et al.
Published: (2025)
Lyrics Matter: Exploiting the Power of Learnt Representations for Music Popularity Prediction
by: Choudhary, Yash, et al.
Published: (2025)
by: Choudhary, Yash, et al.
Published: (2025)
Recon, Answer, Verify: Agents in Search of Truth
by: Shukla, Satyam, et al.
Published: (2025)
by: Shukla, Satyam, et al.
Published: (2025)
CoSTA: Code-Switched Speech Translation using Aligned Speech-Text Interleaving
by: Shankar, Bhavani, et al.
Published: (2024)
by: Shankar, Bhavani, et al.
Published: (2024)
Transformers are Expressive, But Are They Expressive Enough for Regression?
by: Nath, Swaroop, et al.
Published: (2024)
by: Nath, Swaroop, et al.
Published: (2024)
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
by: Ghosh, Poulami, et al.
Published: (2024)
by: Ghosh, Poulami, et al.
Published: (2024)
HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs
by: Dasgupta, Sharanya, et al.
Published: (2025)
by: Dasgupta, Sharanya, et al.
Published: (2025)
Similar Items
-
$\textit{Grahak-Nyay:}$ Consumer Grievance Redressal through Large Language Models
by: Ganatra, Shrey, et al.
Published: (2025) -
Nyay-Darpan: Enhancing Decision Making Through Summarization and Case Retrieval for Consumer Law in India
by: Bhattacharyya, Swapnil, et al.
Published: (2025) -
Timing Matters: Enhancing User Experience through Temporal Prediction in Smart Homes
by: Ganatra, Shrey, et al.
Published: (2024) -
RoundTripOCR: A Data Generation Technique for Enhancing Post-OCR Error Correction in Low-Resource Devanagari Languages
by: Kashid, Harshvivek, et al.
Published: (2024) -
Evaluating the Role of Large Language Models in Legal Practice in India
by: Hemrajani, Rahul
Published: (2025)