MMD-Flagger: Leveraging Maximum Mean Discrepancy to Detect Hallucinations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mitsuzawa, Kensuke, Garreau, Damien |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Word Sense Detection Leveraging Maximum Mean Discrepancy
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025)
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025)
Variable Selection in Maximum Mean Discrepancy for Interpretable Distribution Comparison
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2023)
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2023)
Detecting Machine-Generated Texts by Multi-Population Aware Optimization for Maximum Mean Discrepancy
von: Zhang, Shuhai, et al.
Veröffentlicht: (2024)
von: Zhang, Shuhai, et al.
Veröffentlicht: (2024)
Comparing Feature Importance and Rule Extraction for Interpretability on Text Data
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2022)
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2022)
Faithful and Robust Local Interpretability for Textual Predictions
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)
Towards Understanding Steering Strength
von: Taimeskhanov, Magamed, et al.
Veröffentlicht: (2026)
von: Taimeskhanov, Magamed, et al.
Veröffentlicht: (2026)
Attention Meets Post-hoc Interpretability: A Mathematical Perspective
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2024)
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2024)
Understanding Post-hoc Explainers: The Case of Anchors
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)
A Sea of Words: An In-Depth Analysis of Anchors for Text Data
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2022)
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2022)
Leveraging Graph Structures to Detect Hallucinations in Large Language Models
von: Nonkes, Noa, et al.
Veröffentlicht: (2024)
von: Nonkes, Noa, et al.
Veröffentlicht: (2024)
MMD-OPT : Maximum Mean Discrepancy Based Sample Efficient Collision Risk Minimization for Autonomous Driving
von: Sharma, Basant, et al.
Veröffentlicht: (2024)
von: Sharma, Basant, et al.
Veröffentlicht: (2024)
Leveraging NTPs for Efficient Hallucination Detection in VLMs
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
Time-MMD: Multi-Domain Multimodal Dataset for Time Series Analysis
von: Liu, Haoxin, et al.
Veröffentlicht: (2024)
von: Liu, Haoxin, et al.
Veröffentlicht: (2024)
MALTO at SemEval-2024 Task 6: Leveraging Synthetic Data for LLM Hallucination Detection
von: Borra, Federico, et al.
Veröffentlicht: (2024)
von: Borra, Federico, et al.
Veröffentlicht: (2024)
Sample-Efficient Human Evaluation of Large Language Models via Maximum Discrepancy Competition
von: Feng, Kehua, et al.
Veröffentlicht: (2024)
von: Feng, Kehua, et al.
Veröffentlicht: (2024)
Detecting AI Hallucinations in Finance: An Information-Theoretic Method Cuts Hallucination Rate by 92%
von: Singha, Mainak
Veröffentlicht: (2025)
von: Singha, Mainak
Veröffentlicht: (2025)
On the Optimization Landscape of Maximum Mean Discrepancy
von: Alon, Itai, et al.
Veröffentlicht: (2021)
von: Alon, Itai, et al.
Veröffentlicht: (2021)
HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
von: Du, Xuefeng, et al.
Veröffentlicht: (2024)
von: Du, Xuefeng, et al.
Veröffentlicht: (2024)
Temporal Graph Network: Hallucination Detection in Multi-Turn Conversation
von: Rathore, Vidhi, et al.
Veröffentlicht: (2026)
von: Rathore, Vidhi, et al.
Veröffentlicht: (2026)
A Deterministic Sampling Method via Maximum Mean Discrepancy Flow with Adaptive Kernel
von: Chen, Yindong, et al.
Veröffentlicht: (2021)
von: Chen, Yindong, et al.
Veröffentlicht: (2021)
Leveraging Language Models to Detect Greenwashing
von: Vinella, Avalon, et al.
Veröffentlicht: (2023)
von: Vinella, Avalon, et al.
Veröffentlicht: (2023)
Beyond ROUGE: N-Gram Subspace Features for LLM Hallucination Detection
von: Li, Jerry, et al.
Veröffentlicht: (2025)
von: Li, Jerry, et al.
Veröffentlicht: (2025)
RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration
von: Ridder, Fabian, et al.
Veröffentlicht: (2026)
von: Ridder, Fabian, et al.
Veröffentlicht: (2026)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
von: Yehuda, Yakir, et al.
Veröffentlicht: (2024)
von: Yehuda, Yakir, et al.
Veröffentlicht: (2024)
Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026)
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026)
Resolving Discrepancies in Compute-Optimal Scaling of Language Models
von: Porian, Tomer, et al.
Veröffentlicht: (2024)
von: Porian, Tomer, et al.
Veröffentlicht: (2024)
M3D: Dataset Condensation by Minimizing Maximum Mean Discrepancy
von: Zhang, Hansong, et al.
Veröffentlicht: (2023)
von: Zhang, Hansong, et al.
Veröffentlicht: (2023)
Learning to Reason for Hallucination Span Detection
von: Su, Hsuan, et al.
Veröffentlicht: (2025)
von: Su, Hsuan, et al.
Veröffentlicht: (2025)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Cost-Effective Hallucination Detection for LLMs
von: Valentin, Simon, et al.
Veröffentlicht: (2024)
von: Valentin, Simon, et al.
Veröffentlicht: (2024)
Hallucination Detection and Mitigation with Diffusion in Multi-Variate Time-Series Foundation Models
von: Wichitwechkarn, Vijja, et al.
Veröffentlicht: (2025)
von: Wichitwechkarn, Vijja, et al.
Veröffentlicht: (2025)
Detection Without Correction: A Robust Asymmetry in Activation-Based Hallucination Probing
von: Roy, Dip, et al.
Veröffentlicht: (2026)
von: Roy, Dip, et al.
Veröffentlicht: (2026)
The Risks of Recourse in Binary Classification
von: Fokkema, Hidde, et al.
Veröffentlicht: (2023)
von: Fokkema, Hidde, et al.
Veröffentlicht: (2023)
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
A Practical Introduction to Kernel Discrepancies: MMD, HSIC & KSD
von: Schrab, Antonin
Veröffentlicht: (2025)
von: Schrab, Antonin
Veröffentlicht: (2025)
TraceDet: Hallucination Detection from the Decoding Trace of Diffusion Large Language Models
von: Chang, Shenxu, et al.
Veröffentlicht: (2025)
von: Chang, Shenxu, et al.
Veröffentlicht: (2025)
Detecting Token-Level Hallucinations Using Variance Signals: A Reference-Free Approach
von: Kumar, Keshav
Veröffentlicht: (2025)
von: Kumar, Keshav
Veröffentlicht: (2025)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
CAM-Based Methods Can See through Walls
von: Taimeskhanov, Magamed, et al.
Veröffentlicht: (2024)
von: Taimeskhanov, Magamed, et al.
Veröffentlicht: (2024)
Margin Discrepancy-based Adversarial Training for Multi-Domain Text Classification
von: Wu, Yuan
Veröffentlicht: (2024)
von: Wu, Yuan
Veröffentlicht: (2024)
Ähnliche Einträge
-
Word Sense Detection Leveraging Maximum Mean Discrepancy
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025) -
Variable Selection in Maximum Mean Discrepancy for Interpretable Distribution Comparison
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2023) -
Detecting Machine-Generated Texts by Multi-Population Aware Optimization for Maximum Mean Discrepancy
von: Zhang, Shuhai, et al.
Veröffentlicht: (2024) -
Comparing Feature Importance and Rule Extraction for Interpretability on Text Data
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2022) -
Faithful and Robust Local Interpretability for Textual Predictions
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)