Comparing Hallucination Detection Metrics for Multilingual Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Kang, Haoqiang, Blevins, Terra, Zettlemoyer, Luke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2024)
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2024)
Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
di: Kang, Haoqiang, et al.
Pubblicazione: (2023)
di: Kang, Haoqiang, et al.
Pubblicazione: (2023)
Breaking the Curse of Multilinguality with Cross-lingual Expert Language Models
di: Blevins, Terra, et al.
Pubblicazione: (2024)
di: Blevins, Terra, et al.
Pubblicazione: (2024)
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
di: Ahuja, Sanchit, et al.
Pubblicazione: (2026)
di: Ahuja, Sanchit, et al.
Pubblicazione: (2026)
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
di: Gonen, Hila, et al.
Pubblicazione: (2024)
di: Gonen, Hila, et al.
Pubblicazione: (2024)
Demystifying Prompts in Language Models via Perplexity Estimation
di: Gonen, Hila, et al.
Pubblicazione: (2022)
di: Gonen, Hila, et al.
Pubblicazione: (2022)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
di: Kulkarni, Atharva, et al.
Pubblicazione: (2025)
di: Kulkarni, Atharva, et al.
Pubblicazione: (2025)
Accommodation Goes Both Ways: Studying Linguistic Convergence Between Humans and Language Models
di: Blevins, Terra
Pubblicazione: (2026)
di: Blevins, Terra
Pubblicazione: (2026)
(Mis)Fitting: A Survey of Scaling Laws
di: Li, Margaret, et al.
Pubblicazione: (2025)
di: Li, Margaret, et al.
Pubblicazione: (2025)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
di: Islam, Saad Obaid ul, et al.
Pubblicazione: (2025)
di: Islam, Saad Obaid ul, et al.
Pubblicazione: (2025)
Enhancing Hallucination Detection via Future Context
di: Lee, Joosung, et al.
Pubblicazione: (2025)
di: Lee, Joosung, et al.
Pubblicazione: (2025)
Hallucination Detection and Hallucination Mitigation: An Investigation
di: Luo, Junliang, et al.
Pubblicazione: (2024)
di: Luo, Junliang, et al.
Pubblicazione: (2024)
keepitsimple at SemEval-2025 Task 3: LLM-Uncertainty based Approach for Multilingual Hallucination Span Detection
di: Vemula, Saketh Reddy, et al.
Pubblicazione: (2025)
di: Vemula, Saketh Reddy, et al.
Pubblicazione: (2025)
How to Detect and Defeat Molecular Mirage: A Metric-Driven Benchmark for Hallucination in LLM-based Molecular Comprehension
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples
di: Yu, Fangxu, et al.
Pubblicazione: (2024)
di: Yu, Fangxu, et al.
Pubblicazione: (2024)
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs
di: Bui, Anh Thu Maria, et al.
Pubblicazione: (2024)
di: Bui, Anh Thu Maria, et al.
Pubblicazione: (2024)
Infini-gram: Scaling Unbounded n-gram Language Models to a Trillion Tokens
di: Liu, Jiacheng, et al.
Pubblicazione: (2024)
di: Liu, Jiacheng, et al.
Pubblicazione: (2024)
Memory Layers at Scale
di: Berges, Vincent-Pierre, et al.
Pubblicazione: (2024)
di: Berges, Vincent-Pierre, et al.
Pubblicazione: (2024)
Beyond LLM-as-a-Judge: Deterministic Metrics for Multilingual Generative Text Evaluation
di: Alam, Firoj, et al.
Pubblicazione: (2026)
di: Alam, Firoj, et al.
Pubblicazione: (2026)
Enhancing Knowledge Graph Construction: Evaluating with Emphasis on Hallucination, Omission, and Graph Similarity Metrics
di: Ghanem, Hussam, et al.
Pubblicazione: (2025)
di: Ghanem, Hussam, et al.
Pubblicazione: (2025)
Hallucination Detection with Small Language Models
di: Cheung, Ming
Pubblicazione: (2025)
di: Cheung, Ming
Pubblicazione: (2025)
Hallucination Detection with the Internal Layers of LLMs
di: Preiß, Martin
Pubblicazione: (2025)
di: Preiß, Martin
Pubblicazione: (2025)
Towards Long Context Hallucination Detection
di: Liu, Siyi, et al.
Pubblicazione: (2025)
di: Liu, Siyi, et al.
Pubblicazione: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
Enhancing Hallucination Detection through Perturbation-Based Synthetic Data Generation in System Responses
di: Zhang, Dongxu, et al.
Pubblicazione: (2024)
di: Zhang, Dongxu, et al.
Pubblicazione: (2024)
LettuceDetect: A Hallucination Detection Framework for RAG Applications
di: Kovács, Ádám, et al.
Pubblicazione: (2025)
di: Kovács, Ádám, et al.
Pubblicazione: (2025)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
di: Macko, Dominik, et al.
Pubblicazione: (2023)
di: Macko, Dominik, et al.
Pubblicazione: (2023)
On Early Detection of Hallucinations in Factual Question Answering
di: Snyder, Ben, et al.
Pubblicazione: (2023)
di: Snyder, Ben, et al.
Pubblicazione: (2023)
Unsupervised Hallucination Detection by Inspecting Reasoning Processes
di: Srey, Ponhvoan, et al.
Pubblicazione: (2025)
di: Srey, Ponhvoan, et al.
Pubblicazione: (2025)
FaithLens: Detecting and Explaining Faithfulness Hallucination
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
Sanity Checks for Long-Form Hallucination Detection
di: Zollicoffer, Geigh, et al.
Pubblicazione: (2026)
di: Zollicoffer, Geigh, et al.
Pubblicazione: (2026)
Detecting Pretraining Data from Large Language Models
di: Shi, Weijia, et al.
Pubblicazione: (2023)
di: Shi, Weijia, et al.
Pubblicazione: (2023)
Prompt-Response Semantic Divergence Metrics for Faithfulness Hallucination and Misalignment Detection in Large Language Models
di: Halperin, Igor
Pubblicazione: (2025)
di: Halperin, Igor
Pubblicazione: (2025)
Learning When to Translate for Multilingual Reasoning
di: Kang, Deokhyung, et al.
Pubblicazione: (2026)
di: Kang, Deokhyung, et al.
Pubblicazione: (2026)
ReXTrust: A Model for Fine-Grained Hallucination Detection in AI-Generated Radiology Reports
di: Hardy, Romain, et al.
Pubblicazione: (2024)
di: Hardy, Romain, et al.
Pubblicazione: (2024)
Continual Learning via Sparse Memory Finetuning
di: Lin, Jessy, et al.
Pubblicazione: (2025)
di: Lin, Jessy, et al.
Pubblicazione: (2025)
From Out-of-Distribution Detection to Hallucination Detection: A Geometric View
di: Liu, Litian, et al.
Pubblicazione: (2026)
di: Liu, Litian, et al.
Pubblicazione: (2026)
MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media Texts
di: Macko, Dominik, et al.
Pubblicazione: (2024)
di: Macko, Dominik, et al.
Pubblicazione: (2024)
CEAID: Benchmark of Multilingual Machine-Generated Text Detection Methods for Central European Languages
di: Macko, Dominik, et al.
Pubblicazione: (2025)
di: Macko, Dominik, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
di: Limisiewicz, Tomasz, et al.
Pubblicazione: (2024) -
Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
di: Kang, Haoqiang, et al.
Pubblicazione: (2023) -
Breaking the Curse of Multilinguality with Cross-lingual Expert Language Models
di: Blevins, Terra, et al.
Pubblicazione: (2024) -
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
di: Ahuja, Sanchit, et al.
Pubblicazione: (2026) -
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
di: Gonen, Hila, et al.
Pubblicazione: (2024)