Seeing Through the Fog: A Cost-Effectiveness Analysis of Hallucination Detection Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Thomas, Alexander, Rosen, Seth, Vettrivel, Vishnu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ELECTRA and GPT-4o: Cost-Effective Partners for Sentiment Analysis
di: Beno, James P.
Pubblicazione: (2024)
di: Beno, James P.
Pubblicazione: (2024)
MedHal: An Evaluation Dataset for Medical Hallucination Detection
di: Mehenni, Gaya, et al.
Pubblicazione: (2025)
di: Mehenni, Gaya, et al.
Pubblicazione: (2025)
German also Hallucinates! Inconsistency Detection in News Summaries with the Absinth Dataset
di: Mascarell, Laura, et al.
Pubblicazione: (2024)
di: Mascarell, Laura, et al.
Pubblicazione: (2024)
Detecting Hallucinations in Graph Retrieval-Augmented Generation via Attention Patterns and Semantic Alignment
di: Li, Shanghao, et al.
Pubblicazione: (2025)
di: Li, Shanghao, et al.
Pubblicazione: (2025)
Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Listen to the Layers: Mitigating Hallucinations with Inter-Layer Disagreement
di: Subbalakshmi, Koduvayur, et al.
Pubblicazione: (2026)
di: Subbalakshmi, Koduvayur, et al.
Pubblicazione: (2026)
On the Effectiveness of LLM-Specific Fine-Tuning for Detecting AI-Generated Text
di: Gromadzki, Michał, et al.
Pubblicazione: (2026)
di: Gromadzki, Michał, et al.
Pubblicazione: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
Hallucination or Creativity: How to Evaluate AI-Generated Scientific Stories?
di: Argese, Alex, et al.
Pubblicazione: (2026)
di: Argese, Alex, et al.
Pubblicazione: (2026)
TRACE: Trajectory Correction from Cross-layer Evidence for Hallucination Reduction
di: Ranade, Tej Sanibh
Pubblicazione: (2026)
di: Ranade, Tej Sanibh
Pubblicazione: (2026)
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
Fine-Tuned Large Language Models for Logical Translation: Reducing Hallucinations with Lang2Logic
di: Pan, Muyu, et al.
Pubblicazione: (2025)
di: Pan, Muyu, et al.
Pubblicazione: (2025)
Low-Resource Court Judgment Summarization for Common Law Systems
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
Detecting Hallucinations in Large Language Model Generation: A Token Probability Approach
di: Quevedo, Ernesto, et al.
Pubblicazione: (2024)
di: Quevedo, Ernesto, et al.
Pubblicazione: (2024)
Enhancing Hate Speech Detection on Social Media: A Comparative Analysis of Machine Learning Models and Text Transformation Approaches
di: Mishra, Saurabh, et al.
Pubblicazione: (2026)
di: Mishra, Saurabh, et al.
Pubblicazione: (2026)
Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data
di: de Campos, Andresa Rodrigues, et al.
Pubblicazione: (2026)
di: de Campos, Andresa Rodrigues, et al.
Pubblicazione: (2026)
NinjaLLM: Fast, Scalable and Cost-effective RAG using Amazon SageMaker and AWS Trainium and Inferentia2
di: Xue, Tengfei, et al.
Pubblicazione: (2024)
di: Xue, Tengfei, et al.
Pubblicazione: (2024)
Reducing Hallucinations in Summarization via Reinforcement Learning with Entity Hallucination Index
di: Katwe, Praveenkumar, et al.
Pubblicazione: (2025)
di: Katwe, Praveenkumar, et al.
Pubblicazione: (2025)
Context Is What You Need: The Maximum Effective Context Window for Real World Limits of LLMs
di: Paulsen, Norman
Pubblicazione: (2025)
di: Paulsen, Norman
Pubblicazione: (2025)
PARAPHRASUS : A Comprehensive Benchmark for Evaluating Paraphrase Detection Models
di: Michail, Andrianos, et al.
Pubblicazione: (2024)
di: Michail, Andrianos, et al.
Pubblicazione: (2024)
CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification
di: Ye, Severin, et al.
Pubblicazione: (2026)
di: Ye, Severin, et al.
Pubblicazione: (2026)
Blessing or curse? A survey on the Impact of Generative AI on Fake News
di: Loth, Alexander, et al.
Pubblicazione: (2024)
di: Loth, Alexander, et al.
Pubblicazione: (2024)
Detecting and Steering LLMs' Empathy in Action
di: Cadile, Juan P.
Pubblicazione: (2025)
di: Cadile, Juan P.
Pubblicazione: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
Detecting AI-Generated Texts in Cross-Domains
di: Zhou, You, et al.
Pubblicazione: (2024)
di: Zhou, You, et al.
Pubblicazione: (2024)
Identifying Bias in Machine-generated Text Detection
di: Stowe, Kevin, et al.
Pubblicazione: (2025)
di: Stowe, Kevin, et al.
Pubblicazione: (2025)
Efficient Toxicity Detection in Gaming Chats: A Comparative Study of Embeddings, Fine-Tuned Transformers and LLMs
di: Tereshchenko, Yehor, et al.
Pubblicazione: (2025)
di: Tereshchenko, Yehor, et al.
Pubblicazione: (2025)
Bounding Hallucinations: Information-Theoretic Guarantees for RAG Systems via Merlin-Arthur Protocols
di: Deiseroth, Björn, et al.
Pubblicazione: (2025)
di: Deiseroth, Björn, et al.
Pubblicazione: (2025)
Spotlights and Blindspots: Evaluating Machine-Generated Text Detection
di: Stowe, Kevin, et al.
Pubblicazione: (2026)
di: Stowe, Kevin, et al.
Pubblicazione: (2026)
Detecting Data Contamination in LLMs via In-Context Learning
di: Zawalski, Michał, et al.
Pubblicazione: (2025)
di: Zawalski, Michał, et al.
Pubblicazione: (2025)
How Instruction-Tuning Imparts Length Control: A Cross-Lingual Mechanistic Analysis
di: Rocchetti, Elisabetta, et al.
Pubblicazione: (2025)
di: Rocchetti, Elisabetta, et al.
Pubblicazione: (2025)
Internal Reasoning vs. External Control: A Thermodynamic Analysis of Sycophancy in Large Language Models
di: Chang, Edward Y.
Pubblicazione: (2025)
di: Chang, Edward Y.
Pubblicazione: (2025)
Exploring RWKV for Sentence Embeddings: Layer-wise Analysis and Baseline Comparison for Semantic Similarity
di: Pan, Xinghan
Pubblicazione: (2025)
di: Pan, Xinghan
Pubblicazione: (2025)
Historical Ink: Semantic Shift Detection for 19th Century Spanish
di: Montes, Tony, et al.
Pubblicazione: (2024)
di: Montes, Tony, et al.
Pubblicazione: (2024)
Beyond the Mean: Within-Model Reliable Change Detection for LLM Evaluation
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
SADAS: A Dialogue Assistant System Towards Remediating Norm Violations in Bilingual Socio-Cultural Conversations
di: Hua, Yuncheng, et al.
Pubblicazione: (2024)
di: Hua, Yuncheng, et al.
Pubblicazione: (2024)
Paying Attention to Deflections: Mining Pragmatic Nuances for Whataboutism Detection in Online Discourse
di: Phi, Khiem, et al.
Pubblicazione: (2024)
di: Phi, Khiem, et al.
Pubblicazione: (2024)
AEyeDE: An Attention-Based Attribution Framework for AI-Generated Text Detection
di: Nourbakhsh, Aria, et al.
Pubblicazione: (2026)
di: Nourbakhsh, Aria, et al.
Pubblicazione: (2026)
Detecting Subtle Differences between Human and Model Languages Using Spectrum of Relative Likelihood
di: Xu, Yang, et al.
Pubblicazione: (2024)
di: Xu, Yang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ELECTRA and GPT-4o: Cost-Effective Partners for Sentiment Analysis
di: Beno, James P.
Pubblicazione: (2024) -
MedHal: An Evaluation Dataset for Medical Hallucination Detection
di: Mehenni, Gaya, et al.
Pubblicazione: (2025) -
German also Hallucinates! Inconsistency Detection in News Summaries with the Absinth Dataset
di: Mascarell, Laura, et al.
Pubblicazione: (2024) -
Detecting Hallucinations in Graph Retrieval-Augmented Generation via Attention Patterns and Semantic Alignment
di: Li, Shanghao, et al.
Pubblicazione: (2025) -
Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)