Calibrated Trust in Dealing with LLM Hallucinations: A Qualitative Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ryser, Adrian, Allwein, Florian, Schlippe, Tim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RubiSCoT: A Framework for AI-Supported Academic Assessment
von: Fröhlich, Thorsten, et al.
Veröffentlicht: (2025)
von: Fröhlich, Thorsten, et al.
Veröffentlicht: (2025)
A Cross-Cultural Assessment of Human Ability to Detect LLM-Generated Fake News about South Africa
von: Schlippe, Tim, et al.
Veröffentlicht: (2025)
von: Schlippe, Tim, et al.
Veröffentlicht: (2025)
Exploring ChatGPT's Empathic Abilities
von: Schaaff, Kristina, et al.
Veröffentlicht: (2023)
von: Schaaff, Kristina, et al.
Veröffentlicht: (2023)
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
von: Wagner, Julia, et al.
Veröffentlicht: (2026)
von: Wagner, Julia, et al.
Veröffentlicht: (2026)
Classification of Human- and AI-Generated Texts for English, French, German, and Spanish
von: Schaaff, Kristina, et al.
Veröffentlicht: (2023)
von: Schaaff, Kristina, et al.
Veröffentlicht: (2023)
Language-Independent Sentiment Labelling with Distant Supervision: A Case Study for English, Sepedi and Setswana
von: Mabokela, Koena Ronny, et al.
Veröffentlicht: (2025)
von: Mabokela, Koena Ronny, et al.
Veröffentlicht: (2025)
Large Language Models for Sentiment Analysis to Detect Social Challenges: A Use Case with South African Languages
von: Mabokela, Koena Ronny, et al.
Veröffentlicht: (2025)
von: Mabokela, Koena Ronny, et al.
Veröffentlicht: (2025)
Evaluating Retrieval-Augmented Generation Variants for Natural Language-Based SQL and API Call Generation
von: Marketsmüller, Michael, et al.
Veröffentlicht: (2026)
von: Marketsmüller, Michael, et al.
Veröffentlicht: (2026)
Deep Learning-Based Anomaly Detection in Spacecraft Telemetry on Edge Devices
von: Goetze, Christopher, et al.
Veröffentlicht: (2026)
von: Goetze, Christopher, et al.
Veröffentlicht: (2026)
Assessing Consciousness-Related Behaviors in Large Language Models Using the Maze Test
von: Pimenta, Rui A., et al.
Veröffentlicht: (2025)
von: Pimenta, Rui A., et al.
Veröffentlicht: (2025)
Mitigating LLM Hallucination via Behaviorally Calibrated Reinforcement Learning
von: Wu, Jiayun, et al.
Veröffentlicht: (2025)
von: Wu, Jiayun, et al.
Veröffentlicht: (2025)
Benchmarking NLP-supported Language Sample Analysis for Swiss Children's Speech
von: Ryser, Anja, et al.
Veröffentlicht: (2025)
von: Ryser, Anja, et al.
Veröffentlicht: (2025)
Exploring Trust Calibration in XAI - The Impact of Exposing Model Limitations to Lay Users
von: Ventura, Alfio, et al.
Veröffentlicht: (2026)
von: Ventura, Alfio, et al.
Veröffentlicht: (2026)
Calibrated Language Models Must Hallucinate
von: Kalai, Adam Tauman, et al.
Veröffentlicht: (2023)
von: Kalai, Adam Tauman, et al.
Veröffentlicht: (2023)
Citations and Trust in LLM Generated Responses
von: Ding, Yifan, et al.
Veröffentlicht: (2025)
von: Ding, Yifan, et al.
Veröffentlicht: (2025)
Uncertainty Awareness and Trust in Explainable AI- On Trust Calibration using Local and Global Explanations
von: Newen, Carina, et al.
Veröffentlicht: (2025)
von: Newen, Carina, et al.
Veröffentlicht: (2025)
Epistemic Filtering and Collective Hallucination: A Jury Theorem for Confidence-Calibrated Agents
von: Karge, Jonas
Veröffentlicht: (2026)
von: Karge, Jonas
Veröffentlicht: (2026)
Mitigating LLM Hallucinations with Knowledge Graphs: A Case Study
von: Li, Harry, et al.
Veröffentlicht: (2025)
von: Li, Harry, et al.
Veröffentlicht: (2025)
Hallucination in LLM-Based Code Generation: An Automotive Case Study
von: Pavel, Marc, et al.
Veröffentlicht: (2025)
von: Pavel, Marc, et al.
Veröffentlicht: (2025)
To Trust or Not to Trust: On Calibration in ML-based Resource Allocation for Wireless Networks
von: Raina, Rashika, et al.
Veröffentlicht: (2025)
von: Raina, Rashika, et al.
Veröffentlicht: (2025)
Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations
von: Luo, Jinyuan, et al.
Veröffentlicht: (2025)
von: Luo, Jinyuan, et al.
Veröffentlicht: (2025)
Evaluating Human Trust in LLM-Based Planners: A Preliminary Study
von: Chen, Shenghui, et al.
Veröffentlicht: (2025)
von: Chen, Shenghui, et al.
Veröffentlicht: (2025)
Cross-Modal Attention Calibration for LVLM Hallucination Mitigation
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
Two Is Better Than One: Aligned Representation Pairs for Anomaly Detection
von: Ryser, Alain, et al.
Veröffentlicht: (2024)
von: Ryser, Alain, et al.
Veröffentlicht: (2024)
Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations
von: Cherukuri, Kalyan, et al.
Veröffentlicht: (2026)
von: Cherukuri, Kalyan, et al.
Veröffentlicht: (2026)
HalluLens: LLM Hallucination Benchmark
von: Bang, Yejin, et al.
Veröffentlicht: (2025)
von: Bang, Yejin, et al.
Veröffentlicht: (2025)
Can You Trust an LLM with Your Life-Changing Decision? An Investigation into AI High-Stakes Responses
von: Cahyono, Joshua Adrian, et al.
Veröffentlicht: (2025)
von: Cahyono, Joshua Adrian, et al.
Veröffentlicht: (2025)
MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
Dynamic Trust Calibration Using Contextual Bandits
von: Henrique, Bruno M., et al.
Veröffentlicht: (2025)
von: Henrique, Bruno M., et al.
Veröffentlicht: (2025)
TERMS-Bench: Diagnosing LLM Negotiation Agents Beyond Deal Rate
von: Zhang, Erica, et al.
Veröffentlicht: (2026)
von: Zhang, Erica, et al.
Veröffentlicht: (2026)
HALT-RAG: A Task-Adaptable Framework for Hallucination Detection with Calibrated NLI Ensembles and Abstention
von: Goswami, Saumya, et al.
Veröffentlicht: (2025)
von: Goswami, Saumya, et al.
Veröffentlicht: (2025)
Dealing with Inconsistency for Reasoning over Knowledge Graphs: A Survey
von: Nentidis, Anastasios, et al.
Veröffentlicht: (2025)
von: Nentidis, Anastasios, et al.
Veröffentlicht: (2025)
Banishing LLM Hallucinations Requires Rethinking Generalization
von: Li, Johnny, et al.
Veröffentlicht: (2024)
von: Li, Johnny, et al.
Veröffentlicht: (2024)
Can We Trust LLM Detectors?
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
von: Islam, Saad Obaid ul, et al.
Veröffentlicht: (2025)
LLM-based Agents Suffer from Hallucinations: A Survey of Taxonomy, Methods, and Directions
von: Lin, Xixun, et al.
Veröffentlicht: (2025)
von: Lin, Xixun, et al.
Veröffentlicht: (2025)
Algorithmically Establishing Trust in Evaluators
von: de Wynter, Adrian
Veröffentlicht: (2025)
von: de Wynter, Adrian
Veröffentlicht: (2025)
Learn to Code Sustainably: An Empirical Study on LLM-based Green Code Generation
von: Vartziotis, Tina, et al.
Veröffentlicht: (2024)
von: Vartziotis, Tina, et al.
Veröffentlicht: (2024)
LLM-REVal: Can We Trust LLM Reviewers Yet?
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs
von: Huang, Xiaoyi, et al.
Veröffentlicht: (2026)
von: Huang, Xiaoyi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RubiSCoT: A Framework for AI-Supported Academic Assessment
von: Fröhlich, Thorsten, et al.
Veröffentlicht: (2025) -
A Cross-Cultural Assessment of Human Ability to Detect LLM-Generated Fake News about South Africa
von: Schlippe, Tim, et al.
Veröffentlicht: (2025) -
Exploring ChatGPT's Empathic Abilities
von: Schaaff, Kristina, et al.
Veröffentlicht: (2023) -
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
von: Wagner, Julia, et al.
Veröffentlicht: (2026) -
Classification of Human- and AI-Generated Texts for English, French, German, and Spanish
von: Schaaff, Kristina, et al.
Veröffentlicht: (2023)