Calibrated Trust in Dealing with LLM Hallucinations: A Qualitative Study
Fuente:
arXiv
Salvato in:
| Autori principali: | Ryser, Adrian, Allwein, Florian, Schlippe, Tim |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RubiSCoT: A Framework for AI-Supported Academic Assessment
di: Fröhlich, Thorsten, et al.
Pubblicazione: (2025)
di: Fröhlich, Thorsten, et al.
Pubblicazione: (2025)
A Cross-Cultural Assessment of Human Ability to Detect LLM-Generated Fake News about South Africa
di: Schlippe, Tim, et al.
Pubblicazione: (2025)
di: Schlippe, Tim, et al.
Pubblicazione: (2025)
Exploring ChatGPT's Empathic Abilities
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
di: Wagner, Julia, et al.
Pubblicazione: (2026)
di: Wagner, Julia, et al.
Pubblicazione: (2026)
Classification of Human- and AI-Generated Texts for English, French, German, and Spanish
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)
Language-Independent Sentiment Labelling with Distant Supervision: A Case Study for English, Sepedi and Setswana
di: Mabokela, Koena Ronny, et al.
Pubblicazione: (2025)
di: Mabokela, Koena Ronny, et al.
Pubblicazione: (2025)
Large Language Models for Sentiment Analysis to Detect Social Challenges: A Use Case with South African Languages
di: Mabokela, Koena Ronny, et al.
Pubblicazione: (2025)
di: Mabokela, Koena Ronny, et al.
Pubblicazione: (2025)
Evaluating Retrieval-Augmented Generation Variants for Natural Language-Based SQL and API Call Generation
di: Marketsmüller, Michael, et al.
Pubblicazione: (2026)
di: Marketsmüller, Michael, et al.
Pubblicazione: (2026)
Deep Learning-Based Anomaly Detection in Spacecraft Telemetry on Edge Devices
di: Goetze, Christopher, et al.
Pubblicazione: (2026)
di: Goetze, Christopher, et al.
Pubblicazione: (2026)
Assessing Consciousness-Related Behaviors in Large Language Models Using the Maze Test
di: Pimenta, Rui A., et al.
Pubblicazione: (2025)
di: Pimenta, Rui A., et al.
Pubblicazione: (2025)
Mitigating LLM Hallucination via Behaviorally Calibrated Reinforcement Learning
di: Wu, Jiayun, et al.
Pubblicazione: (2025)
di: Wu, Jiayun, et al.
Pubblicazione: (2025)
Benchmarking NLP-supported Language Sample Analysis for Swiss Children's Speech
di: Ryser, Anja, et al.
Pubblicazione: (2025)
di: Ryser, Anja, et al.
Pubblicazione: (2025)
Exploring Trust Calibration in XAI - The Impact of Exposing Model Limitations to Lay Users
di: Ventura, Alfio, et al.
Pubblicazione: (2026)
di: Ventura, Alfio, et al.
Pubblicazione: (2026)
Calibrated Language Models Must Hallucinate
di: Kalai, Adam Tauman, et al.
Pubblicazione: (2023)
di: Kalai, Adam Tauman, et al.
Pubblicazione: (2023)
Citations and Trust in LLM Generated Responses
di: Ding, Yifan, et al.
Pubblicazione: (2025)
di: Ding, Yifan, et al.
Pubblicazione: (2025)
Uncertainty Awareness and Trust in Explainable AI- On Trust Calibration using Local and Global Explanations
di: Newen, Carina, et al.
Pubblicazione: (2025)
di: Newen, Carina, et al.
Pubblicazione: (2025)
Epistemic Filtering and Collective Hallucination: A Jury Theorem for Confidence-Calibrated Agents
di: Karge, Jonas
Pubblicazione: (2026)
di: Karge, Jonas
Pubblicazione: (2026)
Mitigating LLM Hallucinations with Knowledge Graphs: A Case Study
di: Li, Harry, et al.
Pubblicazione: (2025)
di: Li, Harry, et al.
Pubblicazione: (2025)
Hallucination in LLM-Based Code Generation: An Automotive Case Study
di: Pavel, Marc, et al.
Pubblicazione: (2025)
di: Pavel, Marc, et al.
Pubblicazione: (2025)
To Trust or Not to Trust: On Calibration in ML-based Resource Allocation for Wireless Networks
di: Raina, Rashika, et al.
Pubblicazione: (2025)
di: Raina, Rashika, et al.
Pubblicazione: (2025)
Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations
di: Luo, Jinyuan, et al.
Pubblicazione: (2025)
di: Luo, Jinyuan, et al.
Pubblicazione: (2025)
Evaluating Human Trust in LLM-Based Planners: A Preliminary Study
di: Chen, Shenghui, et al.
Pubblicazione: (2025)
di: Chen, Shenghui, et al.
Pubblicazione: (2025)
Cross-Modal Attention Calibration for LVLM Hallucination Mitigation
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
Two Is Better Than One: Aligned Representation Pairs for Anomaly Detection
di: Ryser, Alain, et al.
Pubblicazione: (2024)
di: Ryser, Alain, et al.
Pubblicazione: (2024)
Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations
di: Cherukuri, Kalyan, et al.
Pubblicazione: (2026)
di: Cherukuri, Kalyan, et al.
Pubblicazione: (2026)
HalluLens: LLM Hallucination Benchmark
di: Bang, Yejin, et al.
Pubblicazione: (2025)
di: Bang, Yejin, et al.
Pubblicazione: (2025)
Can You Trust an LLM with Your Life-Changing Decision? An Investigation into AI High-Stakes Responses
di: Cahyono, Joshua Adrian, et al.
Pubblicazione: (2025)
di: Cahyono, Joshua Adrian, et al.
Pubblicazione: (2025)
MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them
di: Zhang, Weichen, et al.
Pubblicazione: (2025)
di: Zhang, Weichen, et al.
Pubblicazione: (2025)
Dynamic Trust Calibration Using Contextual Bandits
di: Henrique, Bruno M., et al.
Pubblicazione: (2025)
di: Henrique, Bruno M., et al.
Pubblicazione: (2025)
TERMS-Bench: Diagnosing LLM Negotiation Agents Beyond Deal Rate
di: Zhang, Erica, et al.
Pubblicazione: (2026)
di: Zhang, Erica, et al.
Pubblicazione: (2026)
HALT-RAG: A Task-Adaptable Framework for Hallucination Detection with Calibrated NLI Ensembles and Abstention
di: Goswami, Saumya, et al.
Pubblicazione: (2025)
di: Goswami, Saumya, et al.
Pubblicazione: (2025)
Dealing with Inconsistency for Reasoning over Knowledge Graphs: A Survey
di: Nentidis, Anastasios, et al.
Pubblicazione: (2025)
di: Nentidis, Anastasios, et al.
Pubblicazione: (2025)
Banishing LLM Hallucinations Requires Rethinking Generalization
di: Li, Johnny, et al.
Pubblicazione: (2024)
di: Li, Johnny, et al.
Pubblicazione: (2024)
Can We Trust LLM Detectors?
di: Sandhan, Jivnesh, et al.
Pubblicazione: (2026)
di: Sandhan, Jivnesh, et al.
Pubblicazione: (2026)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
di: Islam, Saad Obaid ul, et al.
Pubblicazione: (2025)
di: Islam, Saad Obaid ul, et al.
Pubblicazione: (2025)
LLM-based Agents Suffer from Hallucinations: A Survey of Taxonomy, Methods, and Directions
di: Lin, Xixun, et al.
Pubblicazione: (2025)
di: Lin, Xixun, et al.
Pubblicazione: (2025)
Algorithmically Establishing Trust in Evaluators
di: de Wynter, Adrian
Pubblicazione: (2025)
di: de Wynter, Adrian
Pubblicazione: (2025)
Learn to Code Sustainably: An Empirical Study on LLM-based Green Code Generation
di: Vartziotis, Tina, et al.
Pubblicazione: (2024)
di: Vartziotis, Tina, et al.
Pubblicazione: (2024)
LLM-REVal: Can We Trust LLM Reviewers Yet?
di: Li, Rui, et al.
Pubblicazione: (2025)
di: Li, Rui, et al.
Pubblicazione: (2025)
CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs
di: Huang, Xiaoyi, et al.
Pubblicazione: (2026)
di: Huang, Xiaoyi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
RubiSCoT: A Framework for AI-Supported Academic Assessment
di: Fröhlich, Thorsten, et al.
Pubblicazione: (2025) -
A Cross-Cultural Assessment of Human Ability to Detect LLM-Generated Fake News about South Africa
di: Schlippe, Tim, et al.
Pubblicazione: (2025) -
Exploring ChatGPT's Empathic Abilities
di: Schaaff, Kristina, et al.
Pubblicazione: (2023) -
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
di: Wagner, Julia, et al.
Pubblicazione: (2026) -
Classification of Human- and AI-Generated Texts for English, French, German, and Spanish
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)