HALT: Hallucination Assessment via Log-probs as Time series
Fuente:
arXiv
Salvato in:
| Autori principali: | Shapiro, Ahmad, Taneja, Karan, Goel, Ashok |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
di: Taneja, Karan, et al.
Pubblicazione: (2025)
di: Taneja, Karan, et al.
Pubblicazione: (2025)
Can Active Label Correction Improve LLM-based Modular AI Systems?
di: Taneja, Karan, et al.
Pubblicazione: (2024)
di: Taneja, Karan, et al.
Pubblicazione: (2024)
HALT-RAG: A Task-Adaptable Framework for Hallucination Detection with Calibrated NLI Ensembles and Abstention
di: Goswami, Saumya, et al.
Pubblicazione: (2025)
di: Goswami, Saumya, et al.
Pubblicazione: (2025)
Jill Watson: A Virtual Teaching Assistant powered by ChatGPT
di: Taneja, Karan, et al.
Pubblicazione: (2024)
di: Taneja, Karan, et al.
Pubblicazione: (2024)
Impact of Multimodal and Conversational AI on Learning Outcomes and Experience
di: Taneja, Karan, et al.
Pubblicazione: (2026)
di: Taneja, Karan, et al.
Pubblicazione: (2026)
Towards a Multimodal Document-grounded Conversational AI System for Education
di: Taneja, Karan, et al.
Pubblicazione: (2025)
di: Taneja, Karan, et al.
Pubblicazione: (2025)
Monte Carlo Tree Search for Recipe Generation using GPT-2
di: Taneja, Karan, et al.
Pubblicazione: (2024)
di: Taneja, Karan, et al.
Pubblicazione: (2024)
High Accuracy, Less Talk (HALT): Reliable LLMs through Capability-Aligned Finetuning
di: Franzmeyer, Tim, et al.
Pubblicazione: (2025)
di: Franzmeyer, Tim, et al.
Pubblicazione: (2025)
HIDE and Seek: Detecting Hallucinations in Language Models via Decoupled Representations
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025)
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025)
SymTax: Symbiotic Relationship and Taxonomy Fusion for Effective Citation Recommendation
di: Goyal, Karan, et al.
Pubblicazione: (2024)
di: Goyal, Karan, et al.
Pubblicazione: (2024)
Controlled Automatic Task-Specific Synthetic Data Generation for Hallucination Detection
di: Xie, Yong, et al.
Pubblicazione: (2024)
di: Xie, Yong, et al.
Pubblicazione: (2024)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
di: Liu, Emmy, et al.
Pubblicazione: (2026)
di: Liu, Emmy, et al.
Pubblicazione: (2026)
Evaluating Learner Representations for Differentiation Prior to Instructional Outcomes
di: Park, Junsoo, et al.
Pubblicazione: (2026)
di: Park, Junsoo, et al.
Pubblicazione: (2026)
Learning from Contrastive Prompts: Automated Optimization and Adaptation
di: Li, Mingqi, et al.
Pubblicazione: (2024)
di: Li, Mingqi, et al.
Pubblicazione: (2024)
ChatLog: Carefully Evaluating the Evolution of ChatGPT Across Time
di: Tu, Shangqing, et al.
Pubblicazione: (2023)
di: Tu, Shangqing, et al.
Pubblicazione: (2023)
Enhancing Hallucination Detection via Future Context
di: Lee, Joosung, et al.
Pubblicazione: (2025)
di: Lee, Joosung, et al.
Pubblicazione: (2025)
Hallucination Detection and Hallucination Mitigation: An Investigation
di: Luo, Junliang, et al.
Pubblicazione: (2024)
di: Luo, Junliang, et al.
Pubblicazione: (2024)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
di: Jiang, Xinyan, et al.
Pubblicazione: (2025)
di: Jiang, Xinyan, et al.
Pubblicazione: (2025)
Self-Explanation in Social AI Agents
di: Basappa, Rhea, et al.
Pubblicazione: (2025)
di: Basappa, Rhea, et al.
Pubblicazione: (2025)
AlignCheck: a Semantic Open-Domain Metric for Factual Consistency Assessment
di: Aghaebrahimian, Ahmad
Pubblicazione: (2025)
di: Aghaebrahimian, Ahmad
Pubblicazione: (2025)
Removal of Hallucination on Hallucination: Debate-Augmented RAG
di: Hu, Wentao, et al.
Pubblicazione: (2025)
di: Hu, Wentao, et al.
Pubblicazione: (2025)
AI Hallucination from Students' Perspective: A Thematic Analysis
di: Shoufan, Abdulhadi, et al.
Pubblicazione: (2026)
di: Shoufan, Abdulhadi, et al.
Pubblicazione: (2026)
LLM-CAS: Dynamic Neuron Perturbation for Real-Time Hallucination Correction
di: Zhang, Jensen, et al.
Pubblicazione: (2025)
di: Zhang, Jensen, et al.
Pubblicazione: (2025)
Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits
di: Nielsen, Erik, et al.
Pubblicazione: (2026)
di: Nielsen, Erik, et al.
Pubblicazione: (2026)
HARP: Hallucination Detection via Reasoning Subspace Projection
di: Hu, Junjie, et al.
Pubblicazione: (2025)
di: Hu, Junjie, et al.
Pubblicazione: (2025)
A Unified Definition of Hallucination: It's The World Model, Stupid!
di: Liu, Emmy, et al.
Pubblicazione: (2025)
di: Liu, Emmy, et al.
Pubblicazione: (2025)
The Energy of Falsehood: Detecting Hallucinations via Diffusion Model Likelihoods
di: Gautam, Arpit Singh, et al.
Pubblicazione: (2026)
di: Gautam, Arpit Singh, et al.
Pubblicazione: (2026)
Do Benchmarks Underestimate LLM Performance? Evaluating Hallucination Detection With LLM-First Human-Adjudicated Assessment
di: Atasoy, I. F., et al.
Pubblicazione: (2026)
di: Atasoy, I. F., et al.
Pubblicazione: (2026)
Alleviating Hallucinations of Large Language Models through Induced Hallucinations
di: Zhang, Yue, et al.
Pubblicazione: (2023)
di: Zhang, Yue, et al.
Pubblicazione: (2023)
Memory-Based vs. Context-Only Conditioning Produces Distinct Behavioral Patterns in Stateful Personalization
di: Park, Junsoo, et al.
Pubblicazione: (2026)
di: Park, Junsoo, et al.
Pubblicazione: (2026)
Unsupervised Real-Time Hallucination Detection based on the Internal States of Large Language Models
di: Su, Weihang, et al.
Pubblicazione: (2024)
di: Su, Weihang, et al.
Pubblicazione: (2024)
HausaNLP at SemEval-2025 Task 3: Towards a Fine-Grained Model-Aware Hallucination Detection
di: Bala, Maryam, et al.
Pubblicazione: (2025)
di: Bala, Maryam, et al.
Pubblicazione: (2025)
Black-Box Hallucination Detection via Consistency Under the Uncertain Expression
di: Joo, Seongho, et al.
Pubblicazione: (2025)
di: Joo, Seongho, et al.
Pubblicazione: (2025)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
The Path of Least Resistance: Guiding LLM Reasoning Trajectories with Prefix Consensus
di: Jindal, Ishan, et al.
Pubblicazione: (2026)
di: Jindal, Ishan, et al.
Pubblicazione: (2026)
Don't Let It Hallucinate: Premise Verification via Retrieval-Augmented Logical Reasoning
di: Qin, Yuehan, et al.
Pubblicazione: (2025)
di: Qin, Yuehan, et al.
Pubblicazione: (2025)
Controllable Text Generation in the Instruction-Tuning Era
di: Ashok, Dhananjay, et al.
Pubblicazione: (2024)
di: Ashok, Dhananjay, et al.
Pubblicazione: (2024)
FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs
di: Bao, Forrest Sheng, et al.
Pubblicazione: (2024)
di: Bao, Forrest Sheng, et al.
Pubblicazione: (2024)
Triggering Hallucinations in LLMs: A Quantitative Study of Prompt-Induced Hallucination in Large Language Models
di: Sato, Makoto
Pubblicazione: (2025)
di: Sato, Makoto
Pubblicazione: (2025)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
di: Islam, Saad Obaid ul, et al.
Pubblicazione: (2025)
di: Islam, Saad Obaid ul, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
di: Taneja, Karan, et al.
Pubblicazione: (2025) -
Can Active Label Correction Improve LLM-based Modular AI Systems?
di: Taneja, Karan, et al.
Pubblicazione: (2024) -
HALT-RAG: A Task-Adaptable Framework for Hallucination Detection with Calibrated NLI Ensembles and Abstention
di: Goswami, Saumya, et al.
Pubblicazione: (2025) -
Jill Watson: A Virtual Teaching Assistant powered by ChatGPT
di: Taneja, Karan, et al.
Pubblicazione: (2024) -
Impact of Multimodal and Conversational AI on Learning Outcomes and Experience
di: Taneja, Karan, et al.
Pubblicazione: (2026)