HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Dasgupta, Sharanya, Nath, Sujoy, Basu, Arkaprabha, Shamsolmoali, Pourya, Das, Swagatam |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs
di: Nath, Sujoy, et al.
Pubblicazione: (2025)
di: Nath, Sujoy, et al.
Pubblicazione: (2025)
ARREST: Adversarial Resilient Regulation Enhancing Safety and Truth in Large Language Models
di: Dasgupta, Sharanya, et al.
Pubblicazione: (2026)
di: Dasgupta, Sharanya, et al.
Pubblicazione: (2026)
Zero-Shot Confidence Estimation for Small LLMs: When Supervised Baselines Aren't Worth Training
di: Nguyen, Luong N.
Pubblicazione: (2026)
di: Nguyen, Luong N.
Pubblicazione: (2026)
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits
di: Dikshit, Subrit, et al.
Pubblicazione: (2025)
di: Dikshit, Subrit, et al.
Pubblicazione: (2025)
The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance
di: Mohanty, Anwesha, et al.
Pubblicazione: (2025)
di: Mohanty, Anwesha, et al.
Pubblicazione: (2025)
Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making
di: Amin, Danial
Pubblicazione: (2026)
di: Amin, Danial
Pubblicazione: (2026)
Are Large Language Models Reliable Argument Quality Annotators?
di: Mirzakhmedova, Nailia, et al.
Pubblicazione: (2024)
di: Mirzakhmedova, Nailia, et al.
Pubblicazione: (2024)
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models
di: Ke, Shih-Wen, et al.
Pubblicazione: (2025)
di: Ke, Shih-Wen, et al.
Pubblicazione: (2025)
A Novel Nuanced Conversation Evaluation Framework for Large Language Models in Mental Health
di: Marrapese, Alexander, et al.
Pubblicazione: (2024)
di: Marrapese, Alexander, et al.
Pubblicazione: (2024)
How Many Bytes Can You Take Out Of Brain-To-Text Decoding?
di: Antonello, Richard, et al.
Pubblicazione: (2024)
di: Antonello, Richard, et al.
Pubblicazione: (2024)
Artificial Agency and Large Language Models
di: van Lier, Maud, et al.
Pubblicazione: (2024)
di: van Lier, Maud, et al.
Pubblicazione: (2024)
Can Large Language Models Act as Symbolic Reasoners?
di: Sullivan, Rob, et al.
Pubblicazione: (2024)
di: Sullivan, Rob, et al.
Pubblicazione: (2024)
Quantifying the Effectiveness of Student Organization Activities using Natural Language Processing
di: Taruc, Lyberius Ennio F., et al.
Pubblicazione: (2024)
di: Taruc, Lyberius Ennio F., et al.
Pubblicazione: (2024)
Neurosymbolic Graph Enrichment for Grounded World Models
di: De Giorgis, Stefano, et al.
Pubblicazione: (2024)
di: De Giorgis, Stefano, et al.
Pubblicazione: (2024)
Automated Thematic Analyses Using LLMs: Xylazine Wound Management Social Media Chatter Use Case
di: Hairston, JaMor, et al.
Pubblicazione: (2025)
di: Hairston, JaMor, et al.
Pubblicazione: (2025)
"Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills
di: Liu, Yi, et al.
Pubblicazione: (2026)
di: Liu, Yi, et al.
Pubblicazione: (2026)
BHRAM-IL: A Benchmark for Hallucination Recognition and Assessment in Multiple Indian Languages
di: Terdalkar, Hrishikesh, et al.
Pubblicazione: (2025)
di: Terdalkar, Hrishikesh, et al.
Pubblicazione: (2025)
The Fair Game: Auditing & Debiasing AI Algorithms Over Time
di: Basu, Debabrota, et al.
Pubblicazione: (2025)
di: Basu, Debabrota, et al.
Pubblicazione: (2025)
Exponential Shift: Humans Adapt to AI Economies
di: McNamara, Kevin J, et al.
Pubblicazione: (2025)
di: McNamara, Kevin J, et al.
Pubblicazione: (2025)
Language-Dependent Political Bias in AI: A Study of ChatGPT and Gemini
di: Yuksel, Dogus, et al.
Pubblicazione: (2025)
di: Yuksel, Dogus, et al.
Pubblicazione: (2025)
A Review of Challenges in Speech-based Conversational AI for Elderly Care
di: Klaassen, Willemijn, et al.
Pubblicazione: (2024)
di: Klaassen, Willemijn, et al.
Pubblicazione: (2024)
Can Large Language Models Understand As Well As Apply Patent Regulations to Pass a Hands-On Patent Attorney Test?
di: Khera, Bhakti, et al.
Pubblicazione: (2025)
di: Khera, Bhakti, et al.
Pubblicazione: (2025)
Reflexive Prompt Engineering: A Framework for Responsible Prompt Engineering and Interaction Design
di: Djeffal, Christian
Pubblicazione: (2025)
di: Djeffal, Christian
Pubblicazione: (2025)
Reasoning Abilities of Large Language Models: In-Depth Analysis on the Abstraction and Reasoning Corpus
di: Lee, Seungpil, et al.
Pubblicazione: (2024)
di: Lee, Seungpil, et al.
Pubblicazione: (2024)
Use of AI Tools: Guidelines to Maintain Academic Integrity in Computing Colleges
di: El-boghdadi, Hatem M., et al.
Pubblicazione: (2026)
di: El-boghdadi, Hatem M., et al.
Pubblicazione: (2026)
Which English Do LLMs Prefer? Triangulating Structural Bias Towards American English in Foundation Models
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2026)
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2026)
An Evaluation of LLMs for Detecting Harmful Computing Terms
di: Jacas, Joshua, et al.
Pubblicazione: (2025)
di: Jacas, Joshua, et al.
Pubblicazione: (2025)
InvestAlign: Overcoming Data Scarcity in Aligning Large Language Models with Investor Decision-Making Processes under Herd Behavior
di: Wang, Huisheng, et al.
Pubblicazione: (2025)
di: Wang, Huisheng, et al.
Pubblicazione: (2025)
On the Limitations of Compute Thresholds as a Governance Strategy
di: Hooker, Sara
Pubblicazione: (2024)
di: Hooker, Sara
Pubblicazione: (2024)
An Autonomous GIS Agent Framework for Geospatial Data Retrieval
di: Ning, Huan, et al.
Pubblicazione: (2024)
di: Ning, Huan, et al.
Pubblicazione: (2024)
Teaching Programming in the Age of Generative AI: Insights from Literature, Pedagogical Proposals, and Student Perspectives
di: Rubio-Manzano, Clemente, et al.
Pubblicazione: (2025)
di: Rubio-Manzano, Clemente, et al.
Pubblicazione: (2025)
ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents
di: Lu, Yuxing, et al.
Pubblicazione: (2026)
di: Lu, Yuxing, et al.
Pubblicazione: (2026)
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
di: Liu, Qibang, et al.
Pubblicazione: (2025)
di: Liu, Qibang, et al.
Pubblicazione: (2025)
Automatic Generation of Behavioral Test Cases For Natural Language Processing Using Clustering and Prompting
di: Li, Ying, et al.
Pubblicazione: (2024)
di: Li, Ying, et al.
Pubblicazione: (2024)
Combinatorial Reasoning: Selecting Reasons in Generative AI Pipelines via Combinatorial Optimization
di: Esencan, Mert, et al.
Pubblicazione: (2024)
di: Esencan, Mert, et al.
Pubblicazione: (2024)
Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data
di: Gerstgrasser, Matthias, et al.
Pubblicazione: (2024)
di: Gerstgrasser, Matthias, et al.
Pubblicazione: (2024)
An Auditable Pipeline for Fuzzy Full-Text Screening in Systematic Reviews: Integrating Contrastive Semantic Highlighting and LLM Judgment
di: Mortezaagha, Pouria, et al.
Pubblicazione: (2025)
di: Mortezaagha, Pouria, et al.
Pubblicazione: (2025)
Embedding-Aligned Language Models
di: Tennenholtz, Guy, et al.
Pubblicazione: (2024)
di: Tennenholtz, Guy, et al.
Pubblicazione: (2024)
Synergistic Simulations: Multi-Agent Problem Solving with Large Language Models
di: Sprigler, Asher, et al.
Pubblicazione: (2024)
di: Sprigler, Asher, et al.
Pubblicazione: (2024)
Scaling Multiagent Systems with Process Rewards
di: Li, Ed, et al.
Pubblicazione: (2026)
di: Li, Ed, et al.
Pubblicazione: (2026)
Documenti analoghi
-
HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs
di: Nath, Sujoy, et al.
Pubblicazione: (2025) -
ARREST: Adversarial Resilient Regulation Enhancing Safety and Truth in Large Language Models
di: Dasgupta, Sharanya, et al.
Pubblicazione: (2026) -
Zero-Shot Confidence Estimation for Small LLMs: When Supervised Baselines Aren't Worth Training
di: Nguyen, Luong N.
Pubblicazione: (2026) -
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits
di: Dikshit, Subrit, et al.
Pubblicazione: (2025) -
The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance
di: Mohanty, Anwesha, et al.
Pubblicazione: (2025)