A novel hallucination classification framework
Fuente:
arXiv
Salvato in:
| Autori principali: | Zavhorodnii, Maksym, Dehtiarov, Dmytro, Konovalenko, Anna |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Probabilistic distances-based hallucination detection in LLMs with RAG
di: Oblovatny, Rodion, et al.
Pubblicazione: (2025)
di: Oblovatny, Rodion, et al.
Pubblicazione: (2025)
Halluverse-M^3: A multitask multilingual benchmark for hallucination in LLMs
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
Do LLM hallucination detectors suffer from low-resource effect?
di: Datta, Debtanu, et al.
Pubblicazione: (2026)
di: Datta, Debtanu, et al.
Pubblicazione: (2026)
Ask-EDA: A Design Assistant Empowered by LLM, Hybrid RAG and Abbreviation De-hallucination
di: Shi, Luyao, et al.
Pubblicazione: (2024)
di: Shi, Luyao, et al.
Pubblicazione: (2024)
The impact of fine tuning in LLaMA on hallucinations for named entity extraction in legal documentation
di: Vargas, Francisco, et al.
Pubblicazione: (2025)
di: Vargas, Francisco, et al.
Pubblicazione: (2025)
ACL-Verbatim: hallucination-free question answering for research
di: Recski, Gábor, et al.
Pubblicazione: (2026)
di: Recski, Gábor, et al.
Pubblicazione: (2026)
A comprehensive taxonomy of hallucinations in Large Language Models
di: Cossio, Manuel
Pubblicazione: (2025)
di: Cossio, Manuel
Pubblicazione: (2025)
Beyond Textual Context: Structural Graph Encoding with Adaptive Space Alignment to alleviate the hallucination of LLMs
di: Zhang, Yifang, et al.
Pubblicazione: (2025)
di: Zhang, Yifang, et al.
Pubblicazione: (2025)
SLPL SHROOM at SemEval2024 Task 06: A comprehensive study on models ability to detect hallucination
di: Fallah, Pouya, et al.
Pubblicazione: (2024)
di: Fallah, Pouya, et al.
Pubblicazione: (2024)
Strong hallucinations from negation and how to fix them
di: Asher, Nicholas, et al.
Pubblicazione: (2024)
di: Asher, Nicholas, et al.
Pubblicazione: (2024)
A systematic framework for generating novel experimental hypotheses from language models
di: Misra, Kanishka, et al.
Pubblicazione: (2024)
di: Misra, Kanishka, et al.
Pubblicazione: (2024)
Zero-knowledge LLM hallucination detection and mitigation through fine-grained cross-model consistency
di: Goel, Aman, et al.
Pubblicazione: (2025)
di: Goel, Aman, et al.
Pubblicazione: (2025)
Reducing hallucination in structured outputs via Retrieval-Augmented Generation
di: Béchard, Patrice, et al.
Pubblicazione: (2024)
di: Béchard, Patrice, et al.
Pubblicazione: (2024)
Stance Reasoner: Zero-Shot Stance Detection on Social Media with Explicit Reasoning
di: Taranukhin, Maksym, et al.
Pubblicazione: (2024)
di: Taranukhin, Maksym, et al.
Pubblicazione: (2024)
HalluHard: A Hard Multi-Turn Hallucination Benchmark
di: Fan, Dongyang, et al.
Pubblicazione: (2026)
di: Fan, Dongyang, et al.
Pubblicazione: (2026)
First is Not Really Better Than Last: Evaluating Layer Choice and Aggregation Strategies in Language Model Data Influence Estimation
di: Vitel, Dmytro, et al.
Pubblicazione: (2025)
di: Vitel, Dmytro, et al.
Pubblicazione: (2025)
Empowering Air Travelers: A Chatbot for Canadian Air Passenger Rights
di: Taranukhin, Maksym, et al.
Pubblicazione: (2024)
di: Taranukhin, Maksym, et al.
Pubblicazione: (2024)
Deep Language Geometry: Constructing a Metric Space from LLM Weights
di: Shamrai, Maksym, et al.
Pubblicazione: (2025)
di: Shamrai, Maksym, et al.
Pubblicazione: (2025)
Does Refusal Training in LLMs Generalize to the Past Tense?
di: Andriushchenko, Maksym, et al.
Pubblicazione: (2024)
di: Andriushchenko, Maksym, et al.
Pubblicazione: (2024)
Towards Outcome-Oriented, Task-Agnostic Evaluation of AI Agents
di: AlShikh, Waseem, et al.
Pubblicazione: (2025)
di: AlShikh, Waseem, et al.
Pubblicazione: (2025)
Movie2Story: A framework for understanding videos and telling stories in the form of novel text
di: Li, Kangning, et al.
Pubblicazione: (2024)
di: Li, Kangning, et al.
Pubblicazione: (2024)
Talking with Oompa Loompas: A novel framework for evaluating linguistic acquisition of LLM agents
di: Swain, Sankalp Tattwadarshi, et al.
Pubblicazione: (2025)
di: Swain, Sankalp Tattwadarshi, et al.
Pubblicazione: (2025)
LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics
di: Humblot-Renaux, Galadrielle, et al.
Pubblicazione: (2026)
di: Humblot-Renaux, Galadrielle, et al.
Pubblicazione: (2026)
LLMs as Strategic Actors: Behavioral Alignment, Risk Calibration, and Argumentation Framing in Geopolitical Simulations
di: Solopova, Veronika, et al.
Pubblicazione: (2026)
di: Solopova, Veronika, et al.
Pubblicazione: (2026)
Language Ranker: A Lightweight Ranking framework for LLM Decoding
di: Zhang, Chenheng, et al.
Pubblicazione: (2025)
di: Zhang, Chenheng, et al.
Pubblicazione: (2025)
MOSLIM:Align with diverse preferences in prompts through reward classification
di: Zhang, Yu, et al.
Pubblicazione: (2025)
di: Zhang, Yu, et al.
Pubblicazione: (2025)
Text classification optimization algorithm based on graph neural network
di: Gao, Erdi, et al.
Pubblicazione: (2024)
di: Gao, Erdi, et al.
Pubblicazione: (2024)
A thorough benchmark of automatic text classification: From traditional approaches to large language models
di: Cunha, Washington, et al.
Pubblicazione: (2025)
di: Cunha, Washington, et al.
Pubblicazione: (2025)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
di: Ziki, Batsirayi Mupamhi, et al.
Pubblicazione: (2026)
di: Ziki, Batsirayi Mupamhi, et al.
Pubblicazione: (2026)
A new approach for fine-tuning sentence transformers for intent classification and out-of-scope detection tasks
di: Zhang, Tianyi, et al.
Pubblicazione: (2024)
di: Zhang, Tianyi, et al.
Pubblicazione: (2024)
Is In-Context Learning Sufficient for Instruction Following in LLMs?
di: Zhao, Hao, et al.
Pubblicazione: (2024)
di: Zhao, Hao, et al.
Pubblicazione: (2024)
Stop learning it all to mitigate visual hallucination, Focus on the hallucination target
di: Yoon, Dokyoon, et al.
Pubblicazione: (2025)
di: Yoon, Dokyoon, et al.
Pubblicazione: (2025)
Backtranslation and paraphrasing in the LLM era? Comparing data augmentation methods for emotion classification
di: Radliński, Łukasz, et al.
Pubblicazione: (2025)
di: Radliński, Łukasz, et al.
Pubblicazione: (2025)
Scaling few-shot spoken word classification with generative meta-continual learning
di: Beyers, Louise, et al.
Pubblicazione: (2026)
di: Beyers, Louise, et al.
Pubblicazione: (2026)
One Thousand and One Pairs: A "novel" challenge for long-context language models
di: Karpinska, Marzena, et al.
Pubblicazione: (2024)
di: Karpinska, Marzena, et al.
Pubblicazione: (2024)
Building A Unified AI-centric Language System: analysis, framework and future work
di: Wang, Edward Hong, et al.
Pubblicazione: (2025)
di: Wang, Edward Hong, et al.
Pubblicazione: (2025)
Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
di: Driouich, Ilias, et al.
Pubblicazione: (2025)
di: Driouich, Ilias, et al.
Pubblicazione: (2025)
Employing self-supervised learning models for cross-linguistic child speech maturity classification
di: Zhang, Theo, et al.
Pubblicazione: (2025)
di: Zhang, Theo, et al.
Pubblicazione: (2025)
Trusting CHATGPT: how minor tweaks in the prompts lead to major differences in sentiment classification
di: Cuellar, Jaime E., et al.
Pubblicazione: (2025)
di: Cuellar, Jaime E., et al.
Pubblicazione: (2025)
Ensemble BERT: A student social network text sentiment classification model based on ensemble learning and BERT architecture
di: Jiang, Kai, et al.
Pubblicazione: (2024)
di: Jiang, Kai, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Probabilistic distances-based hallucination detection in LLMs with RAG
di: Oblovatny, Rodion, et al.
Pubblicazione: (2025) -
Halluverse-M^3: A multitask multilingual benchmark for hallucination in LLMs
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026) -
Do LLM hallucination detectors suffer from low-resource effect?
di: Datta, Debtanu, et al.
Pubblicazione: (2026) -
Ask-EDA: A Design Assistant Empowered by LLM, Hybrid RAG and Abbreviation De-hallucination
di: Shi, Luyao, et al.
Pubblicazione: (2024) -
The impact of fine tuning in LLaMA on hallucinations for named entity extraction in legal documentation
di: Vargas, Francisco, et al.
Pubblicazione: (2025)