Not All That Is Fluent Is Factual: Investigating Hallucinations of Large Language Models in Academic Writing
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Khan, Humam, Nafis, Md Tabrez, Sohail, Shahab Saquib, Khalique, Aqeel, Khan, Rehan Hasan |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Hallucination to Truth: A Review of Fact-Checking and Factuality Evaluation in Large Language Models
par: Rahman, Subhey Sadi, et autres
Publié: (2025)
par: Rahman, Subhey Sadi, et autres
Publié: (2025)
Negation Blindness in Large Language Models: Unveiling the NO Syndrome in Image Generation
par: Nadeem, Mohammad, et autres
Publié: (2024)
par: Nadeem, Mohammad, et autres
Publié: (2024)
From Text to Transformation: A Comprehensive Review of Large Language Models' Versatility
par: Kaur, Pravneet, et autres
Publié: (2024)
par: Kaur, Pravneet, et autres
Publié: (2024)
PretrainRL: Alleviating Factuality Hallucination of Large Language Models at the Beginning
par: Liu, Langming, et autres
Publié: (2026)
par: Liu, Langming, et autres
Publié: (2026)
AutoHall: Automated Factuality Hallucination Dataset Generation for Large Language Models
par: Cao, Zouying, et autres
Publié: (2023)
par: Cao, Zouying, et autres
Publié: (2023)
The Dawn After the Dark: An Empirical Study on Factuality Hallucination in Large Language Models
par: Li, Junyi, et autres
Publié: (2024)
par: Li, Junyi, et autres
Publié: (2024)
Owls are wise and foxes are unfaithful: Uncovering animal stereotypes in vision-language models
par: Aman, Tabinda, et autres
Publié: (2025)
par: Aman, Tabinda, et autres
Publié: (2025)
From RAG to Agentic: Validating Islamic-Medicine Responses with LLM Agents
par: Sayeed, Mohammad Amaan, et autres
Publié: (2025)
par: Sayeed, Mohammad Amaan, et autres
Publié: (2025)
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages
par: Azam, Gulfarogh, et autres
Publié: (2025)
par: Azam, Gulfarogh, et autres
Publié: (2025)
Investigating Hallucination in Conversations for Low Resource Languages
par: Das, Amit, et autres
Publié: (2025)
par: Das, Amit, et autres
Publié: (2025)
DHI: Leveraging Diverse Hallucination Induction for Enhanced Contrastive Factuality Control in Large Language Models
par: Guo, Jiani, et autres
Publié: (2026)
par: Guo, Jiani, et autres
Publié: (2026)
Logical Consistency of Large Language Models in Fact-checking
par: Ghosh, Bishwamittra, et autres
Publié: (2024)
par: Ghosh, Bishwamittra, et autres
Publié: (2024)
Investigating Multi-Hop Factual Shortcuts in Knowledge Editing of Large Language Models
par: Ju, Tianjie, et autres
Publié: (2024)
par: Ju, Tianjie, et autres
Publié: (2024)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
par: Yu, Lei, et autres
Publié: (2024)
par: Yu, Lei, et autres
Publié: (2024)
Exploring the Limits of Pruning: Task-Specific Neurons, Model Collapse, and Recovery in Task-Specific Large Language Models
par: Siam, M. K. Khalidi, et autres
Publié: (2026)
par: Siam, M. K. Khalidi, et autres
Publié: (2026)
Unmasking Hallucinations: A Causal Graph-Attention Perspective on Factual Reliability in Large Language Models
par: kurra, Sailesh kiran, et autres
Publié: (2026)
par: kurra, Sailesh kiran, et autres
Publié: (2026)
Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation
par: Ren, Ruiyang, et autres
Publié: (2023)
par: Ren, Ruiyang, et autres
Publié: (2023)
Mitigating Geospatial Knowledge Hallucination in Large Language Models: Benchmarking and Dynamic Factuality Aligning
par: Wang, Shengyuan, et autres
Publié: (2025)
par: Wang, Shengyuan, et autres
Publié: (2025)
Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models
par: Siddiqui, S M Tahmid, et autres
Publié: (2026)
par: Siddiqui, S M Tahmid, et autres
Publié: (2026)
Fluent but Unfeeling: The Emotional Blind Spots of Language Models
par: Shu, Bangzhao, et autres
Publié: (2025)
par: Shu, Bangzhao, et autres
Publié: (2025)
Alternate Preference Optimization for Unlearning Factual Knowledge in Large Language Models
par: Mekala, Anmol, et autres
Publié: (2024)
par: Mekala, Anmol, et autres
Publié: (2024)
Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models
par: Li, Junyi, et autres
Publié: (2025)
par: Li, Junyi, et autres
Publié: (2025)
An Early Investigation into the Utility of Multimodal Large Language Models in Medical Imaging
par: Khan, Sulaiman, et autres
Publié: (2024)
par: Khan, Sulaiman, et autres
Publié: (2024)
PM-LLM-Benchmark: Evaluating Large Language Models on Process Mining Tasks
par: Berti, Alessandro, et autres
Publié: (2024)
par: Berti, Alessandro, et autres
Publié: (2024)
REFIND at SemEval-2025 Task 3: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models
par: Lee, DongGeon, et autres
Publié: (2025)
par: Lee, DongGeon, et autres
Publié: (2025)
Leveraging the Domain Adaptation of Retrieval Augmented Generation Models for Question Answering and Reducing Hallucination
par: Rakin, Salman, et autres
Publié: (2024)
par: Rakin, Salman, et autres
Publié: (2024)
ScholarCopilot: Training Large Language Models for Academic Writing with Accurate Citations
par: Wang, Yubo, et autres
Publié: (2025)
par: Wang, Yubo, et autres
Publié: (2025)
Bridging Domain Knowledge and Process Discovery Using Large Language Models
par: Norouzifar, Ali, et autres
Publié: (2024)
par: Norouzifar, Ali, et autres
Publié: (2024)
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
par: Chrysostomou, George, et autres
Publié: (2023)
par: Chrysostomou, George, et autres
Publié: (2023)
OverleafCopilot: Empowering Academic Writing in Overleaf with Large Language Models
par: Wen, Haomin, et autres
Publié: (2024)
par: Wen, Haomin, et autres
Publié: (2024)
Evaluating Large Language Models on Urdu Idiom Translation
par: Khan, Muhammad Farmal, et autres
Publié: (2025)
par: Khan, Muhammad Farmal, et autres
Publié: (2025)
Factuality of Large Language Models: A Survey
par: Wang, Yuxia, et autres
Publié: (2024)
par: Wang, Yuxia, et autres
Publié: (2024)
One SPACE to Rule Them All: Jointly Mitigating Factuality and Faithfulness Hallucinations in LLMs
par: Wang, Pengbo, et autres
Publié: (2025)
par: Wang, Pengbo, et autres
Publié: (2025)
Exploring the Factual Consistency in Dialogue Comprehension of Large Language Models
par: She, Shuaijie, et autres
Publié: (2023)
par: She, Shuaijie, et autres
Publié: (2023)
IndicFairFace: Balanced Indian Face Dataset for Auditing and Mitigating Geographical Bias in Vision-Language Models
par: Mohsin, Aarish Shah, et autres
Publié: (2026)
par: Mohsin, Aarish Shah, et autres
Publié: (2026)
On Early Detection of Hallucinations in Factual Question Answering
par: Snyder, Ben, et autres
Publié: (2023)
par: Snyder, Ben, et autres
Publié: (2023)
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
par: Chaduvula, Sindhuja, et autres
Publié: (2026)
par: Chaduvula, Sindhuja, et autres
Publié: (2026)
Better Call SAUL: Fluent and Consistent Language Model Editing with Generation Regularization
par: Wang, Mingyang, et autres
Publié: (2024)
par: Wang, Mingyang, et autres
Publié: (2024)
Fluent dreaming for language models
par: Thompson, T. Ben, et autres
Publié: (2024)
par: Thompson, T. Ben, et autres
Publié: (2024)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
par: Zhong, Weihong, et autres
Publié: (2024)
par: Zhong, Weihong, et autres
Publié: (2024)
Documents similaires
-
Hallucination to Truth: A Review of Fact-Checking and Factuality Evaluation in Large Language Models
par: Rahman, Subhey Sadi, et autres
Publié: (2025) -
Negation Blindness in Large Language Models: Unveiling the NO Syndrome in Image Generation
par: Nadeem, Mohammad, et autres
Publié: (2024) -
From Text to Transformation: A Comprehensive Review of Large Language Models' Versatility
par: Kaur, Pravneet, et autres
Publié: (2024) -
PretrainRL: Alleviating Factuality Hallucination of Large Language Models at the Beginning
par: Liu, Langming, et autres
Publié: (2026) -
AutoHall: Automated Factuality Hallucination Dataset Generation for Large Language Models
par: Cao, Zouying, et autres
Publié: (2023)