Do I Really Know? Learning Factual Self-Verification for Hallucination Reduction
Fuente:
arXiv
Saved in:
| Main Authors: | Altinisik, Enes, Fatehkia, Masoomali, Deniz, Fatih, Durrani, Nadir, Hawasly, Majd, Raza, Mohammad, Sencar, Husrev Taha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FanarGuard: A Culturally-Aware Moderation Filter for Arabic Language Models
by: Fatehkia, Masoomali, et al.
Published: (2025)
by: Fatehkia, Masoomali, et al.
Published: (2025)
PAM: Training Policy-Aligned Moderation Filters at Scale
by: Fatehkia, Masoomali, et al.
Published: (2025)
by: Fatehkia, Masoomali, et al.
Published: (2025)
There Is More to Refusal in Large Language Models than a Single Direction
by: Joad, Faaiz, et al.
Published: (2026)
by: Joad, Faaiz, et al.
Published: (2026)
Tool Calling for Arabic LLMs: Data Strategies and Instruction Tuning
by: Ersoy, Asim, et al.
Published: (2025)
by: Ersoy, Asim, et al.
Published: (2025)
Explaining the role of Intrinsic Dimensionality in Adversarial Training
by: Altinisik, Enes, et al.
Published: (2024)
by: Altinisik, Enes, et al.
Published: (2024)
Scaling up Discovery of Latent Concepts in Deep NLP Models
by: Hawasly, Majd, et al.
Published: (2023)
by: Hawasly, Majd, et al.
Published: (2023)
Multimedia Forensics
by: Husrev Taha Sencar
by: Husrev Taha Sencar
Beyond the Leaderboard: Understanding Performance Disparities in Large Language Models via Model Diffing
by: Boughorbel, Sabri, et al.
Published: (2025)
by: Boughorbel, Sabri, et al.
Published: (2025)
Exploring Alignment in Shared Cross-lingual Spaces
by: Mousi, Basel, et al.
Published: (2024)
by: Mousi, Basel, et al.
Published: (2024)
From Text to Actionable Intelligence: Automating STIX Entity and Relationship Extraction
by: Lekssays, Ahmed, et al.
Published: (2025)
by: Lekssays, Ahmed, et al.
Published: (2025)
Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning
by: Saparkhan, Raman, et al.
Published: (2026)
by: Saparkhan, Raman, et al.
Published: (2026)
T-RAG: Lessons from the LLM Trenches
by: Fatehkia, Masoomali, et al.
Published: (2024)
by: Fatehkia, Masoomali, et al.
Published: (2024)
TechniqueRAG: Retrieval Augmented Generation for Adversarial Technique Annotation in Cyber Threat Intelligence Text
by: Lekssays, Ahmed, et al.
Published: (2025)
by: Lekssays, Ahmed, et al.
Published: (2025)
Fanar 2.0: Arabic Generative AI Stack
by: FANAR TEAM, et al.
Published: (2026)
by: FANAR TEAM, et al.
Published: (2026)
Semantic Ranking for Automated Adversarial Technique Annotation in Security Text
by: Kumarasinghe, Udesh, et al.
Published: (2024)
by: Kumarasinghe, Udesh, et al.
Published: (2024)
Improving Language Models Trained on Translated Data with Continual Pre-Training and Dictionary Learning Analysis
by: Boughorbel, Sabri, et al.
Published: (2024)
by: Boughorbel, Sabri, et al.
Published: (2024)
Fanar: An Arabic-Centric Multimodal Generative AI Platform
by: Fanar Team, et al.
Published: (2025)
by: Fanar Team, et al.
Published: (2025)
Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models
by: Mousi, Basel, et al.
Published: (2026)
by: Mousi, Basel, et al.
Published: (2026)
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
by: Chaduvula, Sindhuja, et al.
Published: (2026)
by: Chaduvula, Sindhuja, et al.
Published: (2026)
All I Really Need to Know I Learned in the Library.
by: Brennan, Michael
Published: (1992)
by: Brennan, Michael
Published: (1992)
Editing Across Languages: A Survey of Multilingual Knowledge Editing
by: Durrani, Nadir, et al.
Published: (2025)
by: Durrani, Nadir, et al.
Published: (2025)
An Exploration of Knowledge Editing for Arabic
by: Mousi, Basel, et al.
Published: (2025)
by: Mousi, Basel, et al.
Published: (2025)
Discovering Salient Neurons in Deep NLP Models
by: Durrani, Nadir, et al.
Published: (2022)
by: Durrani, Nadir, et al.
Published: (2022)
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
by: Ferrando, Javier, et al.
Published: (2024)
by: Ferrando, Javier, et al.
Published: (2024)
What Do They Really Need To Know? Adventures in Curriculum Writing.
by: Pritzl, Amy
Published: (2000)
by: Pritzl, Amy
Published: (2000)
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking
by: Dalvi, Fahim, et al.
Published: (2023)
by: Dalvi, Fahim, et al.
Published: (2023)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
by: Zhang, Xiaoying, et al.
Published: (2024)
by: Zhang, Xiaoying, et al.
Published: (2024)
Media Directors and Copyright Issues: How Much Do We Really Know?
by: Wertz, Sandra L, et al.
Published: (1994)
by: Wertz, Sandra L, et al.
Published: (1994)
Pedagogy for Practical Library Instruction: What Do We "Really" Need to Know?
by: Montgomery, Molly
Published: (2015)
by: Montgomery, Molly
Published: (2015)
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality
by: Ren, Baochang, et al.
Published: (2025)
by: Ren, Baochang, et al.
Published: (2025)
What Do Freshmen Really Know about Research? Assess before You Teach
by: Caspers, Jean, et al.
Published: (2005)
by: Caspers, Jean, et al.
Published: (2005)
Do We Really Know Why Carbon Nanotubes Grow? (Small 30/2026)
by: Nicola Verziaggi, et al.
Published: (2026)
by: Nicola Verziaggi, et al.
Published: (2026)
Strengthening Validation and Data Quality in Mobile Phone Surveys of Childhood Mortality
by: Raza Ur Rehman, et al.
Published: (2025)
by: Raza Ur Rehman, et al.
Published: (2025)
Comment on: "Delivering Care Consistent With the Psychosocial Standards: Provider Report From the iSTEPPP Study." Bridging the Gap Between Psychosocial Standards and Real‐World Implementation in Pediatric Oncology
by: Ismaeel Durrani, et al.
Published: (2026)
by: Ismaeel Durrani, et al.
Published: (2026)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
by: Cheang, Chi Seng, et al.
Published: (2025)
by: Cheang, Chi Seng, et al.
Published: (2025)
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks
by: Pan, Wenbo, et al.
Published: (2025)
by: Pan, Wenbo, et al.
Published: (2025)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
by: Durrani, Hamza Ahmed, et al.
Published: (2026)
by: Durrani, Hamza Ahmed, et al.
Published: (2026)
Commentary on “Electrocardiographic Predictors of Major Adverse Cardiovascular Events in Women With Suspected Ischemia and No Obstructive Coronary Artery Disease: Results of the Women's Ischemia Syndrome Evaluation”
by: Fatih Enes Durmaz, et al.
Published: (2026)
by: Fatih Enes Durmaz, et al.
Published: (2026)
Do Language Models Know When They're Hallucinating References?
by: Agrawal, Ayush, et al.
Published: (2023)
by: Agrawal, Ayush, et al.
Published: (2023)
Similar Items
-
FanarGuard: A Culturally-Aware Moderation Filter for Arabic Language Models
by: Fatehkia, Masoomali, et al.
Published: (2025) -
PAM: Training Policy-Aligned Moderation Filters at Scale
by: Fatehkia, Masoomali, et al.
Published: (2025) -
There Is More to Refusal in Large Language Models than a Single Direction
by: Joad, Faaiz, et al.
Published: (2026) -
Tool Calling for Arabic LLMs: Data Strategies and Instruction Tuning
by: Ersoy, Asim, et al.
Published: (2025) -
Explaining the role of Intrinsic Dimensionality in Adversarial Training
by: Altinisik, Enes, et al.
Published: (2024)