FACTOID: FACtual enTailment fOr hallucInation Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rawte, Vipula, Tonmoy, S. M Towhidul Islam, Rajbangshi, Krishnav, Nag, Shravani, Chadha, Aman, Sheth, Amit P., Das, Amitava |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
"Sorry, Come Again?" Prompting -- Enhancing Comprehension and Diminishing Hallucination with [PAUSE]-injected Optimal Paraphrasing
von: Rawte, Vipula, et al.
Veröffentlicht: (2024)
von: Rawte, Vipula, et al.
Veröffentlicht: (2024)
A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models
von: Tonmoy, S. M Towhidul Islam, et al.
Veröffentlicht: (2024)
von: Tonmoy, S. M Towhidul Islam, et al.
Veröffentlicht: (2024)
Visual Hallucination: Definition, Quantification, and Prescriptive Remediations
von: Rani, Anku, et al.
Veröffentlicht: (2024)
von: Rani, Anku, et al.
Veröffentlicht: (2024)
The What, Why, and How of Context Length Extension Techniques in Large Language Models -- A Detailed Survey
von: Pawar, Saurav, et al.
Veröffentlicht: (2024)
von: Pawar, Saurav, et al.
Veröffentlicht: (2024)
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
von: Sinha, Aarush, et al.
Veröffentlicht: (2026)
von: Sinha, Aarush, et al.
Veröffentlicht: (2026)
RADIANT: Retrieval AugmenteD entIty-context AligNmenT -- Introducing RAG-ability and Entity-Context Divergence
von: Rawte, Vipula, et al.
Veröffentlicht: (2025)
von: Rawte, Vipula, et al.
Veröffentlicht: (2025)
SEPSIS: I Can Catch Your Lies -- A New Paradigm for Deception Detection
von: Rani, Anku, et al.
Veröffentlicht: (2023)
von: Rani, Anku, et al.
Veröffentlicht: (2023)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2023)
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2023)
KnowledgePrompts: Exploring the Abilities of Large Language Models to Solve Proportional Analogies via Knowledge-Enhanced Prompting
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2024)
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2024)
A Comprehensive Dataset for Human vs. AI Generated Text Detection
von: Roy, Rajarshi, et al.
Veröffentlicht: (2025)
von: Roy, Rajarshi, et al.
Veröffentlicht: (2025)
SPINAL -- Scaling-law and Preference Integration in Neural Alignment Layers
von: Das, Arion, et al.
Veröffentlicht: (2026)
von: Das, Arion, et al.
Veröffentlicht: (2026)
Overview of Factify5WQA: Fact Verification through 5W Question-Answering
von: Suresh, Suryavardan, et al.
Veröffentlicht: (2024)
von: Suresh, Suryavardan, et al.
Veröffentlicht: (2024)
LLMsAgainstHate @ NLU of Devanagari Script Languages 2025: Hate Speech Detection and Target Identification in Devanagari Languages via Parameter Efficient Fine-Tuning of LLMs
von: Sidibomma, Rushendra, et al.
Veröffentlicht: (2024)
von: Sidibomma, Rushendra, et al.
Veröffentlicht: (2024)
ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models
von: Rawte, Vipula, et al.
Veröffentlicht: (2024)
von: Rawte, Vipula, et al.
Veröffentlicht: (2024)
Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs
von: Saha, Anusa, et al.
Veröffentlicht: (2026)
von: Saha, Anusa, et al.
Veröffentlicht: (2026)
Evidence-backed Fact Checking using RAG and Few-Shot In-Context Learning with LLMs
von: Singhal, Ronit, et al.
Veröffentlicht: (2024)
von: Singhal, Ronit, et al.
Veröffentlicht: (2024)
Counter Turing Test ($CT^2$): Investigating AI-Generated Text Detection for Hindi -- Ranking LLMs based on Hindi AI Detectability Index ($ADI_{hi}$)
von: Kavathekar, Ishan, et al.
Veröffentlicht: (2024)
von: Kavathekar, Ishan, et al.
Veröffentlicht: (2024)
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
von: Sheth, Paras, et al.
Veröffentlicht: (2024)
von: Sheth, Paras, et al.
Veröffentlicht: (2024)
MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models
von: Saha, Partha Pratim, et al.
Veröffentlicht: (2026)
von: Saha, Partha Pratim, et al.
Veröffentlicht: (2026)
MAAT: Multi-phase Adapter-Aware Targeted Unlearning
von: Yagnik, Suryash, et al.
Veröffentlicht: (2026)
von: Yagnik, Suryash, et al.
Veröffentlicht: (2026)
FETILDA: An Effective Framework For Fin-tuned Embeddings For Long Financial Text Documents
von: Xia, Bolun "Namir", et al.
Veröffentlicht: (2022)
von: Xia, Bolun "Namir", et al.
Veröffentlicht: (2022)
SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference
von: Chauhan, Anay, et al.
Veröffentlicht: (2026)
von: Chauhan, Anay, et al.
Veröffentlicht: (2026)
TRACEALIGN -- Tracing the Drift: Attributing Alignment Failures to Training-Time Belief Sources in LLMs
von: Das, Amitava, et al.
Veröffentlicht: (2025)
von: Das, Amitava, et al.
Veröffentlicht: (2025)
Assessing LLM Reliability on Temporally Recent Open-Domain Questions
von: Krishnappa, Pushwitha, et al.
Veröffentlicht: (2026)
von: Krishnappa, Pushwitha, et al.
Veröffentlicht: (2026)
Do Voters Get the Information They Want? Understanding Authentic Voter FAQs in the US and How to Improve for Informed Electoral Participation
von: Rawte, Vipula, et al.
Veröffentlicht: (2024)
von: Rawte, Vipula, et al.
Veröffentlicht: (2024)
PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training
von: Kumar, Harsh, et al.
Veröffentlicht: (2026)
von: Kumar, Harsh, et al.
Veröffentlicht: (2026)
AMBEDKAR-A Multi-level Bias Elimination through a Decoding Approach with Knowledge Augmentation for Robust Constitutional Alignment of Language Models
von: Mukhopadhyay, Snehasis, et al.
Veröffentlicht: (2025)
von: Mukhopadhyay, Snehasis, et al.
Veröffentlicht: (2025)
A Survey of AI-generated Text Forensic Systems: Detection, Attribution, and Characterization
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024)
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024)
Prompt Sensitivity and Answer Consistency of Small Open-Source Language Models for Clinical Question Answering in Low-Resource Healthcare
von: Hariprasad, Shravani
Veröffentlicht: (2026)
von: Hariprasad, Shravani
Veröffentlicht: (2026)
DeHate: A Stable Diffusion-based Multimodal Approach to Mitigate Hate Speech in Images
von: Dalal, Dwip, et al.
Veröffentlicht: (2025)
von: Dalal, Dwip, et al.
Veröffentlicht: (2025)
YINYANG-ALIGN: Benchmarking Contradictory Objectives and Proposing Multi-Objective Optimization based DPO for Text-to-Image Alignment
von: Das, Amitava, et al.
Veröffentlicht: (2025)
von: Das, Amitava, et al.
Veröffentlicht: (2025)
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context
von: Das, Nilanjana, et al.
Veröffentlicht: (2024)
von: Das, Nilanjana, et al.
Veröffentlicht: (2024)
Findings of the Counter Turing Test: AI-Generated Image Detection
von: Roy, Rajarshi, et al.
Veröffentlicht: (2026)
von: Roy, Rajarshi, et al.
Veröffentlicht: (2026)
AlignGuard-LoRA: Alignment-Preserving Fine-Tuning via Fisher-Guided Decomposition and Riemannian-Geodesic Collision Regularization
von: Das, Amitava, et al.
Veröffentlicht: (2025)
von: Das, Amitava, et al.
Veröffentlicht: (2025)
AdversariaL attacK sAfety aLIgnment(ALKALI): Safeguarding LLMs through GRACE: Geometric Representation-Aware Contrastive Enhancement- Introducing Adversarial Vulnerability Quality Index (AVQI)
von: Khanna, Danush, et al.
Veröffentlicht: (2025)
von: Khanna, Danush, et al.
Veröffentlicht: (2025)
Trust but Verify: Introducing DAVinCI -- A Framework for Dual Attribution and Verification in Claim Inference for Language Models
von: Rawte, Vipula, et al.
Veröffentlicht: (2026)
von: Rawte, Vipula, et al.
Veröffentlicht: (2026)
I Think, Therefore I Am Under-Qualified? A Benchmark for Evaluating Linguistic Shibboleth Detection in LLM Hiring Evaluations
von: Kharchenko, Julia, et al.
Veröffentlicht: (2025)
von: Kharchenko, Julia, et al.
Veröffentlicht: (2025)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
von: Chowdhury, Arijit Ghosh, et al.
Veröffentlicht: (2023)
von: Chowdhury, Arijit Ghosh, et al.
Veröffentlicht: (2023)
A Comprehensive Dataset for Human vs. AI Generated Image Detection
von: Roy, Rajarshi, et al.
Veröffentlicht: (2026)
von: Roy, Rajarshi, et al.
Veröffentlicht: (2026)
Breaking Language Barriers: A Question Answering Dataset for Hindi and Marathi
von: Sabane, Maithili, et al.
Veröffentlicht: (2023)
von: Sabane, Maithili, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
"Sorry, Come Again?" Prompting -- Enhancing Comprehension and Diminishing Hallucination with [PAUSE]-injected Optimal Paraphrasing
von: Rawte, Vipula, et al.
Veröffentlicht: (2024) -
A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models
von: Tonmoy, S. M Towhidul Islam, et al.
Veröffentlicht: (2024) -
Visual Hallucination: Definition, Quantification, and Prescriptive Remediations
von: Rani, Anku, et al.
Veröffentlicht: (2024) -
The What, Why, and How of Context Length Extension Techniques in Large Language Models -- A Detailed Survey
von: Pawar, Saurav, et al.
Veröffentlicht: (2024) -
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
von: Sinha, Aarush, et al.
Veröffentlicht: (2026)