Scaling Truth: The Confidence Paradox in AI Fact-Checking
Fuente:
arXiv
Saved in:
| Main Authors: | Qazi, Ihsan A., Khan, Zohaib, Ghani, Abdullah, Raza, Agha A., Qazi, Zafar A., Sajjad, Wassay, Ali, Ayesha, Javaid, Asher, Sohail, Muhammad Abdullah, Azeemi, Abdul H. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Model-Driven Data Pruning Enables Efficient Active Learning
by: Azeemi, Abdul Hameed, et al.
Published: (2024)
by: Azeemi, Abdul Hameed, et al.
Published: (2024)
To Label or Not to Label: Hybrid Active Learning for Neural Machine Translation
by: Azeemi, Abdul Hameed, et al.
Published: (2024)
by: Azeemi, Abdul Hameed, et al.
Published: (2024)
Generalists vs. Specialists: Evaluating Large Language Models for Urdu
by: Arif, Samee, et al.
Published: (2024)
by: Arif, Samee, et al.
Published: (2024)
Rethinking Image Compression on the Web with Generative AI
by: Hassan, Shayan Ali, et al.
Published: (2024)
by: Hassan, Shayan Ali, et al.
Published: (2024)
TweakLLM: A Routing Architecture for Dynamic Tailoring of Cached Responses
by: Cheema, Muhammad Taha, et al.
Published: (2025)
by: Cheema, Muhammad Taha, et al.
Published: (2025)
PakBBQ: A Culturally Adapted Bias Benchmark for QA
by: Hashmat, Abdullah, et al.
Published: (2025)
by: Hashmat, Abdullah, et al.
Published: (2025)
The Fellowship of the LLMs: Multi-Model Workflows for Synthetic Preference Optimization Dataset Generation
by: Arif, Samee, et al.
Published: (2024)
by: Arif, Samee, et al.
Published: (2024)
Investigating Misinformation Dissemination on Social Media in Pakistan
by: Haroon, Danyal, et al.
Published: (2021)
by: Haroon, Danyal, et al.
Published: (2021)
Beyond Uniform Query Distribution: Key-Driven Grouped Query Attention
by: Khan, Zohaib, et al.
Published: (2024)
by: Khan, Zohaib, et al.
Published: (2024)
NeuGen: Amplifying the 'Neural' in Neural Radiance Fields for Domain Generalization
by: Qazi, Ahmed, et al.
Published: (2025)
by: Qazi, Ahmed, et al.
Published: (2025)
With a Grain of SALT: Are LLMs Fair Across Social Dimensions?
by: Arif, Samee, et al.
Published: (2024)
by: Arif, Samee, et al.
Published: (2024)
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
by: Ahmad, Sarfraz, et al.
Published: (2025)
by: Ahmad, Sarfraz, et al.
Published: (2025)
Vanadium-Engineered Co2NiSe4 Nanomaterial: Coupled Thermoelectric, Piezoelectric, and Electronic Optimization via DFT+U for Advanced Energy Applications
by: Riaz, Ayesha, et al.
Published: (2025)
by: Riaz, Ayesha, et al.
Published: (2025)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
by: Chatrath, Veronica, et al.
Published: (2024)
by: Chatrath, Veronica, et al.
Published: (2024)
Factually: Exploring Wearable Fact-Checking for Augmented Truth Discernment
by: Gupta, Chitralekha, et al.
Published: (2025)
by: Gupta, Chitralekha, et al.
Published: (2025)
Profiling-Driven Adaptive Distributed Transformer Inference on Embedded Edge Deployment
by: Qazi, Muhammad Azlan, et al.
Published: (2026)
by: Qazi, Muhammad Azlan, et al.
Published: (2026)
PRISM: Distributed Inference for Foundation Models at Edge
by: Qazi, Muhammad Azlan, et al.
Published: (2025)
by: Qazi, Muhammad Azlan, et al.
Published: (2025)
When retrieval outperforms generation: Dense evidence retrieval for scalable fake news detection
by: Qazi, Alamgir Munir, et al.
Published: (2025)
by: Qazi, Alamgir Munir, et al.
Published: (2025)
Hallucination to Truth: A Review of Fact-Checking and Factuality Evaluation in Large Language Models
by: Rahman, Subhey Sadi, et al.
Published: (2025)
by: Rahman, Subhey Sadi, et al.
Published: (2025)
NeuRN: Neuro-inspired Domain Generalization for Image Classification
by: Jalil, Hamd, et al.
Published: (2025)
by: Jalil, Hamd, et al.
Published: (2025)
Mice to Machines: Neural Representations from Visual Cortex for Domain Generalization
by: Qazi, Ahmed, et al.
Published: (2025)
by: Qazi, Ahmed, et al.
Published: (2025)
Toward an AI-Native Internet: Rethinking the Web Architecture for Semantic Retrieval
by: Bilal, Muhammad, et al.
Published: (2025)
by: Bilal, Muhammad, et al.
Published: (2025)
CheckMate: LLM-Powered Approximate Intermittent Computing
by: Sayyid-Ali, Abdur-Rahman Ibrahim, et al.
Published: (2024)
by: Sayyid-Ali, Abdur-Rahman Ibrahim, et al.
Published: (2024)
Truth with a Twist: The Rhetoric of Persuasion in Professional vs. Community-Authored Fact-Checks
by: Razuvayevskaya, Olesya, et al.
Published: (2026)
by: Razuvayevskaya, Olesya, et al.
Published: (2026)
Mitigating Hallucinations in Healthcare LLMs with Granular Fact-Checking and Domain-Specific Adaptation
by: Zeba, Musarrat, et al.
Published: (2025)
by: Zeba, Musarrat, et al.
Published: (2025)
AnimalFormer: Multimodal Vision Framework for Behavior-based Precision Livestock Farming
by: Qazi, Ahmed, et al.
Published: (2024)
by: Qazi, Ahmed, et al.
Published: (2024)
Quran-MD: A Fine-Grained Multilingual Multimodal Dataset of the Quran
by: Salman, Muhammad Umar, et al.
Published: (2026)
by: Salman, Muhammad Umar, et al.
Published: (2026)
Decoding User Concerns in AI Health Chatbots: An Exploration of Security and Privacy in App Reviews
by: Hassan, Muhammad, et al.
Published: (2025)
by: Hassan, Muhammad, et al.
Published: (2025)
Introducing SDICE: An Index for Assessing Diversity of Synthetic Medical Datasets
by: Alam, Mohammed Talha, et al.
Published: (2024)
by: Alam, Mohammed Talha, et al.
Published: (2024)
PixelConfig: Longitudinal Measurement and Reverse-Engineering of Meta Pixel Configurations
by: Ghani, Abdullah, et al.
Published: (2026)
by: Ghani, Abdullah, et al.
Published: (2026)
ThumbnailTruth: A Multi-Modal LLM Approach for Detecting Misleading YouTube Thumbnails Across Diverse Cultural Settings
by: Naveed, Wajiha, et al.
Published: (2025)
by: Naveed, Wajiha, et al.
Published: (2025)
GPT-generated Text Detection: Benchmark Dataset and Tensor-based Detection Method
by: Qazi, Zubair, et al.
Published: (2024)
by: Qazi, Zubair, et al.
Published: (2024)
Deciphering the Underserved: Benchmarking LLM OCR for Low-Resource Scripts
by: Sohail, Muhammad Abdullah, et al.
Published: (2024)
by: Sohail, Muhammad Abdullah, et al.
Published: (2024)
A Data-Free Analytical Quantization Scheme for Deep Learning Models
by: Luqman, Ahmed, et al.
Published: (2024)
by: Luqman, Ahmed, et al.
Published: (2024)
FactSim: Fact-Checking for Opinion Summarization
by: Anghinoni, Leandro, et al.
Published: (2026)
by: Anghinoni, Leandro, et al.
Published: (2026)
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
by: Russo, Daniel, et al.
Published: (2024)
by: Russo, Daniel, et al.
Published: (2024)
A Phenomenological Study of Semileptonic $B^+$ and $B_s^0$ Decays into Axial-Vector Mesons $\big(D_1(2420),\, D_1^\prime(2430),\, D_{s1}(2460),\, \text{and } D_{s1}^\prime(2536)\big)$ within the Standard Model
by: Khan, Rana, et al.
Published: (2026)
by: Khan, Rana, et al.
Published: (2026)
(Fact) Check Your Bias
by: Bakke, Eivind Morris, et al.
Published: (2025)
by: Bakke, Eivind Morris, et al.
Published: (2025)
Non Asymptotic Mixing Time Analysis of Non-Reversible Markov Chains
by: Naeem, Muhammad Abdullah
Published: (2025)
by: Naeem, Muhammad Abdullah
Published: (2025)
Semantic Caching for Improving Web Affordability
by: Akbar, Hafsa, et al.
Published: (2025)
by: Akbar, Hafsa, et al.
Published: (2025)
Similar Items
-
Language Model-Driven Data Pruning Enables Efficient Active Learning
by: Azeemi, Abdul Hameed, et al.
Published: (2024) -
To Label or Not to Label: Hybrid Active Learning for Neural Machine Translation
by: Azeemi, Abdul Hameed, et al.
Published: (2024) -
Generalists vs. Specialists: Evaluating Large Language Models for Urdu
by: Arif, Samee, et al.
Published: (2024) -
Rethinking Image Compression on the Web with Generative AI
by: Hassan, Shayan Ali, et al.
Published: (2024) -
TweakLLM: A Routing Architecture for Dynamic Tailoring of Cached Responses
by: Cheema, Muhammad Taha, et al.
Published: (2025)