Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
Fuente:
arXiv
Saved in:
| Main Authors: | Chatrath, Veronica, Lotif, Marcelo, Raza, Shaina |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
by: Raza, Shaina, et al.
Published: (2023)
by: Raza, Shaina, et al.
Published: (2023)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
by: Garg, Muskan, et al.
Published: (2024)
by: Garg, Muskan, et al.
Published: (2024)
BEADs: Bias Evaluation Across Domains
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
by: Piot, Paloma, et al.
Published: (2025)
by: Piot, Paloma, et al.
Published: (2025)
Can LLMs Automate Fact-Checking Article Writing?
by: Sahnan, Dhruv, et al.
Published: (2025)
by: Sahnan, Dhruv, et al.
Published: (2025)
TruthStance: An Annotated Dataset of Conversations on Truth Social
by: Ameen, Fathima, et al.
Published: (2026)
by: Ameen, Fathima, et al.
Published: (2026)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Scaling Truth: The Confidence Paradox in AI Fact-Checking
by: Qazi, Ihsan A., et al.
Published: (2025)
by: Qazi, Ihsan A., et al.
Published: (2025)
FairSense-AI: Responsible AI Meets Sustainability
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
by: Khatun, Aisha, et al.
Published: (2024)
by: Khatun, Aisha, et al.
Published: (2024)
Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
Fact or Fiction? Improving Fact Verification with Knowledge Graphs through Simplified Subgraph Retrievals
by: Opsahl, Tobias A.
Published: (2024)
by: Opsahl, Tobias A.
Published: (2024)
The Rise of Small Language Models in Healthcare: A Comprehensive Survey
by: Garg, Muskan, et al.
Published: (2025)
by: Garg, Muskan, et al.
Published: (2025)
SelectLLM: Can LLMs Select Important Instructions to Annotate?
by: Parkar, Ritik Sachin, et al.
Published: (2024)
by: Parkar, Ritik Sachin, et al.
Published: (2024)
Improving Reliability and Explainability of Medical Question Answering through Atomic Fact Checking in Retrieval-Augmented LLMs
by: Vladika, Juraj, et al.
Published: (2025)
by: Vladika, Juraj, et al.
Published: (2025)
VeRO: An Evaluation Harness for Agents to Optimize Agents
by: Ursekar, Varun, et al.
Published: (2026)
by: Ursekar, Varun, et al.
Published: (2026)
On the Relationship between Truth and Political Bias in Language Models
by: Fulay, Suyash, et al.
Published: (2024)
by: Fulay, Suyash, et al.
Published: (2024)
Testing the Limits of Truth Directions in LLMs
by: Poulis, Angelos, et al.
Published: (2026)
by: Poulis, Angelos, et al.
Published: (2026)
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators
by: Bajpai, Prasoon, et al.
Published: (2024)
by: Bajpai, Prasoon, et al.
Published: (2024)
From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?
by: Amin, Hasan, et al.
Published: (2026)
by: Amin, Hasan, et al.
Published: (2026)
How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs
by: Adarsh, Shivam, et al.
Published: (2026)
by: Adarsh, Shivam, et al.
Published: (2026)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
by: Raval, Ananya, et al.
Published: (2025)
by: Raval, Ananya, et al.
Published: (2025)
FakeWatch: A Framework for Detecting Fake News to Ensure Credible Elections
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Truth is Universal: Robust Detection of Lies in LLMs
by: Bürger, Lennart, et al.
Published: (2024)
by: Bürger, Lennart, et al.
Published: (2024)
Can LLMs Detect Their Confabulations? Estimating Reliability in Uncertainty-Aware Language Models
by: Zhou, Tianyi, et al.
Published: (2025)
by: Zhou, Tianyi, et al.
Published: (2025)
Can LLMs Reliably Simulate Real Students' Abilities in Mathematics and Reading Comprehension?
by: Srivatsa, KV Aditya, et al.
Published: (2025)
by: Srivatsa, KV Aditya, et al.
Published: (2025)
Are Chatbots Reliable Text Annotators? Sometimes
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)
Towards Reliable Truth-Aligned Uncertainty Estimation in Large Language Models
by: Srey, Ponhvoan, et al.
Published: (2026)
by: Srey, Ponhvoan, et al.
Published: (2026)
Can Community Notes Replace Professional Fact-Checkers?
by: Borenstein, Nadav, et al.
Published: (2025)
by: Borenstein, Nadav, et al.
Published: (2025)
FASTTRACK: Fast and Accurate Fact Tracing for LLMs
by: Chen, Si, et al.
Published: (2024)
by: Chen, Si, et al.
Published: (2024)
How LLMs Fail to Support Fact-Checking
by: Proma, Adiba Mahbub, et al.
Published: (2025)
by: Proma, Adiba Mahbub, et al.
Published: (2025)
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
by: Karia, Rushang, et al.
Published: (2024)
by: Karia, Rushang, et al.
Published: (2024)
Academic case reports lack diversity: Assessing the presence and diversity of sociodemographic and behavioral factors related to Post COVID-19 Condition
by: Florez, Juan Andres Medina, et al.
Published: (2025)
by: Florez, Juan Andres Medina, et al.
Published: (2025)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
by: Wei, Zhepei, et al.
Published: (2025)
by: Wei, Zhepei, et al.
Published: (2025)
AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM Annotators
by: Ni, Jingwei, et al.
Published: (2024)
by: Ni, Jingwei, et al.
Published: (2024)
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
by: Munir, Sheza, et al.
Published: (2026)
by: Munir, Sheza, et al.
Published: (2026)
Just as Humans Need Vaccines, So Do Models: Model Immunization to Combat Falsehoods
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Debating with More Persuasive LLMs Leads to More Truthful Answers
by: Khan, Akbir, et al.
Published: (2024)
by: Khan, Akbir, et al.
Published: (2024)
Similar Items
-
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
by: Raza, Shaina, et al.
Published: (2023) -
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
by: Raza, Shaina, et al.
Published: (2024) -
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
by: Raza, Shaina, et al.
Published: (2024) -
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
by: Garg, Muskan, et al.
Published: (2024) -
BEADs: Bias Evaluation Across Domains
by: Raza, Shaina, et al.
Published: (2024)