On the Relationship between Truth and Political Bias in Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fulay, Suyash, Brannon, William, Mohanty, Shrestha, Overney, Cassandra, Poole-Dayan, Elinor, Roy, Deb, Kabbara, Jad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2024)
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2024)
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025)
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025)
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
von: Brannon, William, et al.
Veröffentlicht: (2023)
von: Brannon, William, et al.
Veröffentlicht: (2023)
Bridging Context Gaps: Enhancing Comprehension in Long-Form Social Conversations Through Contextualized Excerpts
von: Mohanty, Shrestha, et al.
Veröffentlicht: (2024)
von: Mohanty, Shrestha, et al.
Veröffentlicht: (2024)
Computational Analysis of Conversation Dynamics through Participant Responsivity
von: Hughes, Margaret, et al.
Veröffentlicht: (2025)
von: Hughes, Margaret, et al.
Veröffentlicht: (2025)
AI and Collective Decisions: Strengthening Legitimacy and Losers' Consent
von: Fulay, Suyash, et al.
Veröffentlicht: (2026)
von: Fulay, Suyash, et al.
Veröffentlicht: (2026)
The Empty Chair: Using LLMs to Raise Missing Perspectives in Policy Deliberations
von: Fulay, Suyash, et al.
Veröffentlicht: (2025)
von: Fulay, Suyash, et al.
Veröffentlicht: (2025)
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits
von: Jiang, Hang, et al.
Veröffentlicht: (2023)
von: Jiang, Hang, et al.
Veröffentlicht: (2023)
From Delegates to Trustees: How Optimizing for Long-Term Interests Shapes Bias and Alignment in LLM
von: Fulay, Suyash, et al.
Veröffentlicht: (2025)
von: Fulay, Suyash, et al.
Veröffentlicht: (2025)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
Outgroup Animosity Has Risen for Politicians, Journalists, and a Sample of Partisan Users on Twitter and Reddit
von: Fulay, Suyash, et al.
Veröffentlicht: (2023)
von: Fulay, Suyash, et al.
Veröffentlicht: (2023)
Agora: Teaching the Skill of Consensus-Finding with AI Personas Grounded in Human Voice
von: Ravi, Prerna, et al.
Veröffentlicht: (2026)
von: Ravi, Prerna, et al.
Veröffentlicht: (2026)
Examining the Influence of Political Bias on Large Language Model Performance in Stance Classification
von: Ng, Lynnette Hui Xian, et al.
Veröffentlicht: (2024)
von: Ng, Lynnette Hui Xian, et al.
Veröffentlicht: (2024)
Applying Large Language Models to Characterize Public Narratives
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025)
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025)
Assessing Political Bias in Large Language Models
von: Rettenberger, Luca, et al.
Veröffentlicht: (2024)
von: Rettenberger, Luca, et al.
Veröffentlicht: (2024)
Fine-Grained Bias Detection in LLM: Enhancing detection mechanisms for nuanced biases
von: Mohanty, Suvendu
Veröffentlicht: (2025)
von: Mohanty, Suvendu
Veröffentlicht: (2025)
Data Authenticity, Consent, & Provenance for AI are all broken: what will it take to fix them?
von: Longpre, Shayne, et al.
Veröffentlicht: (2024)
von: Longpre, Shayne, et al.
Veröffentlicht: (2024)
To Tell The Truth: Language of Deception and Language Models
von: Hazra, Sanchaita, et al.
Veröffentlicht: (2023)
von: Hazra, Sanchaita, et al.
Veröffentlicht: (2023)
Toxicity Begets Toxicity: Unraveling Conversational Chains in Political Podcasts
von: Rizwan, Naquee, et al.
Veröffentlicht: (2025)
von: Rizwan, Naquee, et al.
Veröffentlicht: (2025)
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said
von: Bang, Yejin, et al.
Veröffentlicht: (2024)
von: Bang, Yejin, et al.
Veröffentlicht: (2024)
Persuasiveness and Bias in LLM: Investigating the Impact of Persuasiveness and Reinforcement of Bias in Language Models
von: Roy, Saumya
Veröffentlicht: (2025)
von: Roy, Saumya
Veröffentlicht: (2025)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
Benchmarking Overton Pluralism in LLMs
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025)
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025)
Framing Political Bias in Multilingual LLMs Across Pakistani Languages
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
von: Chatrath, Veronica, et al.
Veröffentlicht: (2024)
von: Chatrath, Veronica, et al.
Veröffentlicht: (2024)
Truth Forest: Toward Multi-Scale Truthfulness in Large Language Models through Intervention without Tuning
von: Chen, Zhongzhi, et al.
Veröffentlicht: (2023)
von: Chen, Zhongzhi, et al.
Veröffentlicht: (2023)
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
von: Zhang, Shaolei, et al.
Veröffentlicht: (2024)
von: Zhang, Shaolei, et al.
Veröffentlicht: (2024)
High Risk of Political Bias in Black Box Emotion Inference Models
von: Plisiecki, Hubert, et al.
Veröffentlicht: (2024)
von: Plisiecki, Hubert, et al.
Veröffentlicht: (2024)
Benchmarking Gender and Political Bias in Large Language Models
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
Truth Knows No Language: Evaluating Truthfulness Beyond English
von: Figueras, Blanca Calvo, et al.
Veröffentlicht: (2025)
von: Figueras, Blanca Calvo, et al.
Veröffentlicht: (2025)
Towards Reliable Truth-Aligned Uncertainty Estimation in Large Language Models
von: Srey, Ponhvoan, et al.
Veröffentlicht: (2026)
von: Srey, Ponhvoan, et al.
Veröffentlicht: (2026)
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
von: Wehnert, Sabine, et al.
Veröffentlicht: (2025)
von: Wehnert, Sabine, et al.
Veröffentlicht: (2025)
Personas as a Way to Model Truthfulness in Language Models
von: Joshi, Nitish, et al.
Veröffentlicht: (2023)
von: Joshi, Nitish, et al.
Veröffentlicht: (2023)
Political-LLM: Large Language Models in Political Science
von: Li, Lincan, et al.
Veröffentlicht: (2024)
von: Li, Lincan, et al.
Veröffentlicht: (2024)
Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks
von: Schroeder, Hope, et al.
Veröffentlicht: (2025)
von: Schroeder, Hope, et al.
Veröffentlicht: (2025)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
von: Jenny, David F., et al.
Veröffentlicht: (2023)
von: Jenny, David F., et al.
Veröffentlicht: (2023)
Uncovering Political Bias in Large Language Models using Parliamentary Voting Records
von: Chen, Jieying, et al.
Veröffentlicht: (2026)
von: Chen, Jieying, et al.
Veröffentlicht: (2026)
Steering Towards Fairness: Mitigating Political Bias in LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2023)
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2024) -
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
von: Poole-Dayan, Elinor, et al.
Veröffentlicht: (2025) -
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
von: Brannon, William, et al.
Veröffentlicht: (2023) -
Bridging Context Gaps: Enhancing Comprehension in Long-Form Social Conversations Through Contextualized Excerpts
von: Mohanty, Shrestha, et al.
Veröffentlicht: (2024) -
Computational Analysis of Conversation Dynamics through Participant Responsivity
von: Hughes, Margaret, et al.
Veröffentlicht: (2025)