On the Relationship between Truth and Political Bias in Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Fulay, Suyash, Brannon, William, Mohanty, Shrestha, Overney, Cassandra, Poole-Dayan, Elinor, Roy, Deb, Kabbara, Jad |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2024)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2024)
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
di: Brannon, William, et al.
Pubblicazione: (2023)
di: Brannon, William, et al.
Pubblicazione: (2023)
Bridging Context Gaps: Enhancing Comprehension in Long-Form Social Conversations Through Contextualized Excerpts
di: Mohanty, Shrestha, et al.
Pubblicazione: (2024)
di: Mohanty, Shrestha, et al.
Pubblicazione: (2024)
Computational Analysis of Conversation Dynamics through Participant Responsivity
di: Hughes, Margaret, et al.
Pubblicazione: (2025)
di: Hughes, Margaret, et al.
Pubblicazione: (2025)
AI and Collective Decisions: Strengthening Legitimacy and Losers' Consent
di: Fulay, Suyash, et al.
Pubblicazione: (2026)
di: Fulay, Suyash, et al.
Pubblicazione: (2026)
The Empty Chair: Using LLMs to Raise Missing Perspectives in Policy Deliberations
di: Fulay, Suyash, et al.
Pubblicazione: (2025)
di: Fulay, Suyash, et al.
Pubblicazione: (2025)
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits
di: Jiang, Hang, et al.
Pubblicazione: (2023)
di: Jiang, Hang, et al.
Pubblicazione: (2023)
From Delegates to Trustees: How Optimizing for Long-Term Interests Shapes Bias and Alignment in LLM
di: Fulay, Suyash, et al.
Pubblicazione: (2025)
di: Fulay, Suyash, et al.
Pubblicazione: (2025)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
Outgroup Animosity Has Risen for Politicians, Journalists, and a Sample of Partisan Users on Twitter and Reddit
di: Fulay, Suyash, et al.
Pubblicazione: (2023)
di: Fulay, Suyash, et al.
Pubblicazione: (2023)
Agora: Teaching the Skill of Consensus-Finding with AI Personas Grounded in Human Voice
di: Ravi, Prerna, et al.
Pubblicazione: (2026)
di: Ravi, Prerna, et al.
Pubblicazione: (2026)
Examining the Influence of Political Bias on Large Language Model Performance in Stance Classification
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2024)
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2024)
Applying Large Language Models to Characterize Public Narratives
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
Assessing Political Bias in Large Language Models
di: Rettenberger, Luca, et al.
Pubblicazione: (2024)
di: Rettenberger, Luca, et al.
Pubblicazione: (2024)
Fine-Grained Bias Detection in LLM: Enhancing detection mechanisms for nuanced biases
di: Mohanty, Suvendu
Pubblicazione: (2025)
di: Mohanty, Suvendu
Pubblicazione: (2025)
Data Authenticity, Consent, & Provenance for AI are all broken: what will it take to fix them?
di: Longpre, Shayne, et al.
Pubblicazione: (2024)
di: Longpre, Shayne, et al.
Pubblicazione: (2024)
To Tell The Truth: Language of Deception and Language Models
di: Hazra, Sanchaita, et al.
Pubblicazione: (2023)
di: Hazra, Sanchaita, et al.
Pubblicazione: (2023)
Toxicity Begets Toxicity: Unraveling Conversational Chains in Political Podcasts
di: Rizwan, Naquee, et al.
Pubblicazione: (2025)
di: Rizwan, Naquee, et al.
Pubblicazione: (2025)
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said
di: Bang, Yejin, et al.
Pubblicazione: (2024)
di: Bang, Yejin, et al.
Pubblicazione: (2024)
Persuasiveness and Bias in LLM: Investigating the Impact of Persuasiveness and Reinforcement of Bias in Language Models
di: Roy, Saumya
Pubblicazione: (2025)
di: Roy, Saumya
Pubblicazione: (2025)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
di: Barkett, Emilio, et al.
Pubblicazione: (2025)
di: Barkett, Emilio, et al.
Pubblicazione: (2025)
Benchmarking Overton Pluralism in LLMs
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
Framing Political Bias in Multilingual LLMs Across Pakistani Languages
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
di: Chatrath, Veronica, et al.
Pubblicazione: (2024)
di: Chatrath, Veronica, et al.
Pubblicazione: (2024)
Truth Forest: Toward Multi-Scale Truthfulness in Large Language Models through Intervention without Tuning
di: Chen, Zhongzhi, et al.
Pubblicazione: (2023)
di: Chen, Zhongzhi, et al.
Pubblicazione: (2023)
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
di: Zhang, Shaolei, et al.
Pubblicazione: (2024)
di: Zhang, Shaolei, et al.
Pubblicazione: (2024)
High Risk of Political Bias in Black Box Emotion Inference Models
di: Plisiecki, Hubert, et al.
Pubblicazione: (2024)
di: Plisiecki, Hubert, et al.
Pubblicazione: (2024)
Benchmarking Gender and Political Bias in Large Language Models
di: Yang, Jinrui, et al.
Pubblicazione: (2025)
di: Yang, Jinrui, et al.
Pubblicazione: (2025)
Truth Knows No Language: Evaluating Truthfulness Beyond English
di: Figueras, Blanca Calvo, et al.
Pubblicazione: (2025)
di: Figueras, Blanca Calvo, et al.
Pubblicazione: (2025)
Towards Reliable Truth-Aligned Uncertainty Estimation in Large Language Models
di: Srey, Ponhvoan, et al.
Pubblicazione: (2026)
di: Srey, Ponhvoan, et al.
Pubblicazione: (2026)
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
Personas as a Way to Model Truthfulness in Language Models
di: Joshi, Nitish, et al.
Pubblicazione: (2023)
di: Joshi, Nitish, et al.
Pubblicazione: (2023)
Political-LLM: Large Language Models in Political Science
di: Li, Lincan, et al.
Pubblicazione: (2024)
di: Li, Lincan, et al.
Pubblicazione: (2024)
Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks
di: Schroeder, Hope, et al.
Pubblicazione: (2025)
di: Schroeder, Hope, et al.
Pubblicazione: (2025)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
di: Jenny, David F., et al.
Pubblicazione: (2023)
di: Jenny, David F., et al.
Pubblicazione: (2023)
Uncovering Political Bias in Large Language Models using Parliamentary Voting Records
di: Chen, Jieying, et al.
Pubblicazione: (2026)
di: Chen, Jieying, et al.
Pubblicazione: (2026)
Steering Towards Fairness: Mitigating Political Bias in LLMs
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
di: Sakhawat, Adib, et al.
Pubblicazione: (2026)
di: Sakhawat, Adib, et al.
Pubblicazione: (2026)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
di: Wijesiriwardene, Thilini, et al.
Pubblicazione: (2023)
di: Wijesiriwardene, Thilini, et al.
Pubblicazione: (2023)
Documenti analoghi
-
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2024) -
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025) -
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
di: Brannon, William, et al.
Pubblicazione: (2023) -
Bridging Context Gaps: Enhancing Comprehension in Long-Form Social Conversations Through Contextualized Excerpts
di: Mohanty, Shrestha, et al.
Pubblicazione: (2024) -
Computational Analysis of Conversation Dynamics through Participant Responsivity
di: Hughes, Margaret, et al.
Pubblicazione: (2025)