Political Neutrality as Balanced Approval: A Large-Scale Human Evaluation of AI Responses
Fuente:
arXiv
Saved in:
| Main Authors: | Stray, Jonathan, Yang, David Zhai, Luo, Steven, Takagi, Miu Nicole, Chang, Serina |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Effects of Political Martyrdom on Election Results: The Assassination of Abe
by: Takagi, Miu Nicole
Published: (2023)
by: Takagi, Miu Nicole
Published: (2023)
ChatBench: From Static Benchmarks to Human-AI Evaluation
by: Chang, Serina, et al.
Published: (2025)
by: Chang, Serina, et al.
Published: (2025)
Political Neutrality in AI Is Impossible- But Here Is How to Approximate It
by: Fisher, Jillian, et al.
Published: (2025)
by: Fisher, Jillian, et al.
Published: (2025)
Accurate Measures of Vaccination and Concerns of Vaccine Holdouts from Web Search Logs
by: Chang, Serina, et al.
Published: (2023)
by: Chang, Serina, et al.
Published: (2023)
When Neutral Summaries are not that Neutral: Quantifying Political Neutrality in LLM-Generated News Summaries
by: Vijay, Supriti, et al.
Published: (2024)
by: Vijay, Supriti, et al.
Published: (2024)
Certified Safe: A Schematic for Approval Regulation of Frontier AI
by: Salvador, Cole
Published: (2024)
by: Salvador, Cole
Published: (2024)
Validated Hypotheses as a Lens for Human-Likeness Evaluation in AI Agents
by: Liu, Xuan, et al.
Published: (2026)
by: Liu, Xuan, et al.
Published: (2026)
An FDA for AI? Pitfalls and Plausibility of Approval Regulation for Frontier Artificial Intelligence
by: Carpenter, Daniel, et al.
Published: (2024)
by: Carpenter, Daniel, et al.
Published: (2024)
Generative AI User Experience: Developing Human--AI Epistemic Partnership
by: Zhai, Xiaoming
Published: (2026)
by: Zhai, Xiaoming
Published: (2026)
Human Trust in AI Search: A Large-Scale Experiment
by: Li, Haiwen, et al.
Published: (2025)
by: Li, Haiwen, et al.
Published: (2025)
Multimodal Political Bias Identification and Neutralization
by: Bernard, Cedric, et al.
Published: (2025)
by: Bernard, Cedric, et al.
Published: (2025)
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
by: Wu, Sirui, et al.
Published: (2025)
by: Wu, Sirui, et al.
Published: (2025)
Access Over Deception: Fighting Deceptive Patterns through Accessibility
by: Pellkvist, Tobias, et al.
Published: (2026)
by: Pellkvist, Tobias, et al.
Published: (2026)
Mutual Wanting in Human--AI Interaction: Empirical Evidence from Large-Scale Analysis of GPT Model Transitions
by: Shang, HaoYang, et al.
Published: (2025)
by: Shang, HaoYang, et al.
Published: (2025)
Charting the Future of AI-supported Science Education: A Human-Centered Vision
by: Zhai, Xiaoming, et al.
Published: (2026)
by: Zhai, Xiaoming, et al.
Published: (2026)
AI Ethics by Design: Implementing Customizable Guardrails for Responsible AI Development
by: Šekrst, Kristina, et al.
Published: (2024)
by: Šekrst, Kristina, et al.
Published: (2024)
The 2020s Political Economy of Machine Translation
by: Weber, Steven
Published: (2020)
by: Weber, Steven
Published: (2020)
Investigating Political and Demographic Associations in Large Language Models Through Moral Foundations Theory
by: Smith-Vaniz, Nicole, et al.
Published: (2025)
by: Smith-Vaniz, Nicole, et al.
Published: (2025)
A Framework for Human-AI Q-Matrix Refinement: A NeuralCDM Evaluation
by: Zhang, Ying, et al.
Published: (2026)
by: Zhang, Ying, et al.
Published: (2026)
When AI Gives Advice: Evaluating AI and Human Responses to Online Advice-Seeking for Well-Being
by: Kumar, Harsh, et al.
Published: (2025)
by: Kumar, Harsh, et al.
Published: (2025)
Responsible AI Governance: A Response to UN Interim Report on Governing AI for Humanity
by: Kiden, Sarah, et al.
Published: (2024)
by: Kiden, Sarah, et al.
Published: (2024)
Neutralizing the Narrative: AI-Powered Debiasing of Online News Articles
by: Kuo, Chen Wei, et al.
Published: (2025)
by: Kuo, Chen Wei, et al.
Published: (2025)
Shaping the Future of Social Media with Middleware
by: Hogg, Luke, et al.
Published: (2024)
by: Hogg, Luke, et al.
Published: (2024)
AI4CAREER: Responsible AI for STEM Career Development at Scale in K-16 Education
by: Chawla, Sugana, et al.
Published: (2026)
by: Chawla, Sugana, et al.
Published: (2026)
Human Experts' Evaluation of Generative AI for Contextualizing STEAM Education in the Global South
by: Nyaaba, Matthew, et al.
Published: (2025)
by: Nyaaba, Matthew, et al.
Published: (2025)
Measuring Political Preferences in AI Systems: An Integrative Approach
by: Rozado, David
Published: (2025)
by: Rozado, David
Published: (2025)
Human-AI Collaborative Inductive Thematic Analysis: AI Guided Analysis and Human Interpretive Authority
by: Nyaaba, Matthew, et al.
Published: (2026)
by: Nyaaba, Matthew, et al.
Published: (2026)
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
by: Liu, Yiran, et al.
Published: (2024)
by: Liu, Yiran, et al.
Published: (2024)
Responsible Evaluation of AI for Mental Health
by: Arnaout, Hiba, et al.
Published: (2026)
by: Arnaout, Hiba, et al.
Published: (2026)
Speciesism in AI: Evaluating Discrimination Against Animals in Large Language Models
by: Jotautaitė, Monika, et al.
Published: (2025)
by: Jotautaitė, Monika, et al.
Published: (2025)
Accept or Deny? Evaluating LLM Fairness and Performance in Loan Approval across Table-to-Text Serialization Approaches
by: Azime, Israel Abebe, et al.
Published: (2025)
by: Azime, Israel Abebe, et al.
Published: (2025)
The Rise of AI Search: Implications for Information Markets and Human Judgement at Scale
by: Aral, Sinan, et al.
Published: (2026)
by: Aral, Sinan, et al.
Published: (2026)
General Scales Unlock AI Evaluation with Explanatory and Predictive Power
by: Zhou, Lexin, et al.
Published: (2025)
by: Zhou, Lexin, et al.
Published: (2025)
Evaluating Large Language Models Against Human Annotators in Latent Content Analysis: Sentiment, Political Leaning, Emotional Intensity, and Sarcasm
by: Bojic, Ljubisa, et al.
Published: (2025)
by: Bojic, Ljubisa, et al.
Published: (2025)
Human-Centered Design for AI-based Automatically Generated Assessment Reports: A Systematic Review
by: Latif, Ehsan, et al.
Published: (2024)
by: Latif, Ehsan, et al.
Published: (2024)
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
by: Davani, Aida, et al.
Published: (2025)
by: Davani, Aida, et al.
Published: (2025)
Which Humans? Inclusivity and Representation in Human-Centered AI
by: Mihalcea, Rada, et al.
Published: (2025)
by: Mihalcea, Rada, et al.
Published: (2025)
Generative Large Language Models for Knowledge Representation: A Systematic Review of Concept Map Generation
by: Zhai, Xiaoming
Published: (2025)
by: Zhai, Xiaoming
Published: (2025)
Evaluating Digital Inclusiveness of Digital Agri-Food Tools Using Large Language Models: A Comparative Analysis Between Human and AI-Based Evaluations
by: Pewinya, Githma, et al.
Published: (2026)
by: Pewinya, Githma, et al.
Published: (2026)
Assessing Human Rights Risks in AI: A Framework for Model Evaluation
by: Raman, Vyoma, et al.
Published: (2025)
by: Raman, Vyoma, et al.
Published: (2025)
Similar Items
-
The Effects of Political Martyrdom on Election Results: The Assassination of Abe
by: Takagi, Miu Nicole
Published: (2023) -
ChatBench: From Static Benchmarks to Human-AI Evaluation
by: Chang, Serina, et al.
Published: (2025) -
Political Neutrality in AI Is Impossible- But Here Is How to Approximate It
by: Fisher, Jillian, et al.
Published: (2025) -
Accurate Measures of Vaccination and Concerns of Vaccine Holdouts from Web Search Logs
by: Chang, Serina, et al.
Published: (2023) -
When Neutral Summaries are not that Neutral: Quantifying Political Neutrality in LLM-Generated News Summaries
by: Vijay, Supriti, et al.
Published: (2024)