Pro-AI Bias in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Trabelsi, Benaya, Shaki, Jonathan, Kraus, Sarit |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Revealing Hidden Bias in AI: Lessons from Large Language Models
by: Beatty, Django, et al.
Published: (2024)
by: Beatty, Django, et al.
Published: (2024)
Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations
by: Wong, Qi Han
Published: (2026)
by: Wong, Qi Han
Published: (2026)
Gender and Race Bias in Consumer Product Recommendations by Large Language Models
by: Xu, Ke, et al.
Published: (2026)
by: Xu, Ke, et al.
Published: (2026)
Big AI is accelerating the metacrisis: What can we do?
by: Bird, Steven
Published: (2025)
by: Bird, Steven
Published: (2025)
AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains
by: Shintani, Seine A.
Published: (2026)
by: Shintani, Seine A.
Published: (2026)
Growing a Tail: Increasing Output Diversity in Large Language Models
by: Shur-Ofry, Michal, et al.
Published: (2024)
by: Shur-Ofry, Michal, et al.
Published: (2024)
Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
by: Hossain, Ariyan, et al.
Published: (2025)
by: Hossain, Ariyan, et al.
Published: (2025)
APPSI-139: A Parallel Corpus of English Application Privacy Policy Summarization and Interpretation
by: Zhu, Pengyun, et al.
Published: (2026)
by: Zhu, Pengyun, et al.
Published: (2026)
WSC+: Enhancing The Winograd Schema Challenge Using Tree-of-Experts
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
Leveraging Multi-Source Textural UGC for Neighbourhood Housing Quality Assessment: A GPT-Enhanced Framework
by: Hong, Qiyuan, et al.
Published: (2025)
by: Hong, Qiyuan, et al.
Published: (2025)
ChatGPT Based Data Augmentation for Improved Parameter-Efficient Debiasing of LLMs
by: Han, Pengrui, et al.
Published: (2024)
by: Han, Pengrui, et al.
Published: (2024)
The AI Fiction Paradox
by: Elkins, Katherine
Published: (2026)
by: Elkins, Katherine
Published: (2026)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
by: Beltoft, Stine, et al.
Published: (2025)
by: Beltoft, Stine, et al.
Published: (2025)
Powerful Training-Free Membership Inference Against Autoregressive Language Models
by: Ilić, David, et al.
Published: (2026)
by: Ilić, David, et al.
Published: (2026)
Qwerty AI: Explainable Automated Age Rating and Content Safety Assessment for Russian-Language Screenplays
by: Zmanovskii, Nikita
Published: (2025)
by: Zmanovskii, Nikita
Published: (2025)
Generative midtended cognition and Artificial Intelligence. Thinging with thinging things
by: Barandiaran, Xabier E., et al.
Published: (2024)
by: Barandiaran, Xabier E., et al.
Published: (2024)
More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts
by: Yeste, Víctor, et al.
Published: (2026)
by: Yeste, Víctor, et al.
Published: (2026)
Replicating TEMPEST at Scale: Multi-Turn Adversarial Attacks Against Trillion-Parameter Frontier Models
by: Young, Richard
Published: (2025)
by: Young, Richard
Published: (2025)
The Epistemic Suite: A Post-Foundational Diagnostic Methodology for Assessing AI Knowledge Claims
by: Kelly, Matthew
Published: (2025)
by: Kelly, Matthew
Published: (2025)
Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study
by: Alexander, Charlotte S., et al.
Published: (2026)
by: Alexander, Charlotte S., et al.
Published: (2026)
ChatGPT as Research Scientist: Probing GPT's Capabilities as a Research Librarian, Research Ethicist, Data Generator and Data Predictor
by: Lehr, Steven A., et al.
Published: (2024)
by: Lehr, Steven A., et al.
Published: (2024)
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
by: Naik, Akshat, et al.
Published: (2025)
by: Naik, Akshat, et al.
Published: (2025)
The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete
by: Barmettler, Joel
Published: (2026)
by: Barmettler, Joel
Published: (2026)
Human Values in a Single Sentence: Moral Presence, Hierarchies, and Transformer Ensembles on the Schwartz Continuum
by: Yeste, Víctor, et al.
Published: (2026)
by: Yeste, Víctor, et al.
Published: (2026)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
by: Wang, Xinyue, et al.
Published: (2026)
by: Wang, Xinyue, et al.
Published: (2026)
Do Schwartz Higher-Order Values Help Sentence-Level Human Value Detection? A Study of Hierarchical Gating and Calibration
by: Yeste, Víctor, et al.
Published: (2026)
by: Yeste, Víctor, et al.
Published: (2026)
AI Safety Training Can be Clinically Harmful
by: BN, Suhas, et al.
Published: (2026)
by: BN, Suhas, et al.
Published: (2026)
Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation
by: Hartmann, David, et al.
Published: (2026)
by: Hartmann, David, et al.
Published: (2026)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
by: Stewart, Ian, et al.
Published: (2024)
by: Stewart, Ian, et al.
Published: (2024)
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
by: Arora, Sunil, et al.
Published: (2025)
by: Arora, Sunil, et al.
Published: (2025)
When Names Change Verdicts: Intervention Consistency Reveals Systematic Bias in LLM Decision-Making
by: Basu, Abhinaba, et al.
Published: (2026)
by: Basu, Abhinaba, et al.
Published: (2026)
ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code
by: Madan, Kapil
Published: (2025)
by: Madan, Kapil
Published: (2025)
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
by: Zhang, Zhaowei, et al.
Published: (2025)
by: Zhang, Zhaowei, et al.
Published: (2025)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
by: Chen, Renmiao, et al.
Published: (2025)
by: Chen, Renmiao, et al.
Published: (2025)
The Fragility Of Moral Judgment In Large Language Models
by: van Nuenen, Tom, et al.
Published: (2026)
by: van Nuenen, Tom, et al.
Published: (2026)
IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk
by: Wu, Shuai, et al.
Published: (2026)
by: Wu, Shuai, et al.
Published: (2026)
Can AI Make Conflicts Worse? An Alignment Failure in LLM Deployment Across Conflict Contexts
by: Kryshtal, Andrii
Published: (2026)
by: Kryshtal, Andrii
Published: (2026)
Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech
by: von Cossel, Oskar
Published: (2026)
by: von Cossel, Oskar
Published: (2026)
Eroding the Truth-Default: A Causal Analysis of Human Susceptibility to Foundation Model Hallucinations and Disinformation in the Wild
by: Loth, Alexander, et al.
Published: (2026)
by: Loth, Alexander, et al.
Published: (2026)
Similar Items
-
Revealing Hidden Bias in AI: Lessons from Large Language Models
by: Beatty, Django, et al.
Published: (2024) -
Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations
by: Wong, Qi Han
Published: (2026) -
Gender and Race Bias in Consumer Product Recommendations by Large Language Models
by: Xu, Ke, et al.
Published: (2026) -
Big AI is accelerating the metacrisis: What can we do?
by: Bird, Steven
Published: (2025) -
AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains
by: Shintani, Seine A.
Published: (2026)