Pro-AI Bias in Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Trabelsi, Benaya, Shaki, Jonathan, Kraus, Sarit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Revealing Hidden Bias in AI: Lessons from Large Language Models
por: Beatty, Django, et al.
Publicado: (2024)
por: Beatty, Django, et al.
Publicado: (2024)
Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations
por: Wong, Qi Han
Publicado: (2026)
por: Wong, Qi Han
Publicado: (2026)
Gender and Race Bias in Consumer Product Recommendations by Large Language Models
por: Xu, Ke, et al.
Publicado: (2026)
por: Xu, Ke, et al.
Publicado: (2026)
Big AI is accelerating the metacrisis: What can we do?
por: Bird, Steven
Publicado: (2025)
por: Bird, Steven
Publicado: (2025)
AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains
por: Shintani, Seine A.
Publicado: (2026)
por: Shintani, Seine A.
Publicado: (2026)
Growing a Tail: Increasing Output Diversity in Large Language Models
por: Shur-Ofry, Michal, et al.
Publicado: (2024)
por: Shur-Ofry, Michal, et al.
Publicado: (2024)
Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
por: Hossain, Ariyan, et al.
Publicado: (2025)
por: Hossain, Ariyan, et al.
Publicado: (2025)
APPSI-139: A Parallel Corpus of English Application Privacy Policy Summarization and Interpretation
por: Zhu, Pengyun, et al.
Publicado: (2026)
por: Zhu, Pengyun, et al.
Publicado: (2026)
WSC+: Enhancing The Winograd Schema Challenge Using Tree-of-Experts
por: Zahraei, Pardis Sadat, et al.
Publicado: (2024)
por: Zahraei, Pardis Sadat, et al.
Publicado: (2024)
Leveraging Multi-Source Textural UGC for Neighbourhood Housing Quality Assessment: A GPT-Enhanced Framework
por: Hong, Qiyuan, et al.
Publicado: (2025)
por: Hong, Qiyuan, et al.
Publicado: (2025)
ChatGPT Based Data Augmentation for Improved Parameter-Efficient Debiasing of LLMs
por: Han, Pengrui, et al.
Publicado: (2024)
por: Han, Pengrui, et al.
Publicado: (2024)
The AI Fiction Paradox
por: Elkins, Katherine
Publicado: (2026)
por: Elkins, Katherine
Publicado: (2026)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
por: Beltoft, Stine, et al.
Publicado: (2025)
por: Beltoft, Stine, et al.
Publicado: (2025)
Powerful Training-Free Membership Inference Against Autoregressive Language Models
por: Ilić, David, et al.
Publicado: (2026)
por: Ilić, David, et al.
Publicado: (2026)
Qwerty AI: Explainable Automated Age Rating and Content Safety Assessment for Russian-Language Screenplays
por: Zmanovskii, Nikita
Publicado: (2025)
por: Zmanovskii, Nikita
Publicado: (2025)
Generative midtended cognition and Artificial Intelligence. Thinging with thinging things
por: Barandiaran, Xabier E., et al.
Publicado: (2024)
por: Barandiaran, Xabier E., et al.
Publicado: (2024)
More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts
por: Yeste, Víctor, et al.
Publicado: (2026)
por: Yeste, Víctor, et al.
Publicado: (2026)
Replicating TEMPEST at Scale: Multi-Turn Adversarial Attacks Against Trillion-Parameter Frontier Models
por: Young, Richard
Publicado: (2025)
por: Young, Richard
Publicado: (2025)
The Epistemic Suite: A Post-Foundational Diagnostic Methodology for Assessing AI Knowledge Claims
por: Kelly, Matthew
Publicado: (2025)
por: Kelly, Matthew
Publicado: (2025)
Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study
por: Alexander, Charlotte S., et al.
Publicado: (2026)
por: Alexander, Charlotte S., et al.
Publicado: (2026)
ChatGPT as Research Scientist: Probing GPT's Capabilities as a Research Librarian, Research Ethicist, Data Generator and Data Predictor
por: Lehr, Steven A., et al.
Publicado: (2024)
por: Lehr, Steven A., et al.
Publicado: (2024)
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
por: Naik, Akshat, et al.
Publicado: (2025)
por: Naik, Akshat, et al.
Publicado: (2025)
The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete
por: Barmettler, Joel
Publicado: (2026)
por: Barmettler, Joel
Publicado: (2026)
Human Values in a Single Sentence: Moral Presence, Hierarchies, and Transformer Ensembles on the Schwartz Continuum
por: Yeste, Víctor, et al.
Publicado: (2026)
por: Yeste, Víctor, et al.
Publicado: (2026)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
por: Wang, Xinyue, et al.
Publicado: (2026)
por: Wang, Xinyue, et al.
Publicado: (2026)
Do Schwartz Higher-Order Values Help Sentence-Level Human Value Detection? A Study of Hierarchical Gating and Calibration
por: Yeste, Víctor, et al.
Publicado: (2026)
por: Yeste, Víctor, et al.
Publicado: (2026)
AI Safety Training Can be Clinically Harmful
por: BN, Suhas, et al.
Publicado: (2026)
por: BN, Suhas, et al.
Publicado: (2026)
Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation
por: Hartmann, David, et al.
Publicado: (2026)
por: Hartmann, David, et al.
Publicado: (2026)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
por: Stewart, Ian, et al.
Publicado: (2024)
por: Stewart, Ian, et al.
Publicado: (2024)
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
por: Arora, Sunil, et al.
Publicado: (2025)
por: Arora, Sunil, et al.
Publicado: (2025)
When Names Change Verdicts: Intervention Consistency Reveals Systematic Bias in LLM Decision-Making
por: Basu, Abhinaba, et al.
Publicado: (2026)
por: Basu, Abhinaba, et al.
Publicado: (2026)
ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code
por: Madan, Kapil
Publicado: (2025)
por: Madan, Kapil
Publicado: (2025)
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
por: Zhang, Zhaowei, et al.
Publicado: (2025)
por: Zhang, Zhaowei, et al.
Publicado: (2025)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
por: Chen, Renmiao, et al.
Publicado: (2025)
por: Chen, Renmiao, et al.
Publicado: (2025)
The Fragility Of Moral Judgment In Large Language Models
por: van Nuenen, Tom, et al.
Publicado: (2026)
por: van Nuenen, Tom, et al.
Publicado: (2026)
IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis
por: Bose, Joy
Publicado: (2026)
por: Bose, Joy
Publicado: (2026)
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk
por: Wu, Shuai, et al.
Publicado: (2026)
por: Wu, Shuai, et al.
Publicado: (2026)
Can AI Make Conflicts Worse? An Alignment Failure in LLM Deployment Across Conflict Contexts
por: Kryshtal, Andrii
Publicado: (2026)
por: Kryshtal, Andrii
Publicado: (2026)
Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech
por: von Cossel, Oskar
Publicado: (2026)
por: von Cossel, Oskar
Publicado: (2026)
Eroding the Truth-Default: A Causal Analysis of Human Susceptibility to Foundation Model Hallucinations and Disinformation in the Wild
por: Loth, Alexander, et al.
Publicado: (2026)
por: Loth, Alexander, et al.
Publicado: (2026)
Ejemplares similares
-
Revealing Hidden Bias in AI: Lessons from Large Language Models
por: Beatty, Django, et al.
Publicado: (2024) -
Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations
por: Wong, Qi Han
Publicado: (2026) -
Gender and Race Bias in Consumer Product Recommendations by Large Language Models
por: Xu, Ke, et al.
Publicado: (2026) -
Big AI is accelerating the metacrisis: What can we do?
por: Bird, Steven
Publicado: (2025) -
AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains
por: Shintani, Seine A.
Publicado: (2026)