Saved in:
| Main Authors: | Thorne, William, Shepherd, Rupert, Maynard, Diana |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.26426 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech
by: von Cossel, Oskar
Published: (2026)
by: von Cossel, Oskar
Published: (2026)
ChatGPT as Research Scientist: Probing GPT's Capabilities as a Research Librarian, Research Ethicist, Data Generator and Data Predictor
by: Lehr, Steven A., et al.
Published: (2024)
by: Lehr, Steven A., et al.
Published: (2024)
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
by: Arora, Sunil, et al.
Published: (2025)
by: Arora, Sunil, et al.
Published: (2025)
The Company You Keep: How LLMs Respond to Dark Triad Traits
by: Lu, Zeyi, et al.
Published: (2026)
by: Lu, Zeyi, et al.
Published: (2026)
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
by: Van Clief, Jake, et al.
Published: (2025)
by: Van Clief, Jake, et al.
Published: (2025)
Big AI is accelerating the metacrisis: What can we do?
by: Bird, Steven
Published: (2025)
by: Bird, Steven
Published: (2025)
Revealing Hidden Bias in AI: Lessons from Large Language Models
by: Beatty, Django, et al.
Published: (2024)
by: Beatty, Django, et al.
Published: (2024)
Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study
by: Alexander, Charlotte S., et al.
Published: (2026)
by: Alexander, Charlotte S., et al.
Published: (2026)
WSC+: Enhancing The Winograd Schema Challenge Using Tree-of-Experts
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
Leveraging Multi-Source Textural UGC for Neighbourhood Housing Quality Assessment: A GPT-Enhanced Framework
by: Hong, Qiyuan, et al.
Published: (2025)
by: Hong, Qiyuan, et al.
Published: (2025)
AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains
by: Shintani, Seine A.
Published: (2026)
by: Shintani, Seine A.
Published: (2026)
Auditing Preferences for Brands and Cultures in LLMs
by: Rienecker, Jasmine, et al.
Published: (2026)
by: Rienecker, Jasmine, et al.
Published: (2026)
Whose Name Comes Up? Auditing LLM-Based Scholar Recommendations
by: Barolo, Daniele, et al.
Published: (2025)
by: Barolo, Daniele, et al.
Published: (2025)
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
by: Zhang, Zhaowei, et al.
Published: (2025)
by: Zhang, Zhaowei, et al.
Published: (2025)
The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete
by: Barmettler, Joel
Published: (2026)
by: Barmettler, Joel
Published: (2026)
Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
by: Hossain, Ariyan, et al.
Published: (2025)
by: Hossain, Ariyan, et al.
Published: (2025)
Who Leads in the Shadows? ERGM and Centrality Analysis of Congressional Democrats on Bluesky
by: Hew, Gordon, et al.
Published: (2025)
by: Hew, Gordon, et al.
Published: (2025)
Cross-Subreddit Behavior as Open-Source Indicators of Coordinated Influence: A Case Study of r/Sino & r/China
by: Pilaud, Manon, et al.
Published: (2025)
by: Pilaud, Manon, et al.
Published: (2025)
Growing a Tail: Increasing Output Diversity in Large Language Models
by: Shur-Ofry, Michal, et al.
Published: (2024)
by: Shur-Ofry, Michal, et al.
Published: (2024)
Domain-Specific Constitutional AI: Enhancing Safety in LLM-Powered Mental Health Chatbots
by: Lyu, Chenhan, et al.
Published: (2025)
by: Lyu, Chenhan, et al.
Published: (2025)
Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations
by: Wong, Qi Han
Published: (2026)
by: Wong, Qi Han
Published: (2026)
Replicating TEMPEST at Scale: Multi-Turn Adversarial Attacks Against Trillion-Parameter Frontier Models
by: Young, Richard
Published: (2025)
by: Young, Richard
Published: (2025)
Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation
by: Hartmann, David, et al.
Published: (2026)
by: Hartmann, David, et al.
Published: (2026)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
by: Stewart, Ian, et al.
Published: (2024)
by: Stewart, Ian, et al.
Published: (2024)
Qwerty AI: Explainable Automated Age Rating and Content Safety Assessment for Russian-Language Screenplays
by: Zmanovskii, Nikita
Published: (2025)
by: Zmanovskii, Nikita
Published: (2025)
Generative midtended cognition and Artificial Intelligence. Thinging with thinging things
by: Barandiaran, Xabier E., et al.
Published: (2024)
by: Barandiaran, Xabier E., et al.
Published: (2024)
Reconstruction and Secrecy under Approximate Distance Queries
by: Moran, Shay, et al.
Published: (2025)
by: Moran, Shay, et al.
Published: (2025)
What are People Talking about in #BlackLivesMatter and #StopAsianHate? Exploring and Categorizing Twitter Topics Emerging in Online Social Movements through the Latent Dirichlet Allocation Model
by: Tong, Xin, et al.
Published: (2022)
by: Tong, Xin, et al.
Published: (2022)
Industrialized Deception: The Collateral Effects of LLM-Generated Misinformation on Digital Ecosystems
by: Loth, Alexander, et al.
Published: (2026)
by: Loth, Alexander, et al.
Published: (2026)
Who Shares What? An Empirical Analysis of Security Conference Content Across Academia and Industry
by: Walter, Lukas, et al.
Published: (2024)
by: Walter, Lukas, et al.
Published: (2024)
Evaluating LLM-Based Grant Proposal Review via Structured Perturbations
by: Thorne, William, et al.
Published: (2026)
by: Thorne, William, et al.
Published: (2026)
Beyond classical and contemporary models: a transformative AI framework for student dropout prediction in distance learning using RAG, Prompt engineering, and Cross-modal fusion
by: Mihoubi, Miloud, et al.
Published: (2025)
by: Mihoubi, Miloud, et al.
Published: (2025)
Ideology as a Problem: Lightweight Logit Steering for Annotator-Specific Alignment in Social Media Analysis
by: Xia, Wei, et al.
Published: (2025)
by: Xia, Wei, et al.
Published: (2025)
AI Safety Training Can be Clinically Harmful
by: BN, Suhas, et al.
Published: (2026)
by: BN, Suhas, et al.
Published: (2026)
Can LLMs Understand What We Cannot Say? Measuring Multilevel Alignment Through Abortion Stigma Across Cognitive, Interpersonal, and Structural Levels
by: Sharma, Anika, et al.
Published: (2025)
by: Sharma, Anika, et al.
Published: (2025)
A Field Guide to Decision Making
by: Arthur, Richard B.
Published: (2026)
by: Arthur, Richard B.
Published: (2026)
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
by: Young, Richard J., et al.
Published: (2026)
by: Young, Richard J., et al.
Published: (2026)
Balancing Innovation and Integrity: AI Integration in Liberal Arts College Administration
by: Read, Ian Olivo
Published: (2025)
by: Read, Ian Olivo
Published: (2025)
When Names Change Verdicts: Intervention Consistency Reveals Systematic Bias in LLM Decision-Making
by: Basu, Abhinaba, et al.
Published: (2026)
by: Basu, Abhinaba, et al.
Published: (2026)
Transforming Computer Security and Public Trust Through the Exploration of Fine-Tuning Large Language Models
by: Crumrine, Garrett, et al.
Published: (2024)
by: Crumrine, Garrett, et al.
Published: (2024)
Similar Items
-
Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech
by: von Cossel, Oskar
Published: (2026) -
ChatGPT as Research Scientist: Probing GPT's Capabilities as a Research Librarian, Research Ethicist, Data Generator and Data Predictor
by: Lehr, Steven A., et al.
Published: (2024) -
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
by: Arora, Sunil, et al.
Published: (2025) -
The Company You Keep: How LLMs Respond to Dark Triad Traits
by: Lu, Zeyi, et al.
Published: (2026) -
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
by: Van Clief, Jake, et al.
Published: (2025)