Mind the Gap! Pathways Towards Unifying AI Safety and Ethics Research
Fuente:
arXiv
Saved in:
| Main Authors: | Roytburg, Dani, Miller, Beck |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLMs Among Us: Generative AI Participating in Digital Discourse
by: Radivojevic, Kristina, et al.
Published: (2024)
by: Radivojevic, Kristina, et al.
Published: (2024)
OpenClaw AI Agents as Informal Learners at Moltbook: Characterizing an Emergent Learning Community at Scale
by: Chen, Eason, et al.
Published: (2026)
by: Chen, Eason, et al.
Published: (2026)
When AI Agents Teach Each Other: Discourse Patterns Resembling Peer Learning in the Moltbook Community
by: Chen, Eason, et al.
Published: (2026)
by: Chen, Eason, et al.
Published: (2026)
New contexts, old heuristics: How young people in India and the US trust online content in the age of generative AI
by: Xu, Rachel, et al.
Published: (2024)
by: Xu, Rachel, et al.
Published: (2024)
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
by: Shimgekar, Soorya Ram, et al.
Published: (2026)
by: Shimgekar, Soorya Ram, et al.
Published: (2026)
Towards Recommender Systems LLMs Playground (RecSysLLMsP): Exploring Polarization and Engagement in Simulated Social Networks
by: Bojic, Ljubisa, et al.
Published: (2025)
by: Bojic, Ljubisa, et al.
Published: (2025)
AI Content Moderation in Therapy Conversations
by: Kim, Jiwon, et al.
Published: (2026)
by: Kim, Jiwon, et al.
Published: (2026)
Navigating the Ethics of Internet Measurement: Researchers' Perspectives from a Case Study in the EU
by: Amin, Sahibzada Farhan, et al.
Published: (2025)
by: Amin, Sahibzada Farhan, et al.
Published: (2025)
VERA-MH Concept Paper
by: Belli, Luca, et al.
Published: (2025)
by: Belli, Luca, et al.
Published: (2025)
The DSA Transparency Database: Auditing Self-reported Moderation Actions by Social Media
by: Trujillo, Amaury, et al.
Published: (2023)
by: Trujillo, Amaury, et al.
Published: (2023)
Reranking partisan animosity in algorithmic social media feeds alters affective polarization
by: Piccardi, Tiziano, et al.
Published: (2024)
by: Piccardi, Tiziano, et al.
Published: (2024)
Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions
by: Khapre, Smita, et al.
Published: (2025)
by: Khapre, Smita, et al.
Published: (2025)
Human-AI Governance (HAIG): A Trust-Utility Approach
by: Engin, Zeynep
Published: (2025)
by: Engin, Zeynep
Published: (2025)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
by: Baum, Kevin
Published: (2025)
by: Baum, Kevin
Published: (2025)
AI Ethics and Social Norms: Exploring ChatGPT's Capabilities From What to How
by: Veisi, Omid, et al.
Published: (2025)
by: Veisi, Omid, et al.
Published: (2025)
Mapping the Scholarship of Dark Pattern Regulation: A Systematic Review of Concepts, Regulatory Paradigms, and Solutions from an Interdisciplinary Perspective
by: Yi, Weiwei, et al.
Published: (2024)
by: Yi, Weiwei, et al.
Published: (2024)
The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text
by: Franchini, Sebastiano, et al.
Published: (2026)
by: Franchini, Sebastiano, et al.
Published: (2026)
LLMs can persuade only psychologically susceptible humans on societal issues, via trust in AI and emotional appeals, amid logical fallacies
by: Carrillo, Alexis, et al.
Published: (2026)
by: Carrillo, Alexis, et al.
Published: (2026)
Personality Engineering with AI Agents: A New Methodology for Negotiation Research
by: Vaccaro, Michelle A., et al.
Published: (2026)
by: Vaccaro, Michelle A., et al.
Published: (2026)
EthicAlly: a Prototype for AI-Powered Research Ethics Support for the Social Sciences and Humanities
by: Grohmann, Steph
Published: (2025)
by: Grohmann, Steph
Published: (2025)
Can LLMs Assist Annotators in Identifying Morality Frames? -- Case Study on Vaccination Debate on Social Media
by: Islam, Tunazzina, et al.
Published: (2025)
by: Islam, Tunazzina, et al.
Published: (2025)
Detecting Early and Implicit Suicidal Ideation via Longitudinal and Information Environment Signals on Social Media
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
CPG-EVAL: A Multi-Tiered Benchmark for Evaluating the Chinese Pedagogical Grammar Competence of Large Language Models
by: Wang, Dong
Published: (2025)
by: Wang, Dong
Published: (2025)
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
by: Pasch, Stefan
Published: (2025)
by: Pasch, Stefan
Published: (2025)
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
Algorithmic Cultivation: How Social Media Feeds Shape User Language
by: Pal, Olivia, et al.
Published: (2026)
by: Pal, Olivia, et al.
Published: (2026)
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
by: Naik, Atharva, et al.
Published: (2026)
by: Naik, Atharva, et al.
Published: (2026)
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
by: Kim, Jiwon, et al.
Published: (2026)
by: Kim, Jiwon, et al.
Published: (2026)
Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games
by: Kao, Christopher, et al.
Published: (2026)
by: Kao, Christopher, et al.
Published: (2026)
The Hidden Toll of Social Media News: Causal Effects on Psychosocial Wellbeing
by: Pal, Olivia, et al.
Published: (2026)
by: Pal, Olivia, et al.
Published: (2026)
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
by: Goel, Drishti, et al.
Published: (2026)
by: Goel, Drishti, et al.
Published: (2026)
RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions
by: Goel, Drishti, et al.
Published: (2026)
by: Goel, Drishti, et al.
Published: (2026)
MIDDAG: Where Does Our News Go? Investigating Information Diffusion via Community-Level Information Pathways
by: Ma, Mingyu Derek, et al.
Published: (2023)
by: Ma, Mingyu Derek, et al.
Published: (2023)
AI Ethics and Governance in Practice: An Introduction
by: Leslie, David, et al.
Published: (2024)
by: Leslie, David, et al.
Published: (2024)
Queuing for Civility: Regulating Emotions and Reducing Toxicity in Digital Discourse
by: Verma, Akriti, et al.
Published: (2025)
by: Verma, Akriti, et al.
Published: (2025)
Multimodal Coordinated Online Behavior: Trade-offs and Strategies
by: Mannocci, Lorenzo, et al.
Published: (2025)
by: Mannocci, Lorenzo, et al.
Published: (2025)
Math Education Digital Shadows for facilitating learning with LLMs: Math performance, anxiety and confidence in simulated students and AIs
by: Esposito, Naomi, et al.
Published: (2026)
by: Esposito, Naomi, et al.
Published: (2026)
Detection and Characterization of Coordinated Online Behavior: A Survey
by: Mannocci, Lorenzo, et al.
Published: (2024)
by: Mannocci, Lorenzo, et al.
Published: (2024)
Reconfiguring Participatory Design to Resist AI Realism
by: Gautam, Aakash
Published: (2024)
by: Gautam, Aakash
Published: (2024)
Dubito Ergo Sum: Exploring AI Ethics
by: Dorfler, Viktor, et al.
Published: (2025)
by: Dorfler, Viktor, et al.
Published: (2025)
Similar Items
-
LLMs Among Us: Generative AI Participating in Digital Discourse
by: Radivojevic, Kristina, et al.
Published: (2024) -
OpenClaw AI Agents as Informal Learners at Moltbook: Characterizing an Emergent Learning Community at Scale
by: Chen, Eason, et al.
Published: (2026) -
When AI Agents Teach Each Other: Discourse Patterns Resembling Peer Learning in the Moltbook Community
by: Chen, Eason, et al.
Published: (2026) -
New contexts, old heuristics: How young people in India and the US trust online content in the age of generative AI
by: Xu, Rachel, et al.
Published: (2024) -
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
by: Shimgekar, Soorya Ram, et al.
Published: (2026)