LLM Harms: A Taxonomy and Discussion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Kevin, Afroogh, Saleh, Murali, Abhejay, Atkinson, David, Dhurandhar, Amit, Jiao, Junfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
Evaluating LLM Safety Across Child Development Stages: A Simulated Agent Approach
von: Murali, Abhejay, et al.
Veröffentlicht: (2025)
von: Murali, Abhejay, et al.
Veröffentlicht: (2025)
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
LLMs and Childhood Safety: Identifying Risks and Proposing a Protection Framework for Safe Child-LLM Interaction
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
AI Empathy Erodes Cognitive Autonomy in Younger Users
von: Jiao, Junfeng, et al.
Veröffentlicht: (2026)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2026)
IGGA: A Dataset of Industrial Guidelines and Policy Statements for Generative AIs
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
Generative AI and LLMs in Industry: A text-mining Analysis and Critical Evaluation of Guidelines and Policy Statements Across Fourteen Industrial Sectors
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
The global landscape of academic guidelines for generative AI and Large Language Models
von: Jiao, Junfeng, et al.
Veröffentlicht: (2024)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2024)
Evaluating the Effectiveness of OpenAI's Parental Control System
von: Ersoz, Kerem, et al.
Veröffentlicht: (2026)
von: Ersoz, Kerem, et al.
Veröffentlicht: (2026)
Intelligent Environmental Empathy (IEE): A new power and platform to fostering green obligation for climate peace and justice
von: Afroogh, Saleh, et al.
Veröffentlicht: (2024)
von: Afroogh, Saleh, et al.
Veröffentlicht: (2024)
Mapping out AI Functions in Intelligent Disaster (Mis)Management and AI-Caused Disasters
von: Pouresmaeil, Yasser, et al.
Veröffentlicht: (2025)
von: Pouresmaeil, Yasser, et al.
Veröffentlicht: (2025)
Navigating LLM Ethics: Advancements, Challenges, and Future Directions
von: Jiao, Junfeng, et al.
Veröffentlicht: (2024)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2024)
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue
von: Park, Jihyung, et al.
Veröffentlicht: (2026)
von: Park, Jihyung, et al.
Veröffentlicht: (2026)
Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
von: Park, Jihyung, et al.
Veröffentlicht: (2025)
von: Park, Jihyung, et al.
Veröffentlicht: (2025)
A Legal Risk Taxonomy for Generative Artificial Intelligence
von: Atkinson, David, et al.
Veröffentlicht: (2024)
von: Atkinson, David, et al.
Veröffentlicht: (2024)
A Task-Driven Human-AI Collaboration: When to Automate, When to Collaborate, When to Challenge
von: Afroogh, Saleh, et al.
Veröffentlicht: (2025)
von: Afroogh, Saleh, et al.
Veröffentlicht: (2025)
When Trust is Zero Sum: Automation Threat to Epistemic Agency
von: Malone, Emmie, et al.
Veröffentlicht: (2024)
von: Malone, Emmie, et al.
Veröffentlicht: (2024)
Towards a Harms Taxonomy of AI Likeness Generation
von: Bariach, Ben, et al.
Veröffentlicht: (2024)
von: Bariach, Ben, et al.
Veröffentlicht: (2024)
Trust in AI: Progress, Challenges, and Future Directions
von: Afroogh, Saleh, et al.
Veröffentlicht: (2024)
von: Afroogh, Saleh, et al.
Veröffentlicht: (2024)
Vernacularizing Taxonomies of Harm is Essential for Operationalizing Holistic AI Safety
von: Kennedy, Wm. Matthew, et al.
Veröffentlicht: (2024)
von: Kennedy, Wm. Matthew, et al.
Veröffentlicht: (2024)
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
von: Jo, Claire Wonjeong, et al.
Veröffentlicht: (2024)
von: Jo, Claire Wonjeong, et al.
Veröffentlicht: (2024)
Putting GenAI on Notice: GenAI Exceptionalism and Contract Law
von: Atkinson, David
Veröffentlicht: (2025)
von: Atkinson, David
Veröffentlicht: (2025)
Addressing the Unforeseen Harms of Technology CCC Whitepaper
von: Bliss, Nadya, et al.
Veröffentlicht: (2024)
von: Bliss, Nadya, et al.
Veröffentlicht: (2024)
The Dark Side of AI Companionship: A Taxonomy of Harmful Algorithmic Behaviors in Human-AI Relationships
von: Zhang, Renwen, et al.
Veröffentlicht: (2024)
von: Zhang, Renwen, et al.
Veröffentlicht: (2024)
Not My Voice! A Taxonomy of Ethical and Safety Harms of Speech Generators
von: Hutiri, Wiebke, et al.
Veröffentlicht: (2024)
von: Hutiri, Wiebke, et al.
Veröffentlicht: (2024)
Unfair Learning: GenAI Exceptionalism and Copyright Law
von: Atkinson, David
Veröffentlicht: (2025)
von: Atkinson, David
Veröffentlicht: (2025)
In Quest of an Extensible Multi-Level Harm Taxonomy for Adversarial AI: Heart of Security, Ethical Risk Scoring and Resilience Analytics
von: Khan, Javed I., et al.
Veröffentlicht: (2026)
von: Khan, Javed I., et al.
Veröffentlicht: (2026)
Designing Incident Reporting Systems for Harms from General-Purpose AI
von: Wei, Kevin, et al.
Veröffentlicht: (2025)
von: Wei, Kevin, et al.
Veröffentlicht: (2025)
Echoes of AI Harms: A Human-LLM Synergistic Framework for Bias-Driven Harm Anticipation
von: Tantalaki, Nicoleta, et al.
Veröffentlicht: (2025)
von: Tantalaki, Nicoleta, et al.
Veröffentlicht: (2025)
LLM-based Semantic Augmentation for Harmful Content Detection
von: Meguellati, Elyas, et al.
Veröffentlicht: (2025)
von: Meguellati, Elyas, et al.
Veröffentlicht: (2025)
Unsettled Law: Time to Generate New Approaches?
von: Atkinson, David, et al.
Veröffentlicht: (2024)
von: Atkinson, David, et al.
Veröffentlicht: (2024)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
von: Baum, Kevin
Veröffentlicht: (2025)
von: Baum, Kevin
Veröffentlicht: (2025)
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
In the Mood to Exclude: Revitalizing Trespass to Chattels in the Era of GenAI Scraping
von: Atkinson, David
Veröffentlicht: (2025)
von: Atkinson, David
Veröffentlicht: (2025)
Open Shouldn't Mean Exempt: Open-Source Exceptionalism and Generative AI
von: Atkinson, David
Veröffentlicht: (2025)
von: Atkinson, David
Veröffentlicht: (2025)
LLM Use for Mental Health: Crowdsourcing Users' Sentiment-based Perspectives and Values from Social Discussions
von: Li, Lingyao, et al.
Veröffentlicht: (2025)
von: Li, Lingyao, et al.
Veröffentlicht: (2025)
From Hallucination to Scheming: A Unified Taxonomy and Benchmark Analysis for LLM Deception
von: Shi, Jerick, et al.
Veröffentlicht: (2026)
von: Shi, Jerick, et al.
Veröffentlicht: (2026)
Constitutive vs. Corrective: A Causal Taxonomy of Human Runtime Involvement in AI Systems
von: Baum, Kevin, et al.
Veröffentlicht: (2026)
von: Baum, Kevin, et al.
Veröffentlicht: (2026)
Evaluating Retrieval-Augmented Generation Strategies for Large Language Models in Travel Mode Choice Prediction
von: Xu, Yiming, et al.
Veröffentlicht: (2025)
von: Xu, Yiming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025) -
Evaluating LLM Safety Across Child Development Stages: A Simulated Agent Approach
von: Murali, Abhejay, et al.
Veröffentlicht: (2025) -
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025) -
LLMs and Childhood Safety: Identifying Risks and Proposing a Protection Framework for Safe Child-LLM Interaction
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025) -
AI Empathy Erodes Cognitive Autonomy in Younger Users
von: Jiao, Junfeng, et al.
Veröffentlicht: (2026)