Navigating LLM Ethics: Advancements, Challenges, and Future Directions
Fuente:
arXiv
Guardado en:
| Autores principales: | Jiao, Junfeng, Afroogh, Saleh, Xu, Yiming, Phillips, Connor |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The global landscape of academic guidelines for generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2024)
por: Jiao, Junfeng, et al.
Publicado: (2024)
Trust in AI: Progress, Challenges, and Future Directions
por: Afroogh, Saleh, et al.
Publicado: (2024)
por: Afroogh, Saleh, et al.
Publicado: (2024)
Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
por: Park, Jihyung, et al.
Publicado: (2025)
por: Park, Jihyung, et al.
Publicado: (2025)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions
por: Hu, Yiran, et al.
Publicado: (2026)
por: Hu, Yiran, et al.
Publicado: (2026)
Intelligent Environmental Empathy (IEE): A new power and platform to fostering green obligation for climate peace and justice
por: Afroogh, Saleh, et al.
Publicado: (2024)
por: Afroogh, Saleh, et al.
Publicado: (2024)
Mapping out AI Functions in Intelligent Disaster (Mis)Management and AI-Caused Disasters
por: Pouresmaeil, Yasser, et al.
Publicado: (2025)
por: Pouresmaeil, Yasser, et al.
Publicado: (2025)
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Evaluating Retrieval-Augmented Generation Strategies for Large Language Models in Travel Mode Choice Prediction
por: Xu, Yiming, et al.
Publicado: (2025)
por: Xu, Yiming, et al.
Publicado: (2025)
AI Empathy Erodes Cognitive Autonomy in Younger Users
por: Jiao, Junfeng, et al.
Publicado: (2026)
por: Jiao, Junfeng, et al.
Publicado: (2026)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
por: Backmann, Steffen, et al.
Publicado: (2025)
por: Backmann, Steffen, et al.
Publicado: (2025)
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
por: Li, Minzhi, et al.
Publicado: (2024)
por: Li, Minzhi, et al.
Publicado: (2024)
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue
por: Park, Jihyung, et al.
Publicado: (2026)
por: Park, Jihyung, et al.
Publicado: (2026)
Safe-Child-LLM: A Developmental Benchmark for Evaluating LLM Safety in Child-LLM Interactions
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Ethical Concern Identification in NLP: A Corpus of ACL Anthology Ethics Statements
por: Karamolegkou, Antonia, et al.
Publicado: (2024)
por: Karamolegkou, Antonia, et al.
Publicado: (2024)
Risks from Language Models for Automated Mental Healthcare: Ethics and Structure for Implementation
por: Grabb, Declan, et al.
Publicado: (2024)
por: Grabb, Declan, et al.
Publicado: (2024)
EthicsMH: A Pilot Benchmark for Ethical Reasoning in Mental Health AI
por: Kasu, Sai Kartheek Reddy
Publicado: (2025)
por: Kasu, Sai Kartheek Reddy
Publicado: (2025)
Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study
por: Silva, Jhessica, et al.
Publicado: (2025)
por: Silva, Jhessica, et al.
Publicado: (2025)
LLM Harms: A Taxonomy and Discussion
por: Chen, Kevin, et al.
Publicado: (2025)
por: Chen, Kevin, et al.
Publicado: (2025)
When AI Navigates the Fog of War
por: Li, Ming, et al.
Publicado: (2026)
por: Li, Ming, et al.
Publicado: (2026)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
por: Ranjan, Rajesh, et al.
Publicado: (2024)
por: Ranjan, Rajesh, et al.
Publicado: (2024)
Explainable Artificial Intelligence: A Survey of Needs, Techniques, Applications, and Future Direction
por: Mersha, Melkamu, et al.
Publicado: (2024)
por: Mersha, Melkamu, et al.
Publicado: (2024)
KidLM: Advancing Language Models for Children -- Early Insights and Future Directions
por: Nayeem, Mir Tafseer, et al.
Publicado: (2024)
por: Nayeem, Mir Tafseer, et al.
Publicado: (2024)
Epicure: Navigating the Emergent Geometry of Food Ingredient Embeddings
por: Radzikowski, Jakub, et al.
Publicado: (2026)
por: Radzikowski, Jakub, et al.
Publicado: (2026)
Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions
por: Khapre, Smita, et al.
Publicado: (2025)
por: Khapre, Smita, et al.
Publicado: (2025)
Right to be Forgotten in the Era of Large Language Models: Implications, Challenges, and Solutions
por: Zhang, Dawen, et al.
Publicado: (2023)
por: Zhang, Dawen, et al.
Publicado: (2023)
Evaluating LLM Safety Across Child Development Stages: A Simulated Agent Approach
por: Murali, Abhejay, et al.
Publicado: (2025)
por: Murali, Abhejay, et al.
Publicado: (2025)
How to Protect Yourself from 5G Radiation? Investigating LLM Responses to Implicit Misinformation
por: Guo, Ruohao, et al.
Publicado: (2025)
por: Guo, Ruohao, et al.
Publicado: (2025)
LLMs and Childhood Safety: Identifying Risks and Proposing a Protection Framework for Safe Child-LLM Interaction
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Navigating the Cultural Kaleidoscope: A Hitchhiker's Guide to Sensitivity in Large Language Models
por: Banerjee, Somnath, et al.
Publicado: (2024)
por: Banerjee, Somnath, et al.
Publicado: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
por: Ahmed, Ahmed Haj, et al.
Publicado: (2024)
por: Ahmed, Ahmed Haj, et al.
Publicado: (2024)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
por: Liu, Xiaoze, et al.
Publicado: (2024)
por: Liu, Xiaoze, et al.
Publicado: (2024)
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
por: Wu, Addison J., et al.
Publicado: (2026)
por: Wu, Addison J., et al.
Publicado: (2026)
Denevil: Towards Deciphering and Navigating the Ethical Values of Large Language Models via Instruction Learning
por: Duan, Shitong, et al.
Publicado: (2023)
por: Duan, Shitong, et al.
Publicado: (2023)
Generative AI and LLMs in Industry: A text-mining Analysis and Critical Evaluation of Guidelines and Policy Statements Across Fourteen Industrial Sectors
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
IGGA: A Dataset of Industrial Guidelines and Policy Statements for Generative AIs
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Safe in the Future, Dangerous in the Past: Dissecting Temporal and Linguistic Vulnerabilities in LLMs
por: Said, Muhammad Abdullahi, et al.
Publicado: (2025)
por: Said, Muhammad Abdullahi, et al.
Publicado: (2025)
LLM Nepotism in Organizational Governance
por: Mao, Shunqi, et al.
Publicado: (2026)
por: Mao, Shunqi, et al.
Publicado: (2026)
Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions
por: He, Yuting, et al.
Publicado: (2024)
por: He, Yuting, et al.
Publicado: (2024)
Explainability Through Systematicity: The Hard Systematicity Challenge for Artificial Intelligence
por: Queloz, Matthieu
Publicado: (2025)
por: Queloz, Matthieu
Publicado: (2025)
Ejemplares similares
-
The global landscape of academic guidelines for generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2024) -
Trust in AI: Progress, Challenges, and Future Directions
por: Afroogh, Saleh, et al.
Publicado: (2024) -
Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
por: Park, Jihyung, et al.
Publicado: (2025) -
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2025) -
Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions
por: Hu, Yiran, et al.
Publicado: (2026)