The PacifAIst Benchmark:Would an Artificial Intelligence Choose to Sacrifice Itself for Human Safety?
Fuente:
arXiv
Saved in:
| Main Author: | Herrador, Manuel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ChatGPT-4 and other LLMs in the Turing Test: A Critical Analysis
by: Giunti, Marco
Published: (2025)
by: Giunti, Marco
Published: (2025)
Beyond the Org Chart: AI and the Transformation of Invisible Work
by: Rosenthal, Stephanie, et al.
Published: (2026)
by: Rosenthal, Stephanie, et al.
Published: (2026)
Human-centered trust framework: An HCI perspective
by: Sousa, Sonia, et al.
Published: (2023)
by: Sousa, Sonia, et al.
Published: (2023)
Towards Human Cognition Level-based Experiment Design for Counterfactual Explanations (XAI)
by: Suffian, Muhammad, et al.
Published: (2022)
by: Suffian, Muhammad, et al.
Published: (2022)
A Locally Executable AI System for Improving Preoperative Patient Communication: A Multi-Domain Clinical Evaluation
by: Sato, Motoki, et al.
Published: (2025)
by: Sato, Motoki, et al.
Published: (2025)
IDA: Breaking Barriers in No-code UI Automation Through Large Language Models and Human-Centric Design
by: Shlomov, Segev, et al.
Published: (2024)
by: Shlomov, Segev, et al.
Published: (2024)
Introducing and Interfacing with Cybersecurity -- A Cards Approach
by: Shah, Ryan, et al.
Published: (2023)
by: Shah, Ryan, et al.
Published: (2023)
The Impact of Artificial Intelligence on Human Thought
by: Gesnot, Rénald
Published: (2025)
by: Gesnot, Rénald
Published: (2025)
Artificial Intelligence: Beyound Ocularcentrism, the New Age of Humans Beyond the Spectacle
by: Moussaoui, Mustapha El
Published: (2026)
by: Moussaoui, Mustapha El
Published: (2026)
A Mixed User-Centered Approach to Enable Augmented Intelligence in Intelligent Tutoring Systems: The Case of MathAIde app
by: Guerino, Guilherme, et al.
Published: (2025)
by: Guerino, Guilherme, et al.
Published: (2025)
Would You Rely on an Eerie Agent? A Systematic Review of the Impact of the Uncanny Valley Effect on Trust in Human-Agent Interaction
by: Alipour, Ahdiyeh, et al.
Published: (2025)
by: Alipour, Ahdiyeh, et al.
Published: (2025)
The Use of Artificial Intelligence Tools in Assessing Content Validity: A Comparative Study with Human Experts
by: Gurdil, Hatice, et al.
Published: (2025)
by: Gurdil, Hatice, et al.
Published: (2025)
Inadequacies of Large Language Model Benchmarks in the Era of Generative Artificial Intelligence
by: McIntosh, Timothy R., et al.
Published: (2024)
by: McIntosh, Timothy R., et al.
Published: (2024)
Positive Alignment: Artificial Intelligence for Human Flourishing
by: Laukkonen, Ruben, et al.
Published: (2026)
by: Laukkonen, Ruben, et al.
Published: (2026)
Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis
by: Salehi, Pegah, et al.
Published: (2024)
by: Salehi, Pegah, et al.
Published: (2024)
As Good As A Coin Toss: Human detection of AI-generated images, videos, audio, and audiovisual stimuli
by: Cooke, Di, et al.
Published: (2024)
by: Cooke, Di, et al.
Published: (2024)
Exploring Teachers' Perception of Artificial Intelligence: The Socio-emotional Deficiency as Opportunities and Challenges in Human-AI Complementarity in K-12 Education
by: Oh, Soon-young, et al.
Published: (2024)
by: Oh, Soon-young, et al.
Published: (2024)
A Survey of Accessible Explainable Artificial Intelligence Research
by: Nwokoye, Chukwunonso Henry, et al.
Published: (2024)
by: Nwokoye, Chukwunonso Henry, et al.
Published: (2024)
False Sense of Security in Explainable Artificial Intelligence (XAI)
by: Chung, Neo Christopher, et al.
Published: (2024)
by: Chung, Neo Christopher, et al.
Published: (2024)
Towards Synergistic Teacher-AI Interactions with Generative Artificial Intelligence
by: Cukurova, Mutlu, et al.
Published: (2025)
by: Cukurova, Mutlu, et al.
Published: (2025)
Generative AI Usage of University Students: Navigating Between Education and Business
by: Walke, Fabian, et al.
Published: (2026)
by: Walke, Fabian, et al.
Published: (2026)
Perceptions of Discriminatory Decisions of Artificial Intelligence: Unpacking the Role of Individual Characteristics
by: Kim, Soojong
Published: (2024)
by: Kim, Soojong
Published: (2024)
Artificial Intelligence in Deliberation: The AI Penalty and the Emergence of a New Deliberative Divide
by: Jungherr, Andreas, et al.
Published: (2025)
by: Jungherr, Andreas, et al.
Published: (2025)
Artificial Intelligence for Inclusive Engineering Education: Advancing Equality, Diversity, and Ethical Leadership
by: Ibrahim, Mona G., et al.
Published: (2026)
by: Ibrahim, Mona G., et al.
Published: (2026)
Governance and Regulation of Artificial Intelligence in Developing Countries: A Case Study of Nigeria
by: Okoro, Uloma, et al.
Published: (2026)
by: Okoro, Uloma, et al.
Published: (2026)
Finnish 5th and 6th graders' misconceptions about Artificial Intelligence
by: Mertala, Pekka, et al.
Published: (2023)
by: Mertala, Pekka, et al.
Published: (2023)
GenAI Against Humanity: Nefarious Applications of Generative Artificial Intelligence and Large Language Models
by: Ferrara, Emilio
Published: (2023)
by: Ferrara, Emilio
Published: (2023)
Human/AI Collective Intelligence for Deliberative Democracy: A Human-Centred Design Approach
by: De Liddo, Anna, et al.
Published: (2026)
by: De Liddo, Anna, et al.
Published: (2026)
Situational Awareness as the Imperative Capability for Disaster Resilience in the Era of Complex Hazards and Artificial Intelligence
by: Pak, Hongrak, et al.
Published: (2025)
by: Pak, Hongrak, et al.
Published: (2025)
The Integration of Artificial Intelligence in Undergraduate Medical Education in Spain: Descriptive Analysis and International Perspectives
by: Janeiro, Ana Enériz, et al.
Published: (2025)
by: Janeiro, Ana Enériz, et al.
Published: (2025)
Investigating Collaborative Data Practices: a Case Study on Artificial Intelligence for Healthcare Research
by: Henkin, Rafael, et al.
Published: (2023)
by: Henkin, Rafael, et al.
Published: (2023)
EARN Fairness: Explaining, Asking, Reviewing, and Negotiating Artificial Intelligence Fairness Metrics Among Stakeholders
by: Luo, Lin, et al.
Published: (2024)
by: Luo, Lin, et al.
Published: (2024)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
by: Vaccaro, Michelle, et al.
Published: (2026)
by: Vaccaro, Michelle, et al.
Published: (2026)
Artificial Intelligence as a Training Tool in Clinical Psychology: A Comparison of Text-Based and Avatar Simulations
by: Sawah, V. El, et al.
Published: (2025)
by: Sawah, V. El, et al.
Published: (2025)
Mapping Public Perception of Artificial Intelligence: Expectations, Risk-Benefit Tradeoffs, and Value As Determinants for Societal Acceptance
by: Brauner, Philipp, et al.
Published: (2024)
by: Brauner, Philipp, et al.
Published: (2024)
Who Would Chatbots Vote For? Political Preferences of ChatGPT and Gemini in the 2024 European Union Elections
by: Haman, Michael, et al.
Published: (2024)
by: Haman, Michael, et al.
Published: (2024)
Enhancing Mental Health Support through Human-AI Collaboration: Toward Secure and Empathetic AI-enabled chatbots
by: AlMakinah, Rawan, et al.
Published: (2024)
by: AlMakinah, Rawan, et al.
Published: (2024)
Cognitive Dissonance Artificial Intelligence (CD-AI): The Mind at War with Itself. Harnessing Discomfort to Sharpen Critical Thinking
by: Deliu, Delia
Published: (2025)
by: Deliu, Delia
Published: (2025)
AI Assistants for Spaceflight Procedures: Combining Generative Pre-Trained Transformer and Retrieval-Augmented Generation on Knowledge Graphs With Augmented Reality Cues
by: Bensch, Oliver, et al.
Published: (2024)
by: Bensch, Oliver, et al.
Published: (2024)
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis
by: Petrov, Nikolay B, et al.
Published: (2024)
by: Petrov, Nikolay B, et al.
Published: (2024)
Similar Items
-
ChatGPT-4 and other LLMs in the Turing Test: A Critical Analysis
by: Giunti, Marco
Published: (2025) -
Beyond the Org Chart: AI and the Transformation of Invisible Work
by: Rosenthal, Stephanie, et al.
Published: (2026) -
Human-centered trust framework: An HCI perspective
by: Sousa, Sonia, et al.
Published: (2023) -
Towards Human Cognition Level-based Experiment Design for Counterfactual Explanations (XAI)
by: Suffian, Muhammad, et al.
Published: (2022) -
A Locally Executable AI System for Improving Preoperative Patient Communication: A Multi-Domain Clinical Evaluation
by: Sato, Motoki, et al.
Published: (2025)