Human-AI Safety: A Descendant of Generative AI and Control Systems Safety
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bajcsy, Andrea, Fisac, Jaime F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sentiment Simulation using Generative AI Agents
von: Tia, Melrose, et al.
Veröffentlicht: (2025)
von: Tia, Melrose, et al.
Veröffentlicht: (2025)
Our Cars Can Talk: How IoT Brings AI to Vehicles
von: Agrawal, Amod Kant
Veröffentlicht: (2025)
von: Agrawal, Amod Kant
Veröffentlicht: (2025)
Will Agents Replace Us? Perceptions of Autonomous Multi-Agent AI
von: Balic, Nikola
Veröffentlicht: (2025)
von: Balic, Nikola
Veröffentlicht: (2025)
From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysis
von: Dietrich, Juergen
Veröffentlicht: (2026)
von: Dietrich, Juergen
Veröffentlicht: (2026)
Pedagogical Safety in Educational Reinforcement Learning: Formalizing and Detecting Reward Hacking in AI Tutoring Systems
von: Olukola, Oluseyi, et al.
Veröffentlicht: (2026)
von: Olukola, Oluseyi, et al.
Veröffentlicht: (2026)
Navigating Ethical Challenges in Generative AI-Enhanced Research: The ETHICAL Framework for Responsible Generative AI Use
von: Eacersall, Douglas, et al.
Veröffentlicht: (2024)
von: Eacersall, Douglas, et al.
Veröffentlicht: (2024)
Trustworthy AI: UK Air Traffic Control Revisited
von: Procter, Rob, et al.
Veröffentlicht: (2025)
von: Procter, Rob, et al.
Veröffentlicht: (2025)
Limits of Safe AI Deployment: Differentiating Oversight and Control
von: Manheim, David, et al.
Veröffentlicht: (2025)
von: Manheim, David, et al.
Veröffentlicht: (2025)
Navigating the sociotechnical labyrinth: Dynamic certification for responsible embodied AI
von: Bakirtzis, Georgios, et al.
Veröffentlicht: (2024)
von: Bakirtzis, Georgios, et al.
Veröffentlicht: (2024)
Safety Margins for Reinforcement Learning
von: Grushin, Alexander, et al.
Veröffentlicht: (2023)
von: Grushin, Alexander, et al.
Veröffentlicht: (2023)
Physics-inspired Neural Networks for Parameter Learning of Adaptive Cruise Control Systems
von: Apostolakis, Theocharis, et al.
Veröffentlicht: (2023)
von: Apostolakis, Theocharis, et al.
Veröffentlicht: (2023)
Penalizing Transparency? How AI Disclosure and Author Demographics Shape Human and AI Judgments About Writing
von: Cheong, Inyoung, et al.
Veröffentlicht: (2025)
von: Cheong, Inyoung, et al.
Veröffentlicht: (2025)
Provocations from the Humanities for Generative AI Research
von: Klein, Lauren, et al.
Veröffentlicht: (2025)
von: Klein, Lauren, et al.
Veröffentlicht: (2025)
Criticality and Safety Margins for Reinforcement Learning
von: Grushin, Alexander, et al.
Veröffentlicht: (2024)
von: Grushin, Alexander, et al.
Veröffentlicht: (2024)
AI for All: Identifying AI incidents Related to Diversity and Inclusion
von: Shams, Rifat Ara, et al.
Veröffentlicht: (2024)
von: Shams, Rifat Ara, et al.
Veröffentlicht: (2024)
Self-Emotion Blended Dialogue Generation in Social Simulation Agents
von: Zhang, Qiang, et al.
Veröffentlicht: (2024)
von: Zhang, Qiang, et al.
Veröffentlicht: (2024)
Positive AI: Key Challenges in Designing Artificial Intelligence for Wellbeing
von: van der Maden, Willem, et al.
Veröffentlicht: (2023)
von: van der Maden, Willem, et al.
Veröffentlicht: (2023)
Who Gets the Kidney? Human-AI Alignment, Indecision, and Moral Values
von: Dickerson, John P., et al.
Veröffentlicht: (2025)
von: Dickerson, John P., et al.
Veröffentlicht: (2025)
Characterizing AI Agents for Alignment and Governance
von: Kasirzadeh, Atoosa, et al.
Veröffentlicht: (2025)
von: Kasirzadeh, Atoosa, et al.
Veröffentlicht: (2025)
The SMART+ Framework for AI Systems
von: Kandikatla, Laxmiraju, et al.
Veröffentlicht: (2025)
von: Kandikatla, Laxmiraju, et al.
Veröffentlicht: (2025)
Prescriptive Artificial Intelligence: A Formal Paradigm for Auditing Human Decisions Under Uncertainty
von: Passos, Pedro, et al.
Veröffentlicht: (2025)
von: Passos, Pedro, et al.
Veröffentlicht: (2025)
The Ethical Implications of AI in Creative Industries: A Focus on AI-Generated Art
von: Khatiwada, Prerana, et al.
Veröffentlicht: (2025)
von: Khatiwada, Prerana, et al.
Veröffentlicht: (2025)
AI and the Problem of Knowledge Collapse
von: Peterson, Andrew J.
Veröffentlicht: (2024)
von: Peterson, Andrew J.
Veröffentlicht: (2024)
Autonomous AI and Ownership Rules
von: Fagan, Frank
Veröffentlicht: (2026)
von: Fagan, Frank
Veröffentlicht: (2026)
Structured AI Decision-Making in Disaster Management
von: Dcruz, Julian Gerald, et al.
Veröffentlicht: (2025)
von: Dcruz, Julian Gerald, et al.
Veröffentlicht: (2025)
Generative AI in Science: Applications, Challenges, and Emerging Questions
von: Harries, Ryan, et al.
Veröffentlicht: (2025)
von: Harries, Ryan, et al.
Veröffentlicht: (2025)
Ensuring Both Positivity and Stability Using Sector-Bounded Nonlinearity for Systems with Neural Network Controllers
von: Hedesh, Hamidreza Montazeri, et al.
Veröffentlicht: (2024)
von: Hedesh, Hamidreza Montazeri, et al.
Veröffentlicht: (2024)
Systematic Hazard Analysis for Frontier AI using STPA
von: Mylius, Simon
Veröffentlicht: (2025)
von: Mylius, Simon
Veröffentlicht: (2025)
Understanding Gen Alpha Digital Language: Evaluation of LLM Safety Systems for Content Moderation
von: Mehta, Manisha, et al.
Veröffentlicht: (2025)
von: Mehta, Manisha, et al.
Veröffentlicht: (2025)
The human biological advantage over AI
von: Stewart, William
Veröffentlicht: (2025)
von: Stewart, William
Veröffentlicht: (2025)
SCOR: A Framework for Responsible AI Innovation in Digital Ecosystems
von: Torkestani, Mohammad Saleh, et al.
Veröffentlicht: (2025)
von: Torkestani, Mohammad Saleh, et al.
Veröffentlicht: (2025)
GAIA: A General Agency Interaction Architecture for LLM-Human B2B Negotiation & Screening
von: Zhao, Siming, et al.
Veröffentlicht: (2025)
von: Zhao, Siming, et al.
Veröffentlicht: (2025)
Towards Friendly AI: A Comprehensive Review and New Perspectives on Human-AI Alignment
von: Sun, Qiyang, et al.
Veröffentlicht: (2024)
von: Sun, Qiyang, et al.
Veröffentlicht: (2024)
AI Literacy in K-12 and Higher Education in the Wake of Generative AI: An Integrative Review
von: Gu, Xingjian, et al.
Veröffentlicht: (2025)
von: Gu, Xingjian, et al.
Veröffentlicht: (2025)
BCause: Human-AI collaboration to improve hybrid mapping and ideation in argumentation-grounded deliberation
von: Anastasiou, Lucas, et al.
Veröffentlicht: (2025)
von: Anastasiou, Lucas, et al.
Veröffentlicht: (2025)
Estimating the Increase in Emissions caused by AI-augmented Search
von: Vanderbauwhede, Wim
Veröffentlicht: (2024)
von: Vanderbauwhede, Wim
Veröffentlicht: (2024)
Agentic AI for Scaling Diagnosis and Care in Neurodegenerative Disease
von: Breithaupt, Andrew G., et al.
Veröffentlicht: (2025)
von: Breithaupt, Andrew G., et al.
Veröffentlicht: (2025)
Cognitive Amplification vs Cognitive Delegation in Human-AI Systems: A Metric Framework
von: Di Santi, Eduardo
Veröffentlicht: (2026)
von: Di Santi, Eduardo
Veröffentlicht: (2026)
AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-Produce Supply Chains
von: Vakhnovskyi, Andrii
Veröffentlicht: (2026)
von: Vakhnovskyi, Andrii
Veröffentlicht: (2026)
Strategizing Equitable Transit Evacuations: A Data-Driven Reinforcement Learning Approach
von: Tang, Fang, et al.
Veröffentlicht: (2024)
von: Tang, Fang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Sentiment Simulation using Generative AI Agents
von: Tia, Melrose, et al.
Veröffentlicht: (2025) -
Our Cars Can Talk: How IoT Brings AI to Vehicles
von: Agrawal, Amod Kant
Veröffentlicht: (2025) -
Will Agents Replace Us? Perceptions of Autonomous Multi-Agent AI
von: Balic, Nikola
Veröffentlicht: (2025) -
From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysis
von: Dietrich, Juergen
Veröffentlicht: (2026) -
Pedagogical Safety in Educational Reinforcement Learning: Formalizing and Detecting Reward Hacking in AI Tutoring Systems
von: Olukola, Oluseyi, et al.
Veröffentlicht: (2026)