Human-AI Safety: A Descendant of Generative AI and Control Systems Safety
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bajcsy, Andrea, Fisac, Jaime F. |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Sentiment Simulation using Generative AI Agents
par: Tia, Melrose, et autres
Publié: (2025)
par: Tia, Melrose, et autres
Publié: (2025)
Our Cars Can Talk: How IoT Brings AI to Vehicles
par: Agrawal, Amod Kant
Publié: (2025)
par: Agrawal, Amod Kant
Publié: (2025)
Will Agents Replace Us? Perceptions of Autonomous Multi-Agent AI
par: Balic, Nikola
Publié: (2025)
par: Balic, Nikola
Publié: (2025)
From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysis
par: Dietrich, Juergen
Publié: (2026)
par: Dietrich, Juergen
Publié: (2026)
Pedagogical Safety in Educational Reinforcement Learning: Formalizing and Detecting Reward Hacking in AI Tutoring Systems
par: Olukola, Oluseyi, et autres
Publié: (2026)
par: Olukola, Oluseyi, et autres
Publié: (2026)
Navigating Ethical Challenges in Generative AI-Enhanced Research: The ETHICAL Framework for Responsible Generative AI Use
par: Eacersall, Douglas, et autres
Publié: (2024)
par: Eacersall, Douglas, et autres
Publié: (2024)
Trustworthy AI: UK Air Traffic Control Revisited
par: Procter, Rob, et autres
Publié: (2025)
par: Procter, Rob, et autres
Publié: (2025)
Limits of Safe AI Deployment: Differentiating Oversight and Control
par: Manheim, David, et autres
Publié: (2025)
par: Manheim, David, et autres
Publié: (2025)
Navigating the sociotechnical labyrinth: Dynamic certification for responsible embodied AI
par: Bakirtzis, Georgios, et autres
Publié: (2024)
par: Bakirtzis, Georgios, et autres
Publié: (2024)
Safety Margins for Reinforcement Learning
par: Grushin, Alexander, et autres
Publié: (2023)
par: Grushin, Alexander, et autres
Publié: (2023)
Physics-inspired Neural Networks for Parameter Learning of Adaptive Cruise Control Systems
par: Apostolakis, Theocharis, et autres
Publié: (2023)
par: Apostolakis, Theocharis, et autres
Publié: (2023)
Penalizing Transparency? How AI Disclosure and Author Demographics Shape Human and AI Judgments About Writing
par: Cheong, Inyoung, et autres
Publié: (2025)
par: Cheong, Inyoung, et autres
Publié: (2025)
Provocations from the Humanities for Generative AI Research
par: Klein, Lauren, et autres
Publié: (2025)
par: Klein, Lauren, et autres
Publié: (2025)
Criticality and Safety Margins for Reinforcement Learning
par: Grushin, Alexander, et autres
Publié: (2024)
par: Grushin, Alexander, et autres
Publié: (2024)
AI for All: Identifying AI incidents Related to Diversity and Inclusion
par: Shams, Rifat Ara, et autres
Publié: (2024)
par: Shams, Rifat Ara, et autres
Publié: (2024)
Self-Emotion Blended Dialogue Generation in Social Simulation Agents
par: Zhang, Qiang, et autres
Publié: (2024)
par: Zhang, Qiang, et autres
Publié: (2024)
Positive AI: Key Challenges in Designing Artificial Intelligence for Wellbeing
par: van der Maden, Willem, et autres
Publié: (2023)
par: van der Maden, Willem, et autres
Publié: (2023)
Who Gets the Kidney? Human-AI Alignment, Indecision, and Moral Values
par: Dickerson, John P., et autres
Publié: (2025)
par: Dickerson, John P., et autres
Publié: (2025)
Characterizing AI Agents for Alignment and Governance
par: Kasirzadeh, Atoosa, et autres
Publié: (2025)
par: Kasirzadeh, Atoosa, et autres
Publié: (2025)
The SMART+ Framework for AI Systems
par: Kandikatla, Laxmiraju, et autres
Publié: (2025)
par: Kandikatla, Laxmiraju, et autres
Publié: (2025)
Prescriptive Artificial Intelligence: A Formal Paradigm for Auditing Human Decisions Under Uncertainty
par: Passos, Pedro, et autres
Publié: (2025)
par: Passos, Pedro, et autres
Publié: (2025)
The Ethical Implications of AI in Creative Industries: A Focus on AI-Generated Art
par: Khatiwada, Prerana, et autres
Publié: (2025)
par: Khatiwada, Prerana, et autres
Publié: (2025)
AI and the Problem of Knowledge Collapse
par: Peterson, Andrew J.
Publié: (2024)
par: Peterson, Andrew J.
Publié: (2024)
Autonomous AI and Ownership Rules
par: Fagan, Frank
Publié: (2026)
par: Fagan, Frank
Publié: (2026)
Structured AI Decision-Making in Disaster Management
par: Dcruz, Julian Gerald, et autres
Publié: (2025)
par: Dcruz, Julian Gerald, et autres
Publié: (2025)
Generative AI in Science: Applications, Challenges, and Emerging Questions
par: Harries, Ryan, et autres
Publié: (2025)
par: Harries, Ryan, et autres
Publié: (2025)
Ensuring Both Positivity and Stability Using Sector-Bounded Nonlinearity for Systems with Neural Network Controllers
par: Hedesh, Hamidreza Montazeri, et autres
Publié: (2024)
par: Hedesh, Hamidreza Montazeri, et autres
Publié: (2024)
Systematic Hazard Analysis for Frontier AI using STPA
par: Mylius, Simon
Publié: (2025)
par: Mylius, Simon
Publié: (2025)
Understanding Gen Alpha Digital Language: Evaluation of LLM Safety Systems for Content Moderation
par: Mehta, Manisha, et autres
Publié: (2025)
par: Mehta, Manisha, et autres
Publié: (2025)
The human biological advantage over AI
par: Stewart, William
Publié: (2025)
par: Stewart, William
Publié: (2025)
SCOR: A Framework for Responsible AI Innovation in Digital Ecosystems
par: Torkestani, Mohammad Saleh, et autres
Publié: (2025)
par: Torkestani, Mohammad Saleh, et autres
Publié: (2025)
GAIA: A General Agency Interaction Architecture for LLM-Human B2B Negotiation & Screening
par: Zhao, Siming, et autres
Publié: (2025)
par: Zhao, Siming, et autres
Publié: (2025)
Towards Friendly AI: A Comprehensive Review and New Perspectives on Human-AI Alignment
par: Sun, Qiyang, et autres
Publié: (2024)
par: Sun, Qiyang, et autres
Publié: (2024)
AI Literacy in K-12 and Higher Education in the Wake of Generative AI: An Integrative Review
par: Gu, Xingjian, et autres
Publié: (2025)
par: Gu, Xingjian, et autres
Publié: (2025)
BCause: Human-AI collaboration to improve hybrid mapping and ideation in argumentation-grounded deliberation
par: Anastasiou, Lucas, et autres
Publié: (2025)
par: Anastasiou, Lucas, et autres
Publié: (2025)
Estimating the Increase in Emissions caused by AI-augmented Search
par: Vanderbauwhede, Wim
Publié: (2024)
par: Vanderbauwhede, Wim
Publié: (2024)
Agentic AI for Scaling Diagnosis and Care in Neurodegenerative Disease
par: Breithaupt, Andrew G., et autres
Publié: (2025)
par: Breithaupt, Andrew G., et autres
Publié: (2025)
Cognitive Amplification vs Cognitive Delegation in Human-AI Systems: A Metric Framework
par: Di Santi, Eduardo
Publié: (2026)
par: Di Santi, Eduardo
Publié: (2026)
AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-Produce Supply Chains
par: Vakhnovskyi, Andrii
Publié: (2026)
par: Vakhnovskyi, Andrii
Publié: (2026)
Strategizing Equitable Transit Evacuations: A Data-Driven Reinforcement Learning Approach
par: Tang, Fang, et autres
Publié: (2024)
par: Tang, Fang, et autres
Publié: (2024)
Documents similaires
-
Sentiment Simulation using Generative AI Agents
par: Tia, Melrose, et autres
Publié: (2025) -
Our Cars Can Talk: How IoT Brings AI to Vehicles
par: Agrawal, Amod Kant
Publié: (2025) -
Will Agents Replace Us? Perceptions of Autonomous Multi-Agent AI
par: Balic, Nikola
Publié: (2025) -
From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysis
par: Dietrich, Juergen
Publié: (2026) -
Pedagogical Safety in Educational Reinforcement Learning: Formalizing and Detecting Reward Hacking in AI Tutoring Systems
par: Olukola, Oluseyi, et autres
Publié: (2026)