AI and Human Oversight: A Risk-Based Framework for Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Kandikatla, Laxmiraju, Radeljic, Branislav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The SMART+ Framework for AI Systems
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025)
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025)
Genocide by Algorithm in Gaza: Artificial Intelligence, Countervailing Responsibility, and the Corruption of Public Discourse
di: Radeljic, Branislav
Pubblicazione: (2026)
di: Radeljic, Branislav
Pubblicazione: (2026)
Unheard in the Digital Age: Rethinking AI Bias and Speech Diversity
di: Amaechi-Okorie, Onyedikachi Hope, et al.
Pubblicazione: (2026)
di: Amaechi-Okorie, Onyedikachi Hope, et al.
Pubblicazione: (2026)
Levers of Power in the Field of AI
di: Mackenzie, Tammy, et al.
Pubblicazione: (2025)
di: Mackenzie, Tammy, et al.
Pubblicazione: (2025)
Governance and Regulation of Artificial Intelligence in Developing Countries: A Case Study of Nigeria
di: Okoro, Uloma, et al.
Pubblicazione: (2026)
di: Okoro, Uloma, et al.
Pubblicazione: (2026)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
di: Gaube, Susanne, et al.
Pubblicazione: (2026)
di: Gaube, Susanne, et al.
Pubblicazione: (2026)
Between Innovation and Oversight: A Cross-Regional Study of AI Risk Management Frameworks in the EU, U.S., UK, and China
di: Al-Maamari, Amir
Pubblicazione: (2025)
di: Al-Maamari, Amir
Pubblicazione: (2025)
Exploring Moral Exercises for Human Oversight of AI systems: Insights from Three Pilot Studies
di: Crafa, Silvia, et al.
Pubblicazione: (2025)
di: Crafa, Silvia, et al.
Pubblicazione: (2025)
Oversight Structures for Agentic AI in Public-Sector Organizations
di: Schmitz, Chris, et al.
Pubblicazione: (2025)
di: Schmitz, Chris, et al.
Pubblicazione: (2025)
Human Control Is the Anchor, Not the Answer: Early Divergence of Oversight in Agentic AI Communities
di: Shi, Hanjing, et al.
Pubblicazione: (2026)
di: Shi, Hanjing, et al.
Pubblicazione: (2026)
AI-Driven Document Redaction in UK Public Authorities: Implementation Gaps, Regulatory Challenges, and the Human Oversight Imperative
di: Chen, Yijun
Pubblicazione: (2025)
di: Chen, Yijun
Pubblicazione: (2025)
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026)
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026)
Beyond Procedural Compliance: Human Oversight as a Dimension of Well-being Efficacy in AI Governance
di: Xie, Yao, et al.
Pubblicazione: (2025)
di: Xie, Yao, et al.
Pubblicazione: (2025)
Understanding the Process of Human-AI Value Alignment
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
di: Barthwal, Ankur, et al.
Pubblicazione: (2025)
di: Barthwal, Ankur, et al.
Pubblicazione: (2025)
Accountability Capture: How Record-Keeping to Support AI Transparency and Accountability (Re)shapes Algorithmic Oversight
di: Chappidi, Shreya, et al.
Pubblicazione: (2025)
di: Chappidi, Shreya, et al.
Pubblicazione: (2025)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
di: Motnikar, Lenart, et al.
Pubblicazione: (2025)
di: Motnikar, Lenart, et al.
Pubblicazione: (2025)
War Elephants: Rethinking Combat AI and Human Oversight
di: Feldman, Philip, et al.
Pubblicazione: (2024)
di: Feldman, Philip, et al.
Pubblicazione: (2024)
Scaling Laws For Scalable Oversight
di: Engels, Joshua, et al.
Pubblicazione: (2025)
di: Engels, Joshua, et al.
Pubblicazione: (2025)
Humanity in the Age of AI: Reassessing 2025's Existential-Risk Narratives
di: Louadi, Mohamed El
Pubblicazione: (2025)
di: Louadi, Mohamed El
Pubblicazione: (2025)
Risk Alignment in Agentic AI Systems
di: Clatterbuck, Hayley, et al.
Pubblicazione: (2024)
di: Clatterbuck, Hayley, et al.
Pubblicazione: (2024)
The AI Alignment Paradox
di: West, Robert, et al.
Pubblicazione: (2024)
di: West, Robert, et al.
Pubblicazione: (2024)
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
di: Caputo, Nicholas A.
Pubblicazione: (2024)
di: Caputo, Nicholas A.
Pubblicazione: (2024)
AI Cards: Towards an Applied Framework for Machine-Readable AI and Risk Documentation Inspired by the EU AI Act
di: Golpayegani, Delaram, et al.
Pubblicazione: (2024)
di: Golpayegani, Delaram, et al.
Pubblicazione: (2024)
Rethinking AI Cultural Alignment
di: Bravansky, Michal, et al.
Pubblicazione: (2025)
di: Bravansky, Michal, et al.
Pubblicazione: (2025)
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
di: Sornette, Didier, et al.
Pubblicazione: (2026)
di: Sornette, Didier, et al.
Pubblicazione: (2026)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
di: Shen, Hua
Pubblicazione: (2025)
di: Shen, Hua
Pubblicazione: (2025)
A Framework for Human-AI Q-Matrix Refinement: A NeuralCDM Evaluation
di: Zhang, Ying, et al.
Pubblicazione: (2026)
di: Zhang, Ying, et al.
Pubblicazione: (2026)
Justifications for Democratizing AI Alignment and Their Prospects
di: Steingrüber, André, et al.
Pubblicazione: (2025)
di: Steingrüber, André, et al.
Pubblicazione: (2025)
Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment
di: Takahashi, Toru
Pubblicazione: (2026)
di: Takahashi, Toru
Pubblicazione: (2026)
CERN for AI: A Theoretical Framework for Autonomous Simulation-Based Artificial Intelligence Testing and Alignment
di: Bojic, Ljubisa, et al.
Pubblicazione: (2023)
di: Bojic, Ljubisa, et al.
Pubblicazione: (2023)
Antisocial Analagous Behavior, Alignment and Human Impact of Google AI Systems: Evaluating through the lens of modified Antisocial Behavior Criteria by Human Interaction, Independent LLM Analysis, and AI Self-Reflection
di: Ogilvie, Alan D.
Pubblicazione: (2024)
di: Ogilvie, Alan D.
Pubblicazione: (2024)
SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
Echoes of AI Harms: A Human-LLM Synergistic Framework for Bias-Driven Harm Anticipation
di: Tantalaki, Nicoleta, et al.
Pubblicazione: (2025)
di: Tantalaki, Nicoleta, et al.
Pubblicazione: (2025)
The ASIR Courage Model: A Phase-Dynamic Framework for Truth Transitions in Human and AI Systems
di: Kim, Hyo Jin
Pubblicazione: (2026)
di: Kim, Hyo Jin
Pubblicazione: (2026)
Societal Capacity Assessment Framework: Measuring Resilience to Inform Advanced AI Risk Management
di: Gandhi, Milan, et al.
Pubblicazione: (2025)
di: Gandhi, Milan, et al.
Pubblicazione: (2025)
A First-Principles Based Risk Assessment Framework and the IEEE P3396 Standard
di: Tong, Richard J., et al.
Pubblicazione: (2025)
di: Tong, Richard J., et al.
Pubblicazione: (2025)
Community-Led AI Integration for Wildfire Risk Assessment: A Participatory AI Literacy and Explainability Integration (PALEI) Framework in Los Angeles, CA
di: Hosseini, Sanaz Sadat, et al.
Pubblicazione: (2026)
di: Hosseini, Sanaz Sadat, et al.
Pubblicazione: (2026)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
di: Chen, Benjamin Minhao, et al.
Pubblicazione: (2026)
di: Chen, Benjamin Minhao, et al.
Pubblicazione: (2026)
Societal Alignment Frameworks Can Improve LLM Alignment
di: Stańczak, Karolina, et al.
Pubblicazione: (2025)
di: Stańczak, Karolina, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The SMART+ Framework for AI Systems
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025) -
Genocide by Algorithm in Gaza: Artificial Intelligence, Countervailing Responsibility, and the Corruption of Public Discourse
di: Radeljic, Branislav
Pubblicazione: (2026) -
Unheard in the Digital Age: Rethinking AI Bias and Speech Diversity
di: Amaechi-Okorie, Onyedikachi Hope, et al.
Pubblicazione: (2026) -
Levers of Power in the Field of AI
di: Mackenzie, Tammy, et al.
Pubblicazione: (2025) -
Governance and Regulation of Artificial Intelligence in Developing Countries: A Case Study of Nigeria
di: Okoro, Uloma, et al.
Pubblicazione: (2026)