Open Problems in Technical AI Governance
Fuente:
arXiv
Saved in:
| Main Authors: | Reuel, Anka, Bucknall, Ben, Casper, Stephen, Fist, Tim, Soder, Lisa, Aarne, Onni, Hammond, Lewis, Ibrahim, Lujain, Chan, Alan, Wills, Peter, Anderljung, Markus, Garfinkel, Ben, Heim, Lennart, Trask, Andrew, Mukobi, Gabriel, Schaeffer, Rylan, Baker, Mauricio, Hooker, Sara, Solaiman, Irene, Luccioni, Alexandra Sasha, Rajkumar, Nitarshan, Moës, Nicolas, Ladish, Jeffrey, Bau, David, Bricman, Paul, Guha, Neel, Newman, Jessica, Bengio, Yoshua, South, Tobin, Pentland, Alex, Koyejo, Sanmi, Kochenderfer, Mykel J., Trager, Robert |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Position Paper: Technical Research and Talent is Needed for Effective AI Governance
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
by: Haupt, Andreas, et al.
Published: (2026)
by: Haupt, Andreas, et al.
Published: (2026)
More than Marketing? On the Information Value of AI Benchmarks for Practitioners
by: Hardy, Amelia, et al.
Published: (2024)
by: Hardy, Amelia, et al.
Published: (2024)
International Security Applications of Flexible Hardware-Enabled Guarantees
by: Aarne, Onni, et al.
Published: (2025)
by: Aarne, Onni, et al.
Published: (2025)
Technical Options for Flexible Hardware-Enabled Guarantees
by: Petrie, James, et al.
Published: (2025)
by: Petrie, James, et al.
Published: (2025)
AI Integrity: Defending Against Backdoors and Secret Loyalties
by: Banerjee, Dave, et al.
Published: (2026)
by: Banerjee, Dave, et al.
Published: (2026)
BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Responsible AI in the Global Context: Maturity Model and Survey
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
by: Bucknall, Ben, et al.
Published: (2025)
by: Bucknall, Ben, et al.
Published: (2025)
From Principles to Rules: A Regulatory Approach for Frontier AI
by: Schuett, Jonas, et al.
Published: (2024)
by: Schuett, Jonas, et al.
Published: (2024)
Escalation Risks from Language Models in Military and Diplomatic Decision-Making
by: Rivera, Juan-Pablo, et al.
Published: (2024)
by: Rivera, Juan-Pablo, et al.
Published: (2024)
Analyzing And Editing Inner Mechanisms Of Backdoored Language Models
by: Lamparth, Max, et al.
Published: (2023)
by: Lamparth, Max, et al.
Published: (2023)
Fairness in Reinforcement Learning: A Survey
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Flexible Hardware-Enabled Guarantees for AI Compute
by: Petrie, James, et al.
Published: (2025)
by: Petrie, James, et al.
Published: (2025)
Generative AI Needs Adaptive Governance
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Measurement to Meaning: A Validity-Centered Framework for AI Evaluation
by: Salaudeen, Olawale, et al.
Published: (2025)
by: Salaudeen, Olawale, et al.
Published: (2025)
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization
by: Chaubard, Francois, et al.
Published: (2025)
by: Chaubard, Francois, et al.
Published: (2025)
Societal Adaptation to Advanced AI
by: Bernardi, Jamie, et al.
Published: (2024)
by: Bernardi, Jamie, et al.
Published: (2024)
An Adaptive Responsible AI Governance Framework for Decentralized Organizations
by: Meimandi, Kiana Jafari, et al.
Published: (2025)
by: Meimandi, Kiana Jafari, et al.
Published: (2025)
The Synergy Between Optimal Transport Theory and Multi-Agent Reinforcement Learning
by: Baheri, Ali, et al.
Published: (2024)
by: Baheri, Ali, et al.
Published: (2024)
Brahe: A Modern Astrodynamics Library for Research and Engineering Applications
by: Eddy, Duncan, et al.
Published: (2026)
by: Eddy, Duncan, et al.
Published: (2026)
Audit Cards: Contextualizing AI Evaluations
by: Staufer, Leon, et al.
Published: (2025)
by: Staufer, Leon, et al.
Published: (2025)
Why Has Predicting Downstream Capabilities of Frontier AI Models with Scale Remained Elusive?
by: Schaeffer, Rylan, et al.
Published: (2024)
by: Schaeffer, Rylan, et al.
Published: (2024)
Fantastic Bugs and Where to Find Them in AI Benchmarks
by: Truong, Sang, et al.
Published: (2025)
by: Truong, Sang, et al.
Published: (2025)
In-Context Learning of Energy Functions
by: Schaeffer, Rylan, et al.
Published: (2024)
by: Schaeffer, Rylan, et al.
Published: (2024)
Trajectory Optimization for Adaptive Informative Path Planning with Multimodal Sensing
by: Ott, Joshua, et al.
Published: (2024)
by: Ott, Joshua, et al.
Published: (2024)
Efficient Multiagent Planning via Shared Action Suggestions
by: Asmar, Dylan M., et al.
Published: (2024)
by: Asmar, Dylan M., et al.
Published: (2024)
Conditional Deep Generative Models for Belief State Planning
by: Bigeard, Antoine, et al.
Published: (2025)
by: Bigeard, Antoine, et al.
Published: (2025)
An Iterative Bayesian Approach for System Identification based on Linear Gaussian Models
by: Tzikas, Alexandros E., et al.
Published: (2025)
by: Tzikas, Alexandros E., et al.
Published: (2025)
Position Paper: Model Access should be a Key Concern in AI Governance
by: Kembery, Edward, et al.
Published: (2024)
by: Kembery, Edward, et al.
Published: (2024)
Towards interactive evaluations for interaction harms in human-AI systems
by: Ibrahim, Lujain, et al.
Published: (2024)
by: Ibrahim, Lujain, et al.
Published: (2024)
Reasons to Doubt the Impact of AI Risk Evaluations
by: Mukobi, Gabriel
Published: (2024)
by: Mukobi, Gabriel
Published: (2024)
Of mermaids and monsters: Transgender history and the boundaries of the human in eighteenth‐ and early‐nineteenth‐century Britain
by: Onni Gust
Published: (2024)
by: Onni Gust
Published: (2024)
Beyond Gradient Averaging in Parallel Optimization: Improved Robustness through Gradient Agreement Filtering
by: Chaubard, Francois, et al.
Published: (2024)
by: Chaubard, Francois, et al.
Published: (2024)
Optimal Ground Station Selection for Low-Earth Orbiting Satellites
by: Eddy, Duncan, et al.
Published: (2024)
by: Eddy, Duncan, et al.
Published: (2024)
Hierarchical Framework for Optimizing Wildfire Surveillance and Suppression using Human-Autonomous Teaming
by: Al-Husseini, Mahdi, et al.
Published: (2024)
by: Al-Husseini, Mahdi, et al.
Published: (2024)
Fault-Aware MPC for Robotic Fleet Communications Scheduling
by: Schreiber, Carlo, et al.
Published: (2026)
by: Schreiber, Carlo, et al.
Published: (2026)
Satisfiability.jl: Satisfiability Modulo Theories in Julia
by: Soroka, Emiko, et al.
Published: (2023)
by: Soroka, Emiko, et al.
Published: (2023)
Optimizing Falsification for Learning-Based Control Systems: A Multi-Fidelity Bayesian Approach
by: Shahrooei, Zahra, et al.
Published: (2024)
by: Shahrooei, Zahra, et al.
Published: (2024)
Inferring Traffic Models in Terminal Airspace from Flight Tracks and Procedures
by: Jung, Soyeon, et al.
Published: (2023)
by: Jung, Soyeon, et al.
Published: (2023)
Similar Items
-
Position Paper: Technical Research and Talent is Needed for Effective AI Governance
by: Reuel, Anka, et al.
Published: (2024) -
Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
by: Haupt, Andreas, et al.
Published: (2026) -
More than Marketing? On the Information Value of AI Benchmarks for Practitioners
by: Hardy, Amelia, et al.
Published: (2024) -
International Security Applications of Flexible Hardware-Enabled Guarantees
by: Aarne, Onni, et al.
Published: (2025) -
Technical Options for Flexible Hardware-Enabled Guarantees
by: Petrie, James, et al.
Published: (2025)