Managing extreme AI risks amid rapid progress
Fuente:
arXiv
Guardado en:
| Autores principales: | Bengio, Yoshua, Hinton, Geoffrey, Yao, Andrew, Song, Dawn, Abbeel, Pieter, Darrell, Trevor, Harari, Yuval Noah, Zhang, Ya-Qin, Xue, Lan, Shalev-Shwartz, Shai, Hadfield, Gillian, Clune, Jeff, Maharaj, Tegan, Hutter, Frank, Baydin, Atılım Güneş, McIlraith, Sheila, Gao, Qiqi, Acharya, Ashwin, Krueger, David, Dragan, Anca, Torr, Philip, Russell, Stuart, Kahneman, Daniel, Brauner, Jan, Mindermann, Sören |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers
por: Lifshitz, Shalev, et al.
Publicado: (2025)
por: Lifshitz, Shalev, et al.
Publicado: (2025)
Closing the Gap Between SGP4 and High-Precision Propagation via Differentiable Programming
por: Acciarini, Giacomo, et al.
Publicado: (2024)
por: Acciarini, Giacomo, et al.
Publicado: (2024)
From Reasoning to Super-Intelligence: A Search-Theoretic Perspective
por: Shalev-Shwartz, Shai, et al.
Publicado: (2025)
por: Shalev-Shwartz, Shai, et al.
Publicado: (2025)
STEVE-1: A Generative Model for Text-to-Behavior in Minecraft
por: Lifshitz, Shalev, et al.
Publicado: (2023)
por: Lifshitz, Shalev, et al.
Publicado: (2023)
Second-Order Forward-Mode Automatic Differentiation for Optimization
por: Cobb, Adam D., et al.
Publicado: (2024)
por: Cobb, Adam D., et al.
Publicado: (2024)
Gaussian Processes for Probabilistic Estimates of Earthquake Ground Shaking: A 1-D Proof-of-Concept
por: Scivier, Sam A., et al.
Publicado: (2024)
por: Scivier, Sam A., et al.
Publicado: (2024)
Research Program: Theory of Learning in Dynamical Systems
por: Hazan, Elad, et al.
Publicado: (2025)
por: Hazan, Elad, et al.
Publicado: (2025)
Contextual Plackett-Luce: An Efficient Neural Model for Probabilistic Sequence Selection under Ambiguity
por: Mizrachi, Noam, et al.
Publicado: (2026)
por: Mizrachi, Noam, et al.
Publicado: (2026)
Untangling Lariats: Subgradient Following of Variationally Penalized Objectives
por: Mo, Kai-Chia, et al.
Publicado: (2024)
por: Mo, Kai-Chia, et al.
Publicado: (2024)
Cooperative Inverse Reinforcement Learning
por: Hadfield-Menell, Dylan, et al.
Publicado: (2016)
por: Hadfield-Menell, Dylan, et al.
Publicado: (2016)
Now, Later, and Lasting: Ten Priorities for AI Research, Policy, and Practice
por: Horvitz, Eric, et al.
Publicado: (2024)
por: Horvitz, Eric, et al.
Publicado: (2024)
Why Hawks Win
por: Kahneman, Daniel
Publicado: (2007)
por: Kahneman, Daniel
Publicado: (2007)
Pluralistic Alignment Over Time
por: Klassen, Toryn Q., et al.
Publicado: (2024)
por: Klassen, Toryn Q., et al.
Publicado: (2024)
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems
por: Alamdari, Parand A., et al.
Publicado: (2026)
por: Alamdari, Parand A., et al.
Publicado: (2026)
Machine learning and information theory concepts towards an AI Mathematician
por: Bengio, Yoshua, et al.
Publicado: (2024)
por: Bengio, Yoshua, et al.
Publicado: (2024)
Probabilistic Forecasting of Radiation Exposure for Spaceflight
por: Gurav, Rutuja, et al.
Publicado: (2024)
por: Gurav, Rutuja, et al.
Publicado: (2024)
Language Models For Generalised PDDL Planning: Synthesising Sound and Programmatic Policies
por: Chen, Dillon Z., et al.
Publicado: (2025)
por: Chen, Dillon Z., et al.
Publicado: (2025)
Baking Symmetry into GFlowNets
por: Ma, George, et al.
Publicado: (2024)
por: Ma, George, et al.
Publicado: (2024)
Satisficing and Optimal Generalised Planning via Goal Regression (Extended Version)
por: Chen, Dillon Z., et al.
Publicado: (2025)
por: Chen, Dillon Z., et al.
Publicado: (2025)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
por: Alamdari, Parand A., et al.
Publicado: (2023)
por: Alamdari, Parand A., et al.
Publicado: (2023)
Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning
por: Chen, Dillon Z., et al.
Publicado: (2026)
por: Chen, Dillon Z., et al.
Publicado: (2026)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
por: Alamdari, Parand A., et al.
Publicado: (2024)
por: Alamdari, Parand A., et al.
Publicado: (2024)
Noticing the Watcher: LLM Agents Can Infer CoT Monitoring from Blocking Feedback
por: Jiralerspong, Thomas, et al.
Publicado: (2026)
por: Jiralerspong, Thomas, et al.
Publicado: (2026)
Forecasting the Ionosphere from Sparse GNSS Data with Temporal-Fusion Transformers
por: Acciarini, Giacomo, et al.
Publicado: (2025)
por: Acciarini, Giacomo, et al.
Publicado: (2025)
Better Training Data Attribution via Better Inverse Hessian-Vector Products
por: Wang, Andrew, et al.
Publicado: (2025)
por: Wang, Andrew, et al.
Publicado: (2025)
Psicología de las preferencias
por: Kahneman, Daniel y Tversky, Amos
Publicado: (1982)
por: Kahneman, Daniel y Tversky, Amos
Publicado: (1982)
Beyond Predictive Algorithms in Child Welfare
por: Moon, Erina Seh-Young, et al.
Publicado: (2024)
por: Moon, Erina Seh-Young, et al.
Publicado: (2024)
Implicit meta-learning may lead language models to trust more reliable sources
por: Krasheninnikov, Dmitrii, et al.
Publicado: (2023)
por: Krasheninnikov, Dmitrii, et al.
Publicado: (2023)
Ground-Compose-Reinforce: Grounding Language in Agentic Behaviours using Limited Data
por: Li, Andrew C., et al.
Publicado: (2025)
por: Li, Andrew C., et al.
Publicado: (2025)
Pushdown Reward Machines for Reinforcement Learning
por: Varricchione, Giovanni, et al.
Publicado: (2025)
por: Varricchione, Giovanni, et al.
Publicado: (2025)
Single-Frame Super-Resolution of Solar Magnetograms: Investigating Physics-Based Metrics & Losses
por: Jungbluth, Anna, et al.
Publicado: (2019)
por: Jungbluth, Anna, et al.
Publicado: (2019)
A Foundation Model for the Solar Dynamics Observatory
por: Walsh, James, et al.
Publicado: (2024)
por: Walsh, James, et al.
Publicado: (2024)
On Generalization for Generative Flow Networks
por: Krichel, Anas, et al.
Publicado: (2024)
por: Krichel, Anas, et al.
Publicado: (2024)
Interventional Causal Representation Learning
por: Ahuja, Kartik, et al.
Publicado: (2022)
por: Ahuja, Kartik, et al.
Publicado: (2022)
A Complexity-Based Theory of Compositionality
por: Elmoznino, Eric, et al.
Publicado: (2024)
por: Elmoznino, Eric, et al.
Publicado: (2024)
Visual symbolic mechanisms: Emergent symbol processing in vision language models
por: Assouel, Rim, et al.
Publicado: (2025)
por: Assouel, Rim, et al.
Publicado: (2025)
Relative Trajectory Balance is equivalent to Trust-PCL
por: Deleu, Tristan, et al.
Publicado: (2025)
por: Deleu, Tristan, et al.
Publicado: (2025)
Fast Monte Carlo Tree Diffusion: 100x Speedup via Parallel Sparse Planning
por: Yoon, Jaesik, et al.
Publicado: (2025)
por: Yoon, Jaesik, et al.
Publicado: (2025)
In-Context Parametric Inference: Point or Distribution Estimators?
por: Mittal, Sarthak, et al.
Publicado: (2025)
por: Mittal, Sarthak, et al.
Publicado: (2025)
GFlowNet Foundations
por: Bengio, Yoshua, et al.
Publicado: (2021)
por: Bengio, Yoshua, et al.
Publicado: (2021)
Ejemplares similares
-
Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers
por: Lifshitz, Shalev, et al.
Publicado: (2025) -
Closing the Gap Between SGP4 and High-Precision Propagation via Differentiable Programming
por: Acciarini, Giacomo, et al.
Publicado: (2024) -
From Reasoning to Super-Intelligence: A Search-Theoretic Perspective
por: Shalev-Shwartz, Shai, et al.
Publicado: (2025) -
STEVE-1: A Generative Model for Text-to-Behavior in Minecraft
por: Lifshitz, Shalev, et al.
Publicado: (2023) -
Second-Order Forward-Mode Automatic Differentiation for Optimization
por: Cobb, Adam D., et al.
Publicado: (2024)