The Shutdown Problem: An AI Engineering Puzzle for Decision Theorists
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Thornley, Elliott |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Shutdownable Agents through POST-Agency
von: Thornley, Elliott
Veröffentlicht: (2025)
von: Thornley, Elliott
Veröffentlicht: (2025)
Towards Shutdownable Agents via Stochastic Choice
von: Thornley, Elliott, et al.
Veröffentlicht: (2024)
von: Thornley, Elliott, et al.
Veröffentlicht: (2024)
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
von: Cullen, Carissa, et al.
Veröffentlicht: (2026)
von: Cullen, Carissa, et al.
Veröffentlicht: (2026)
Shutdown Safety Valves for Advanced AI
von: Conitzer, Vincent
Veröffentlicht: (2026)
von: Conitzer, Vincent
Veröffentlicht: (2026)
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
von: Moya, Christian, et al.
Veröffentlicht: (2026)
von: Moya, Christian, et al.
Veröffentlicht: (2026)
Explaining Hitori Puzzles: Neurosymbolic Proof Staging for Sequential Decisions
von: Pacheco, Maria Leonor, et al.
Veröffentlicht: (2025)
von: Pacheco, Maria Leonor, et al.
Veröffentlicht: (2025)
Incomplete Tasks Induce Shutdown Resistance in Some Frontier LLMs
von: Schlatter, Jeremy, et al.
Veröffentlicht: (2025)
von: Schlatter, Jeremy, et al.
Veröffentlicht: (2025)
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
von: Williams, Kai, et al.
Veröffentlicht: (2025)
von: Williams, Kai, et al.
Veröffentlicht: (2025)
Towards Clinical AI Fairness: Filling Gaps in the Puzzle
von: Liu, Mingxuan, et al.
Veröffentlicht: (2024)
von: Liu, Mingxuan, et al.
Veröffentlicht: (2024)
Executable Archaeology: Reanimating the Logic Theorist from its IPL-V Source
von: Shrager, Jeff
Veröffentlicht: (2026)
von: Shrager, Jeff
Veröffentlicht: (2026)
GRAD-SUM: Leveraging Gradient Summarization for Optimal Prompt Engineering
von: Austin, Derek, et al.
Veröffentlicht: (2024)
von: Austin, Derek, et al.
Veröffentlicht: (2024)
PuzzlePlex: Benchmarking Foundation Models on Reasoning and Planning with Puzzles
von: Long, Yitao, et al.
Veröffentlicht: (2025)
von: Long, Yitao, et al.
Veröffentlicht: (2025)
Bongard in Wonderland: Visual Puzzles that Still Make AI Go Mad?
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
Temporal Fairness in Decision Making Problems
von: Torres, Manuel R., et al.
Veröffentlicht: (2024)
von: Torres, Manuel R., et al.
Veröffentlicht: (2024)
ArabIcros: AI-Powered Arabic Crossword Puzzle Generation for Educational Applications
von: Zeinalipour, Kamyar, et al.
Veröffentlicht: (2023)
von: Zeinalipour, Kamyar, et al.
Veröffentlicht: (2023)
Evidence-Driven Decision Support for AI Model Selection in Research Software Engineering
von: Joonbakhsh, Alireza, et al.
Veröffentlicht: (2025)
von: Joonbakhsh, Alireza, et al.
Veröffentlicht: (2025)
Moderating Model Marketplaces: Platform Governance Puzzles for AI Intermediaries
von: Gorwa, Robert, et al.
Veröffentlicht: (2023)
von: Gorwa, Robert, et al.
Veröffentlicht: (2023)
Advancing Problem-Based Learning in Biomedical Engineering in the Era of Generative AI
von: Nnamdi, Micky C., et al.
Veröffentlicht: (2025)
von: Nnamdi, Micky C., et al.
Veröffentlicht: (2025)
Benchmarking Content-Based Puzzle Solvers on Corrupted Jigsaw Puzzles
von: Dirauf, Richard, et al.
Veröffentlicht: (2025)
von: Dirauf, Richard, et al.
Veröffentlicht: (2025)
Generating Creative Chess Puzzles
von: Feng, Xidong, et al.
Veröffentlicht: (2025)
von: Feng, Xidong, et al.
Veröffentlicht: (2025)
From Frustration to Fun: An Adaptive Problem-Solving Puzzle Game Powered by Genetic Algorithm
von: McConnell, Matthew, et al.
Veröffentlicht: (2025)
von: McConnell, Matthew, et al.
Veröffentlicht: (2025)
A Puzzle-Based Dataset for Natural Language Inference
von: Szomiu, Roxana, et al.
Veröffentlicht: (2021)
von: Szomiu, Roxana, et al.
Veröffentlicht: (2021)
Preference Elicitation for Step-Wise Explanations in Logic Puzzles
von: Foschini, Marco, et al.
Veröffentlicht: (2025)
von: Foschini, Marco, et al.
Veröffentlicht: (2025)
Mathematical Definition and Systematization of Puzzle Rules
von: Maeda, Itsuki, et al.
Veröffentlicht: (2024)
von: Maeda, Itsuki, et al.
Veröffentlicht: (2024)
The Token Games: Evaluating Language Model Reasoning with Puzzle Duels
von: Henniger, Simon, et al.
Veröffentlicht: (2026)
von: Henniger, Simon, et al.
Veröffentlicht: (2026)
SHACL2FOL: An FOL Toolkit for SHACL Decision Problems
von: Pareti, Paolo
Veröffentlicht: (2024)
von: Pareti, Paolo
Veröffentlicht: (2024)
Unified Bayesian Frameworks for Multi-criteria Decision-making Problems
von: Mohammadi, Majid
Veröffentlicht: (2022)
von: Mohammadi, Majid
Veröffentlicht: (2022)
A Knowledge-Informed Large Language Model Framework for U.S. Nuclear Power Plant Shutdown Initiating Event Classification for Probabilistic Risk Assessment
von: Xian, Min, et al.
Veröffentlicht: (2024)
von: Xian, Min, et al.
Veröffentlicht: (2024)
Synthetic Trust Attacks: Modeling How Generative AI Manipulates Human Decisions in Social Engineering Fraud
von: Ashraf, Muhammad Tahir
Veröffentlicht: (2026)
von: Ashraf, Muhammad Tahir
Veröffentlicht: (2026)
A Workflow for Full Traceability of AI Decisions
von: Wenzel, Julius, et al.
Veröffentlicht: (2025)
von: Wenzel, Julius, et al.
Veröffentlicht: (2025)
Stateless Decision Memory for Enterprise AI Agents
von: Srinivasan, Vasundra
Veröffentlicht: (2026)
von: Srinivasan, Vasundra
Veröffentlicht: (2026)
Architectural Design Decisions in AI Agent Harnesses
von: Wei, Hu
Veröffentlicht: (2026)
von: Wei, Hu
Veröffentlicht: (2026)
Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning
von: Ghosal, Deepanway, et al.
Veröffentlicht: (2024)
von: Ghosal, Deepanway, et al.
Veröffentlicht: (2024)
Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models
von: Lyu, Zesen, et al.
Veröffentlicht: (2025)
von: Lyu, Zesen, et al.
Veröffentlicht: (2025)
LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?
von: Tang, Kexian, et al.
Veröffentlicht: (2025)
von: Tang, Kexian, et al.
Veröffentlicht: (2025)
Measuring Iterative Temporal Reasoning with Time Puzzles
von: Wang, Zhengxiang, et al.
Veröffentlicht: (2026)
von: Wang, Zhengxiang, et al.
Veröffentlicht: (2026)
Evaluating Game Difficulty in Tetris Block Puzzle
von: Wang, Chun-Jui, et al.
Veröffentlicht: (2026)
von: Wang, Chun-Jui, et al.
Veröffentlicht: (2026)
PuzzleJAX: A Benchmark for Reasoning and Learning
von: Earle, Sam, et al.
Veröffentlicht: (2025)
von: Earle, Sam, et al.
Veröffentlicht: (2025)
A Quantum-Inspired Algorithm for Solving Sudoku Puzzles and the MaxCut Problem
von: Zhao, Max B., et al.
Veröffentlicht: (2025)
von: Zhao, Max B., et al.
Veröffentlicht: (2025)
PuzzleBench: A Fully Dynamic Evaluation Framework for Large Multimodal Models on Puzzle Solving
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Shutdownable Agents through POST-Agency
von: Thornley, Elliott
Veröffentlicht: (2025) -
Towards Shutdownable Agents via Stochastic Choice
von: Thornley, Elliott, et al.
Veröffentlicht: (2024) -
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
von: Cullen, Carissa, et al.
Veröffentlicht: (2026) -
Shutdown Safety Valves for Advanced AI
von: Conitzer, Vincent
Veröffentlicht: (2026) -
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
von: Moya, Christian, et al.
Veröffentlicht: (2026)