Shutdownable Agents through POST-Agency
Fuente:
arXiv
Saved in:
| Main Author: | Thornley, Elliott |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Shutdown Problem: An AI Engineering Puzzle for Decision Theorists
by: Thornley, Elliott
Published: (2024)
by: Thornley, Elliott
Published: (2024)
Towards Shutdownable Agents via Stochastic Choice
by: Thornley, Elliott, et al.
Published: (2024)
by: Thornley, Elliott, et al.
Published: (2024)
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
by: Cullen, Carissa, et al.
Published: (2026)
by: Cullen, Carissa, et al.
Published: (2026)
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
by: Moya, Christian, et al.
Published: (2026)
by: Moya, Christian, et al.
Published: (2026)
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
by: Williams, Kai, et al.
Published: (2025)
by: Williams, Kai, et al.
Published: (2025)
Shutdown Safety Valves for Advanced AI
by: Conitzer, Vincent
Published: (2026)
by: Conitzer, Vincent
Published: (2026)
Incomplete Tasks Induce Shutdown Resistance in Some Frontier LLMs
by: Schlatter, Jeremy, et al.
Published: (2025)
by: Schlatter, Jeremy, et al.
Published: (2025)
Regulating the Agency of LLM-based Agents
by: Boddy, Seán, et al.
Published: (2025)
by: Boddy, Seán, et al.
Published: (2025)
Prediction and Empowerment: A Theory of Agency through Bridge Interfaces
by: Csaky, Richard
Published: (2026)
by: Csaky, Richard
Published: (2026)
Agency Is Frame-Dependent
by: Abel, David, et al.
Published: (2025)
by: Abel, David, et al.
Published: (2025)
Evaluating Language Model Agency through Negotiations
by: Davidson, Tim R., et al.
Published: (2024)
by: Davidson, Tim R., et al.
Published: (2024)
AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts
by: Li, Keyu, et al.
Published: (2026)
by: Li, Keyu, et al.
Published: (2026)
AIs and Humans with Agency
by: Mumford, David
Published: (2026)
by: Mumford, David
Published: (2026)
LIMI: Less is More for Agency
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
A Knowledge-Informed Large Language Model Framework for U.S. Nuclear Power Plant Shutdown Initiating Event Classification for Probabilistic Risk Assessment
by: Xian, Min, et al.
Published: (2024)
by: Xian, Min, et al.
Published: (2024)
Active Inference as a Model of Agency
by: Da Costa, Lancelot, et al.
Published: (2024)
by: Da Costa, Lancelot, et al.
Published: (2024)
Internalizing Agency from Reflective Experience
by: Ge, Rui, et al.
Published: (2026)
by: Ge, Rui, et al.
Published: (2026)
Agency in the Age of AI
by: Swarup, Samarth
Published: (2025)
by: Swarup, Samarth
Published: (2025)
Beyond Automation: Socratic AI, Epistemic Agency, and the Implications of the Emergence of Orchestrated Multi-Agent Learning Architectures
by: Degen, Peer-Benedikt, et al.
Published: (2025)
by: Degen, Peer-Benedikt, et al.
Published: (2025)
Agency in Artificial Intelligence Systems
by: Das, Parashar
Published: (2025)
by: Das, Parashar
Published: (2025)
The STAR-XAI Protocol: A Framework for Inducing and Verifying Agency, Reasoning, and Reliability in AI Agents
by: Guasch, Antoni, et al.
Published: (2025)
by: Guasch, Antoni, et al.
Published: (2025)
Value-Sensitive AI for Prayer: Balancing the Agencies Between Human and AI Agents in Spiritual Context
by: Kwon, Soonho, et al.
Published: (2026)
by: Kwon, Soonho, et al.
Published: (2026)
Improving Language Agents through BREW
by: Kirtania, Shashank, et al.
Published: (2025)
by: Kirtania, Shashank, et al.
Published: (2025)
Agent Mentor: Framing Agent Knowledge through Semantic Trajectory Analysis
by: Ben-Gigi, Roi, et al.
Published: (2026)
by: Ben-Gigi, Roi, et al.
Published: (2026)
Artificial Intelligent Disobedience: Rethinking the Agency of Our Artificial Teammates
by: Mirsky, Reuth
Published: (2025)
by: Mirsky, Reuth
Published: (2025)
Active Inference: A method for Phenotyping Agency in AI systems?
by: Wilson, Philip, et al.
Published: (2026)
by: Wilson, Philip, et al.
Published: (2026)
Autonomy and Agency in Agentic AI: Architectural Tactics for Regulated Contexts
by: Safin, Damir, et al.
Published: (2026)
by: Safin, Damir, et al.
Published: (2026)
Reflection-Bench: Evaluating Epistemic Agency in Large Language Models
by: Li, Lingyu, et al.
Published: (2024)
by: Li, Lingyu, et al.
Published: (2024)
Situational Agency: The Framework for Designing Behavior in Agent-based art
by: Huang, Ary-Yue, et al.
Published: (2025)
by: Huang, Ary-Yue, et al.
Published: (2025)
Agency Among Agents: Designing with Hypertextual Friction in the Algorithmic Web
by: Liu, Sophia, et al.
Published: (2025)
by: Liu, Sophia, et al.
Published: (2025)
daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
by: Jiang, Mohan, et al.
Published: (2026)
by: Jiang, Mohan, et al.
Published: (2026)
Agency, Affordances, and Enculturation of Augmentation Technologies
by: Duin, Ann Hill, et al.
Published: (2025)
by: Duin, Ann Hill, et al.
Published: (2025)
A Mathematical Theory of Agency and Intelligence
by: Hafez, Wael, et al.
Published: (2026)
by: Hafez, Wael, et al.
Published: (2026)
MorphAgent: Empowering Agents through Self-Evolving Profiles and Decentralized Collaboration
by: Lu, Siyuan, et al.
Published: (2024)
by: Lu, Siyuan, et al.
Published: (2024)
The Comprehension-Gated Agent Economy: A Robustness-First Architecture for AI Economic Agency
by: Baxi, Rahul
Published: (2026)
by: Baxi, Rahul
Published: (2026)
POST: Prior-Observation Adversarial Learning of Spatio-Temporal Associations for Multivariate Time Series Anomaly Detection
by: Zhang, Suofei, et al.
Published: (2026)
by: Zhang, Suofei, et al.
Published: (2026)
AIR: Improving Agent Safety through Incident Response
by: Xiao, Zibo, et al.
Published: (2026)
by: Xiao, Zibo, et al.
Published: (2026)
Seed1.8 Model Card: Towards Generalized Real-World Agency
by: Seed, Bytedance
Published: (2026)
by: Seed, Bytedance
Published: (2026)
Intrinsic Memory Agents: Heterogeneous Multi-Agent LLM Systems through Structured Contextual Memory
by: Yuen, Sizhe, et al.
Published: (2025)
by: Yuen, Sizhe, et al.
Published: (2025)
StepMathAgent: A Step-Wise Agent for Evaluating Mathematical Processes through Tree-of-Error
by: Yang, Shu-Xun, et al.
Published: (2025)
by: Yang, Shu-Xun, et al.
Published: (2025)
Similar Items
-
The Shutdown Problem: An AI Engineering Puzzle for Decision Theorists
by: Thornley, Elliott
Published: (2024) -
Towards Shutdownable Agents via Stochastic Choice
by: Thornley, Elliott, et al.
Published: (2024) -
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
by: Cullen, Carissa, et al.
Published: (2026) -
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
by: Moya, Christian, et al.
Published: (2026) -
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
by: Williams, Kai, et al.
Published: (2025)