Saved in:
| Main Authors: | Thornley, Elliott, Roman, Alexander, Ziakas, Christos, Ho, Leyton, Thomson, Louis |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.00805 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
by: Cullen, Carissa, et al.
Published: (2026)
by: Cullen, Carissa, et al.
Published: (2026)
Shutdownable Agents through POST-Agency
by: Thornley, Elliott
Published: (2025)
by: Thornley, Elliott
Published: (2025)
The Shutdown Problem: An AI Engineering Puzzle for Decision Theorists
by: Thornley, Elliott
Published: (2024)
by: Thornley, Elliott
Published: (2024)
VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision-Language Models
by: Ziakas, Christos, et al.
Published: (2025)
by: Ziakas, Christos, et al.
Published: (2025)
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
by: Moya, Christian, et al.
Published: (2026)
by: Moya, Christian, et al.
Published: (2026)
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
by: Williams, Kai, et al.
Published: (2025)
by: Williams, Kai, et al.
Published: (2025)
Shutdown Safety Valves for Advanced AI
by: Conitzer, Vincent
Published: (2026)
by: Conitzer, Vincent
Published: (2026)
Incomplete Tasks Induce Shutdown Resistance in Some Frontier LLMs
by: Schlatter, Jeremy, et al.
Published: (2025)
by: Schlatter, Jeremy, et al.
Published: (2025)
Understanding Iterative Combinatorial Auction Designs via Multi-Agent Reinforcement Learning
by: d'Eon, Greg, et al.
Published: (2024)
by: d'Eon, Greg, et al.
Published: (2024)
Utilitarian Algorithm Configuration for Infinite Parameter Spaces
by: Graham, Devon, et al.
Published: (2024)
by: Graham, Devon, et al.
Published: (2024)
Annealing Optimization for Progressive Learning with Stochastic Approximation
by: Mavridis, Christos, et al.
Published: (2022)
by: Mavridis, Christos, et al.
Published: (2022)
Understanding Understanding: A Pragmatic Framework Motivated by Large Language Models
by: Leyton-Brown, Kevin, et al.
Published: (2024)
by: Leyton-Brown, Kevin, et al.
Published: (2024)
Orchestral AI: A Framework for Agent Orchestration
by: Roman, Alexander, et al.
Published: (2026)
by: Roman, Alexander, et al.
Published: (2026)
Practical, Utilitarian Algorithm Configuration
by: Graham, Devon, et al.
Published: (2025)
by: Graham, Devon, et al.
Published: (2025)
Differentiating Choices via Commonality for Multiple-Choice Question Answering
by: Deng, Wenqing, et al.
Published: (2024)
by: Deng, Wenqing, et al.
Published: (2024)
Grounding Generated Videos in Feasible Plans via World Models
by: Ziakas, Christos, et al.
Published: (2026)
by: Ziakas, Christos, et al.
Published: (2026)
Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols
by: Griffin, Charlie, et al.
Published: (2024)
by: Griffin, Charlie, et al.
Published: (2024)
Rational Tuning of LLM Cascades via Probabilistic Modeling
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Predictive Probability Density Mapping for Search and Rescue Using An Agent-Based Approach with Sparse Data
by: Ewers, Jan-Hendrik, et al.
Published: (2024)
by: Ewers, Jan-Hendrik, et al.
Published: (2024)
Herd: Using multiple, smaller LLMs to match the performances of proprietary, large LLMs via an intelligent composer
by: Hari, Surya Narayanan, et al.
Published: (2023)
by: Hari, Surya Narayanan, et al.
Published: (2023)
Evaluating Stochasticity in Deep Research Agents
by: Zhai, Haotian, et al.
Published: (2026)
by: Zhai, Haotian, et al.
Published: (2026)
Towards a Multi-Agent Simulation of Cyber-attackers and Cyber-defenders Battles
by: Soulé, Julien, et al.
Published: (2025)
by: Soulé, Julien, et al.
Published: (2025)
Toward Efficient Convolutional Neural Networks With Structured Ternary Patterns
by: Kyrkou, Christos
Published: (2024)
by: Kyrkou, Christos
Published: (2024)
EvoAgent: Towards Automatic Multi-Agent Generation via Evolutionary Algorithms
by: Yuan, Siyu, et al.
Published: (2024)
by: Yuan, Siyu, et al.
Published: (2024)
A Knowledge-Informed Large Language Model Framework for U.S. Nuclear Power Plant Shutdown Initiating Event Classification for Probabilistic Risk Assessment
by: Xian, Min, et al.
Published: (2024)
by: Xian, Min, et al.
Published: (2024)
Modeling Choice via Self-Attention
by: Ko, Joohwan, et al.
Published: (2023)
by: Ko, Joohwan, et al.
Published: (2023)
Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
by: Winston, Caleb, et al.
Published: (2026)
by: Winston, Caleb, et al.
Published: (2026)
AI Agents and Hard Choices
by: Wang, Kangyu
Published: (2025)
by: Wang, Kangyu
Published: (2025)
Beyond Rule-Based Workflows: An Information-Flow-Orchestrated Multi-Agents Paradigm via Agent-to-Agent Communication from CORAL
by: Ren, Xinxing, et al.
Published: (2026)
by: Ren, Xinxing, et al.
Published: (2026)
Memory-Induced Supra-Competitive Outcomes Between Deep Reinforcement Learning Agents in Optimal Trade Execution
by: Koulouris, Christos Spyridon, et al.
Published: (2026)
by: Koulouris, Christos Spyridon, et al.
Published: (2026)
Prompt Baking
by: Bhargava, Aman, et al.
Published: (2024)
by: Bhargava, Aman, et al.
Published: (2024)
Instrumental Choices: Measuring the Propensity of LLM Agents to Pursue Instrumental Behaviors
by: Wiedermann-Möller, Jonas, et al.
Published: (2026)
by: Wiedermann-Möller, Jonas, et al.
Published: (2026)
UNSAT Solver Synthesis via Monte Carlo Forest Search
by: Cameron, Chris, et al.
Published: (2022)
by: Cameron, Chris, et al.
Published: (2022)
Structured Cognitive Loop for Behavioral Intelligence in Large Language Model Agents
by: Kim, Myung Ho
Published: (2025)
by: Kim, Myung Ho
Published: (2025)
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
by: Lee, Yooseop, et al.
Published: (2025)
by: Lee, Yooseop, et al.
Published: (2025)
Evaluating Agents using Social Choice Theory
by: Lanctot, Marc, et al.
Published: (2023)
by: Lanctot, Marc, et al.
Published: (2023)
Towards Autonomous Sustainability Assessment via Multimodal AI Agents
by: Zhang, Zhihan, et al.
Published: (2025)
by: Zhang, Zhihan, et al.
Published: (2025)
KompeteAI: Accelerated Autonomous Multi-Agent System for End-to-End Pipeline Generation for Machine Learning Problems
by: Kulibaba, Stepan, et al.
Published: (2025)
by: Kulibaba, Stepan, et al.
Published: (2025)
TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Economic Evaluation of LLMs
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Similar Items
-
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
by: Cullen, Carissa, et al.
Published: (2026) -
Shutdownable Agents through POST-Agency
by: Thornley, Elliott
Published: (2025) -
The Shutdown Problem: An AI Engineering Puzzle for Decision Theorists
by: Thornley, Elliott
Published: (2024) -
VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision-Language Models
by: Ziakas, Christos, et al.
Published: (2025) -
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
by: Moya, Christian, et al.
Published: (2026)