The AI off-switch problem as a signalling game: bounded rationality and incomparability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Benavoli, Alessio, Facchini, Alessandro, Zaffalon, Marco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025)
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025)
Connecting classical finite exchangeability to quantum theory
von: Benavoli, Alessio, et al.
Veröffentlicht: (2023)
von: Benavoli, Alessio, et al.
Veröffentlicht: (2023)
dynoGP: Deep Gaussian Processes for dynamic system identification
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025)
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025)
A tutorial on learning from preferences and choices with Gaussian Processes
von: Benavoli, Alessio, et al.
Veröffentlicht: (2024)
von: Benavoli, Alessio, et al.
Veröffentlicht: (2024)
Constraint- and Score-Based Nonlinear Granger Causality Discovery with Kernels
von: Murphy, Fiona, et al.
Veröffentlicht: (2026)
von: Murphy, Fiona, et al.
Veröffentlicht: (2026)
A Note on Bayesian Networks with Latent Root Variables
von: Zaffalon, Marco, et al.
Veröffentlicht: (2024)
von: Zaffalon, Marco, et al.
Veröffentlicht: (2024)
Automatic Causal Fairness Analysis with LLM-Generated Reporting
von: Berarducci, Alessia, et al.
Veröffentlicht: (2026)
von: Berarducci, Alessia, et al.
Veröffentlicht: (2026)
COPA: Comparing the incomparable in multi-objective model evaluation
von: Javaloy, Adrián, et al.
Veröffentlicht: (2025)
von: Javaloy, Adrián, et al.
Veröffentlicht: (2025)
Epistemology gives a Future to Complementarity in Human-AI Interactions
von: Ferrario, Andrea, et al.
Veröffentlicht: (2026)
von: Ferrario, Andrea, et al.
Veröffentlicht: (2026)
Quantum Wiener architecture for quantum reservoir computing
von: Benavoli, Alessio, et al.
Veröffentlicht: (2026)
von: Benavoli, Alessio, et al.
Veröffentlicht: (2026)
A singular Riemannian Geometry Approach to Deep Neural Networks III. Piecewise Differentiable Layers and Random Walks on $n$-dimensional Classes
von: Benfenati, Alessandro, et al.
Veröffentlicht: (2024)
von: Benfenati, Alessandro, et al.
Veröffentlicht: (2024)
On the influence of dependent features in classification problems: a game-theoretic perspective
von: Davila-Pena, Laura, et al.
Veröffentlicht: (2024)
von: Davila-Pena, Laura, et al.
Veröffentlicht: (2024)
Prediction and Causality of functional MRI and synthetic signal using a Zero-Shot Time-Series Foundation Model
von: Crimi, Alessandro, et al.
Veröffentlicht: (2025)
von: Crimi, Alessandro, et al.
Veröffentlicht: (2025)
AlphaViT: A flexible game-playing AI for multiple games and variable board sizes
von: Fujita, Kazuhisa
Veröffentlicht: (2024)
von: Fujita, Kazuhisa
Veröffentlicht: (2024)
Additive decomposition of one-dimensional signals using Transformers
von: Salti, Samuele, et al.
Veröffentlicht: (2025)
von: Salti, Samuele, et al.
Veröffentlicht: (2025)
A finite-sample bound for identifying partially observed linear switched systems from a single trajectory
von: Racz, Daniel, et al.
Veröffentlicht: (2025)
von: Racz, Daniel, et al.
Veröffentlicht: (2025)
Causal Discovery on Higher-Order Interactions
von: Zanga, Alessio, et al.
Veröffentlicht: (2025)
von: Zanga, Alessio, et al.
Veröffentlicht: (2025)
Two failure modes of deep transformers and how to avoid them: a unified theory of signal propagation at initialisation
von: Giorlandino, Alessio, et al.
Veröffentlicht: (2025)
von: Giorlandino, Alessio, et al.
Veröffentlicht: (2025)
Trading off rewards and errors in multi-armed bandits
von: Erraqabi, Akram, et al.
Veröffentlicht: (2026)
von: Erraqabi, Akram, et al.
Veröffentlicht: (2026)
Function Based Isolation Forest (FuBIF): A Unifying Framework for Interpretable Isolation-Based Anomaly Detection
von: Arcudi, Alessio, et al.
Veröffentlicht: (2025)
von: Arcudi, Alessio, et al.
Veröffentlicht: (2025)
Sink vs. diagonal patterns as mechanisms for attention switch and oversmoothing prevention
von: Súkeník, Peter, et al.
Veröffentlicht: (2026)
von: Súkeník, Peter, et al.
Veröffentlicht: (2026)
Trade-offs in Ensembling, Merging and Routing Among Parameter-Efficient Experts
von: Lotfi, Sanae, et al.
Veröffentlicht: (2026)
von: Lotfi, Sanae, et al.
Veröffentlicht: (2026)
Modelling bounded rational decision-making through Wasserstein constraints
von: Evans, Benjamin Patrick, et al.
Veröffentlicht: (2025)
von: Evans, Benjamin Patrick, et al.
Veröffentlicht: (2025)
The logic of rational graph neural networks
von: Khalife, Sammy
Veröffentlicht: (2023)
von: Khalife, Sammy
Veröffentlicht: (2023)
An adaptive approach to Bayesian Optimization with switching costs
von: Pricopie, Stefan, et al.
Veröffentlicht: (2024)
von: Pricopie, Stefan, et al.
Veröffentlicht: (2024)
The representation landscape of few-shot learning and fine-tuning in large language models
von: Doimo, Diego, et al.
Veröffentlicht: (2024)
von: Doimo, Diego, et al.
Veröffentlicht: (2024)
Solving Bayesian inverse problems with diffusion priors and off-policy RL
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
Approximate information maximization for bandit games
von: Barbier-Chebbah, Alex, et al.
Veröffentlicht: (2023)
von: Barbier-Chebbah, Alex, et al.
Veröffentlicht: (2023)
An Improved Algorithm for Learning Drifting Discrete Distributions
von: Mazzetto, Alessio
Veröffentlicht: (2024)
von: Mazzetto, Alessio
Veröffentlicht: (2024)
Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems
von: Yang, Xianjin, et al.
Veröffentlicht: (2025)
von: Yang, Xianjin, et al.
Veröffentlicht: (2025)
Unified theory of upper confidence bound policies for bandit problems targeting total reward, maximal reward, and more
von: Kikkawa, Nobuaki, et al.
Veröffentlicht: (2024)
von: Kikkawa, Nobuaki, et al.
Veröffentlicht: (2024)
On rankings in multiplayer games with an application to the game of Whist
von: Coyette, Alexis, et al.
Veröffentlicht: (2026)
von: Coyette, Alexis, et al.
Veröffentlicht: (2026)
AI Agents as Universal Task Solvers
von: Achille, Alessandro, et al.
Veröffentlicht: (2025)
von: Achille, Alessandro, et al.
Veröffentlicht: (2025)
Unified continuous-time q-learning for mean-field game and mean-field control problems
von: Wei, Xiaoli, et al.
Veröffentlicht: (2024)
von: Wei, Xiaoli, et al.
Veröffentlicht: (2024)
Small transformer architectures for task switching
von: Gros, Claudius
Veröffentlicht: (2025)
von: Gros, Claudius
Veröffentlicht: (2025)
Computations and ML for surjective rational maps
von: Karzhemanov, Ilya
Veröffentlicht: (2025)
von: Karzhemanov, Ilya
Veröffentlicht: (2025)
Planning in entropy-regularized Markov decision processes and games
von: Grill, Jean-Bastien, et al.
Veröffentlicht: (2026)
von: Grill, Jean-Bastien, et al.
Veröffentlicht: (2026)
Tackling the Accuracy-Interpretability Trade-off in a Hierarchy of Machine Learning Models for the Prediction of Extreme Heatwaves
von: Lovo, Alessandro, et al.
Veröffentlicht: (2024)
von: Lovo, Alessandro, et al.
Veröffentlicht: (2024)
When to Forget? Complexity Trade-offs in Machine Unlearning
von: Van Waerebeke, Martin, et al.
Veröffentlicht: (2025)
von: Van Waerebeke, Martin, et al.
Veröffentlicht: (2025)
Water Quality Estimation Through Machine Learning Multivariate Analysis
von: Cardia, Marco, et al.
Veröffentlicht: (2025)
von: Cardia, Marco, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025) -
Connecting classical finite exchangeability to quantum theory
von: Benavoli, Alessio, et al.
Veröffentlicht: (2023) -
dynoGP: Deep Gaussian Processes for dynamic system identification
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025) -
A tutorial on learning from preferences and choices with Gaussian Processes
von: Benavoli, Alessio, et al.
Veröffentlicht: (2024) -
Constraint- and Score-Based Nonlinear Granger Causality Discovery with Kernels
von: Murphy, Fiona, et al.
Veröffentlicht: (2026)