Causally Abstracted Multi-armed Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Zennaro, Fabio Massimo, Bishop, Nicholas, Dyer, Joel, Felekis, Yorgos, Calinescu, Anisoara, Wooldridge, Michael, Damoulas, Theodoros |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Distributionally Robust Causal Abstractions
by: Felekis, Yorgos, et al.
Published: (2025)
by: Felekis, Yorgos, et al.
Published: (2025)
Sandbagging in a Simple Survival Bandit Problem
by: Dyer, Joel, et al.
Published: (2025)
by: Dyer, Joel, et al.
Published: (2025)
Using causal abstractions to accelerate decision-making in complex bandit problems
by: Dyer, Joel, et al.
Published: (2025)
by: Dyer, Joel, et al.
Published: (2025)
Bayesian Decision Making around Experts
by: Ornia, Daniel Jarne, et al.
Published: (2025)
by: Ornia, Daniel Jarne, et al.
Published: (2025)
Causal Abstraction Learning based on the Semantic Embedding Principle
by: D'Acunto, Gabriele, et al.
Published: (2025)
by: D'Acunto, Gabriele, et al.
Published: (2025)
Automatic Differentiation of Agent-Based Models
by: Quera-Bofarull, Arnau, et al.
Published: (2025)
by: Quera-Bofarull, Arnau, et al.
Published: (2025)
Multi-Level Causal Embeddings
by: Schooltink, Willem, et al.
Published: (2026)
by: Schooltink, Willem, et al.
Published: (2026)
Emergent Risk Awareness in Rational Agents under Resource Constraints
by: Ornia, Daniel Jarne, et al.
Published: (2025)
by: Ornia, Daniel Jarne, et al.
Published: (2025)
A multi-objective combinatorial optimisation framework for large scale hierarchical population synthesis
by: Mahmood, Imran, et al.
Published: (2024)
by: Mahmood, Imran, et al.
Published: (2024)
Neural Network-Based Parameter Estimation of a Labour Market Agent-Based Model
by: Alves, M Lopes, et al.
Published: (2026)
by: Alves, M Lopes, et al.
Published: (2026)
Deceptive Exploration in Multi-armed Bandits
by: Vurankaya, I. Arda, et al.
Published: (2025)
by: Vurankaya, I. Arda, et al.
Published: (2025)
Multi-View Spectral Clustering for Graphs with Multiple View Structures
by: Tsitsikas, Yorgos, et al.
Published: (2025)
by: Tsitsikas, Yorgos, et al.
Published: (2025)
Aligning Graphical and Functional Causal Abstractions
by: Schooltink, Willem, et al.
Published: (2024)
by: Schooltink, Willem, et al.
Published: (2024)
Improving Thompson Sampling via Information Relaxation for Budgeted Multi-armed Bandits
by: Jeong, Woojin, et al.
Published: (2024)
by: Jeong, Woojin, et al.
Published: (2024)
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
by: Zhao, Yunfan, et al.
Published: (2023)
by: Zhao, Yunfan, et al.
Published: (2023)
Teleological Inference in Structural Causal Models via Intentional Interventions
by: Compagno, Dario, et al.
Published: (2026)
by: Compagno, Dario, et al.
Published: (2026)
Rethinking Reinforcement fine-tuning of LLMs: A Multi-armed Bandit Learning Perspective
by: Hu, Xiao, et al.
Published: (2026)
by: Hu, Xiao, et al.
Published: (2026)
Combinatorial Multi-armed Bandits: Arm Selection via Group Testing
by: Mukherjee, Arpan, et al.
Published: (2024)
by: Mukherjee, Arpan, et al.
Published: (2024)
Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values
by: Sharma, Shradha, et al.
Published: (2026)
by: Sharma, Shradha, et al.
Published: (2026)
SynEHRgy: Synthesizing Mixed-Type Structured Electronic Health Records using Decoder-Only Transformers
by: Karami, Hojjat, et al.
Published: (2024)
by: Karami, Hojjat, et al.
Published: (2024)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
by: Xu, Tianyi, et al.
Published: (2025)
by: Xu, Tianyi, et al.
Published: (2025)
Causal Contextual Bandits with Adaptive Context
by: Madhavan, Rahul, et al.
Published: (2024)
by: Madhavan, Rahul, et al.
Published: (2024)
The Minimal Search Space for Conditional Causal Bandits
by: Simoes, Francisco N. F. Q., et al.
Published: (2025)
by: Simoes, Francisco N. F. Q., et al.
Published: (2025)
Robust Uncertainty Quantification Using Conformalised Monte Carlo Prediction
by: Bethell, Daniel, et al.
Published: (2023)
by: Bethell, Daniel, et al.
Published: (2023)
FeatEHR-LLM: Leveraging Large Language Models for Feature Engineering in Electronic Health Records
by: Karami, Hojjat, et al.
Published: (2026)
by: Karami, Hojjat, et al.
Published: (2026)
Fixed Point Explainability
by: La Malfa, Emanuele, et al.
Published: (2025)
by: La Malfa, Emanuele, et al.
Published: (2025)
Robust Bayesian Inference for Measurement Error Misspecification: The Berkson and Classical Cases
by: Dellaporta, Charita, et al.
Published: (2023)
by: Dellaporta, Charita, et al.
Published: (2023)
Flickering Multi-Armed Bandits
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
CausalARC: Abstract Reasoning with Causal World Models
by: Maasch, Jacqueline, et al.
Published: (2025)
by: Maasch, Jacqueline, et al.
Published: (2025)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
by: Bethell, Daniel, et al.
Published: (2024)
by: Bethell, Daniel, et al.
Published: (2024)
Diffusion and Flow-based Copulas: Forgetting and Remembering Dependencies
by: Huk, David, et al.
Published: (2025)
by: Huk, David, et al.
Published: (2025)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
by: Veness, Joel, et al.
Published: (2025)
by: Veness, Joel, et al.
Published: (2025)
Certified Policy Optimisation for Nested Causal Bandits via PAC-Bayes Risk
by: Woydt, Tim, et al.
Published: (2026)
by: Woydt, Tim, et al.
Published: (2026)
End-to-end PDDL Planning with Hardcoded and Dynamic Agents
by: La Malfa, Emanuele, et al.
Published: (2025)
by: La Malfa, Emanuele, et al.
Published: (2025)
Multi-Armed Bandits With Best-Action Queries
by: Bacchiocchi, Francesco, et al.
Published: (2026)
by: Bacchiocchi, Francesco, et al.
Published: (2026)
A Scalable, Causal, and Energy Efficient Framework for Neural Decoding with Spiking Neural Networks
by: Mentzelopoulos, Georgios, et al.
Published: (2025)
by: Mentzelopoulos, Georgios, et al.
Published: (2025)
Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments
by: Zhang, Ziyuan, et al.
Published: (2025)
by: Zhang, Ziyuan, et al.
Published: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
CIVeX: Causal Intervention Verification for Language Agents
by: Rovai, Fabio
Published: (2026)
by: Rovai, Fabio
Published: (2026)
Networked Restless Multi-Arm Bandits with Reinforcement Learning
by: Zhang, Hanmo, et al.
Published: (2025)
by: Zhang, Hanmo, et al.
Published: (2025)
Similar Items
-
Distributionally Robust Causal Abstractions
by: Felekis, Yorgos, et al.
Published: (2025) -
Sandbagging in a Simple Survival Bandit Problem
by: Dyer, Joel, et al.
Published: (2025) -
Using causal abstractions to accelerate decision-making in complex bandit problems
by: Dyer, Joel, et al.
Published: (2025) -
Bayesian Decision Making around Experts
by: Ornia, Daniel Jarne, et al.
Published: (2025) -
Causal Abstraction Learning based on the Semantic Embedding Principle
by: D'Acunto, Gabriele, et al.
Published: (2025)