Pure Exploration in Bandits with Linear Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Carlsson, Emil, Basu, Debabrota, Johansson, Fredrik D., Dubhashi, Devdatt |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Active Preference Learning for Ordering Items In- and Out-of-sample
by: Bergström, Herman, et al.
Published: (2024)
by: Bergström, Herman, et al.
Published: (2024)
Variational Quantum Optimization with Continuous Bandits
by: Wanner, Marc, et al.
Published: (2025)
by: Wanner, Marc, et al.
Published: (2025)
Learning Approximate and Exact Numeral Systems via Reinforcement Learning
by: Carlsson, Emil, et al.
Published: (2021)
by: Carlsson, Emil, et al.
Published: (2021)
Latent Preference Bandits
by: Mwai, Newton, et al.
Published: (2025)
by: Mwai, Newton, et al.
Published: (2025)
Latent Order Bandits
by: Carlsson, Emil, et al.
Published: (2026)
by: Carlsson, Emil, et al.
Published: (2026)
Preference-based Pure Exploration
by: Shukla, Apurv, et al.
Published: (2024)
by: Shukla, Apurv, et al.
Published: (2024)
Learning to Explore with Lagrangians for Bandits under Unknown Linear Constraints
by: Das, Udvas, et al.
Published: (2024)
by: Das, Udvas, et al.
Published: (2024)
Identifiable Latent Bandits: Leveraging observational data for personalized decision-making
by: Balcıoğlu, Ahmet Zahid, et al.
Published: (2024)
by: Balcıoğlu, Ahmet Zahid, et al.
Published: (2024)
Cultural evolution via iterated learning and communication explains efficient color naming systems
by: Carlsson, Emil, et al.
Published: (2023)
by: Carlsson, Emil, et al.
Published: (2023)
FLIPHAT: Joint Differential Privacy for High Dimensional Sparse Linear Bandits
by: Chakraborty, Sunrit, et al.
Published: (2024)
by: Chakraborty, Sunrit, et al.
Published: (2024)
FraPPE: Fast and Efficient Preference-based Pure Exploration
by: Das, Udvas, et al.
Published: (2025)
by: Das, Udvas, et al.
Published: (2025)
Learning Contextual Runtime Monitors for Safe AI-Based Autonomy
by: Luque-Cerpa, Alejandro, et al.
Published: (2026)
by: Luque-Cerpa, Alejandro, et al.
Published: (2026)
Concentrated Differential Privacy for Bandits
by: Azize, Achraf, et al.
Published: (2023)
by: Azize, Achraf, et al.
Published: (2023)
Learning Efficient Recursive Numeral Systems via Reinforcement Learning
by: Silvi, Andrea, et al.
Published: (2024)
by: Silvi, Andrea, et al.
Published: (2024)
Stochastic Online Instrumental Variable Regression: Regrets for Endogeneity and Bandit Feedback
by: Della Vecchia, Riccardo, et al.
Published: (2023)
by: Della Vecchia, Riccardo, et al.
Published: (2023)
Predicting Ground State Properties: Constant Sample Complexity and Deep Learning Algorithms
by: Wanner, Marc, et al.
Published: (2024)
by: Wanner, Marc, et al.
Published: (2024)
Pure Exploration in Asynchronous Federated Bandits
by: Wang, Zichen, et al.
Published: (2023)
by: Wang, Zichen, et al.
Published: (2023)
The Batch Complexity of Bandit Pure Exploration
by: Tuynman, Adrienne, et al.
Published: (2025)
by: Tuynman, Adrienne, et al.
Published: (2025)
FlashHead: Efficient Drop-In Replacement for the Classification Head in Language Model Inference
by: Tranheden, Wilhelm, et al.
Published: (2026)
by: Tranheden, Wilhelm, et al.
Published: (2026)
Near Optimal Pure Exploration in Logistic Bandits
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
Some Targets Are Harder to Identify than Others: Quantifying the Target-dependent Membership Leakage
by: Azize, Achraf, et al.
Published: (2024)
by: Azize, Achraf, et al.
Published: (2024)
Auditing Fairness under Model Updates: Fundamental Complexity and Property-Preserving Updates
by: Ajarra, Ayoub, et al.
Published: (2026)
by: Ajarra, Ayoub, et al.
Published: (2026)
Isoperimetry is All We Need: Langevin Posterior Sampling for RL with Sublinear Regret
by: Jorge, Emilio, et al.
Published: (2024)
by: Jorge, Emilio, et al.
Published: (2024)
Infrequent Exploration in Linear Bandits
by: Lee, Harin, et al.
Published: (2025)
by: Lee, Harin, et al.
Published: (2025)
Sublinear Algorithms for Wasserstein and Total Variation Distances: Applications to Fairness and Privacy Auditing
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback
by: Li, Zitian, et al.
Published: (2026)
by: Li, Zitian, et al.
Published: (2026)
Optimal Regret of Bernoulli Bandits under Global Differential Privacy
by: Azize, Achraf, et al.
Published: (2025)
by: Azize, Achraf, et al.
Published: (2025)
Sparse Linear Bandits with Blocking Constraints
by: Jain, Adit, et al.
Published: (2024)
by: Jain, Adit, et al.
Published: (2024)
IncomeSCM: From tabular data set to time-series simulator and causal estimation benchmark
by: Johansson, Fredrik D.
Published: (2024)
by: Johansson, Fredrik D.
Published: (2024)
Augmented Bayesian Policy Search
by: Kallel, Mahdi, et al.
Published: (2024)
by: Kallel, Mahdi, et al.
Published: (2024)
PACE: Procedural Abstractions for Communicating Efficiently
by: Thomas, Jonathan D., et al.
Published: (2024)
by: Thomas, Jonathan D., et al.
Published: (2024)
Active Fourier Auditor for Estimating Distributional Properties of ML Models
by: Ajarra, Ayoub, et al.
Published: (2024)
by: Ajarra, Ayoub, et al.
Published: (2024)
Sequential Membership Inference Attacks
by: Michel, Thomas, et al.
Published: (2026)
by: Michel, Thomas, et al.
Published: (2026)
DP-SPRT: Differentially Private Sequential Probability Ratio Tests
by: Michel, Thomas, et al.
Published: (2025)
by: Michel, Thomas, et al.
Published: (2025)
Lagrangian-based Equilibrium Propagation: generalisation to arbitrary boundary conditions & equivalence with Hamiltonian Echo Learning
by: Pourcel, Guillaume, et al.
Published: (2025)
by: Pourcel, Guillaume, et al.
Published: (2025)
Distributed Linear Bandits under Communication Constraints
by: Salgia, Sudeep, et al.
Published: (2022)
by: Salgia, Sudeep, et al.
Published: (2022)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
by: Nakamura, Shintaro, et al.
Published: (2023)
by: Nakamura, Shintaro, et al.
Published: (2023)
When Witnesses Defend: A Witness Graph Topological Layer for Adversarial Graph Learning
by: Arafat, Naheed Anjum, et al.
Published: (2024)
by: Arafat, Naheed Anjum, et al.
Published: (2024)
Dynamical-VAE-based Hindsight to Learn the Causal Dynamics of Factored-POMDPs
by: Han, Chao, et al.
Published: (2024)
by: Han, Chao, et al.
Published: (2024)
How does Your RL Agent Explore? An Optimal Transport Analysis of Occupancy Measure Trajectories
by: Nkhumise, Reabetswe M., et al.
Published: (2024)
by: Nkhumise, Reabetswe M., et al.
Published: (2024)
Similar Items
-
Active Preference Learning for Ordering Items In- and Out-of-sample
by: Bergström, Herman, et al.
Published: (2024) -
Variational Quantum Optimization with Continuous Bandits
by: Wanner, Marc, et al.
Published: (2025) -
Learning Approximate and Exact Numeral Systems via Reinforcement Learning
by: Carlsson, Emil, et al.
Published: (2021) -
Latent Preference Bandits
by: Mwai, Newton, et al.
Published: (2025) -
Latent Order Bandits
by: Carlsson, Emil, et al.
Published: (2026)