Saved in:
| Main Authors: | Kone, Cyrille, Kaufmann, Emilie, Richert, Laura |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.08127 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bandit Pareto Set Identification: the Fixed Budget Setting
by: Kone, Cyrille, et al.
Published: (2023)
by: Kone, Cyrille, et al.
Published: (2023)
Bandit Pareto Set Identification in a Multi-Output Linear Model
by: Kone, Cyrille, et al.
Published: (2025)
by: Kone, Cyrille, et al.
Published: (2025)
Pareto Set Identification With Posterior Sampling
by: Kone, Cyrille, et al.
Published: (2024)
by: Kone, Cyrille, et al.
Published: (2024)
Robust Pareto Set Identification with Contaminated Bandit Feedback
by: Korkmaz, İlter Onat, et al.
Published: (2022)
by: Korkmaz, İlter Onat, et al.
Published: (2022)
Optimal Posterior Sampling for Policy Identification in Tabular Markov Decision Processes
by: Kone, Cyrille, et al.
Published: (2026)
by: Kone, Cyrille, et al.
Published: (2026)
Best-Arm Identification in Unimodal Bandits
by: Poiani, Riccardo, et al.
Published: (2024)
by: Poiani, Riccardo, et al.
Published: (2024)
Constrained Best Arm Identification in Grouped Bandits
by: Dharod, Sahil, et al.
Published: (2024)
by: Dharod, Sahil, et al.
Published: (2024)
On Pareto Optimality for Parametric Choice Bandits
by: Zuo, Jierui, et al.
Published: (2025)
by: Zuo, Jierui, et al.
Published: (2025)
Online Nonsubmodular Optimization with Delayed Feedback in the Bandit Setting
by: Yang, Sifan, et al.
Published: (2025)
by: Yang, Sifan, et al.
Published: (2025)
Efficient Online Set-valued Classification with Bandit Feedback
by: Wang, Zhou, et al.
Published: (2024)
by: Wang, Zhou, et al.
Published: (2024)
Multi-thresholding Good Arm Identification with Bandit Feedback
by: Jiang, Xuanke, et al.
Published: (2025)
by: Jiang, Xuanke, et al.
Published: (2025)
Constrained Feedback Learning for Non-Stationary Multi-Armed Bandits
by: Li, Shaoang, et al.
Published: (2025)
by: Li, Shaoang, et al.
Published: (2025)
Optimal Multi-Fidelity Best-Arm Identification
by: Poiani, Riccardo, et al.
Published: (2024)
by: Poiani, Riccardo, et al.
Published: (2024)
Fixed-Budget Constrained Best Arm Identification in Grouped Bandits
by: Mukherjee, Raunak, et al.
Published: (2026)
by: Mukherjee, Raunak, et al.
Published: (2026)
Sequential Learning of the Pareto Front for Multi-objective Bandits
by: Crépon, Elise, et al.
Published: (2025)
by: Crépon, Elise, et al.
Published: (2025)
Fixed-Confidence Multiple Change Point Identification under Bandit Feedback
by: Lazzaro, Joseph, et al.
Published: (2025)
by: Lazzaro, Joseph, et al.
Published: (2025)
Finding good policies in average-reward Markov Decision Processes without prior knowledge
by: Tuynman, Adrienne, et al.
Published: (2024)
by: Tuynman, Adrienne, et al.
Published: (2024)
The Sample Complexity of Multiple Change Point Identification under Bandit Feedback
by: Graf, Maximilian, et al.
Published: (2026)
by: Graf, Maximilian, et al.
Published: (2026)
Replicable Constrained Bandits
by: Bollini, Matteo, et al.
Published: (2026)
by: Bollini, Matteo, et al.
Published: (2026)
Amortized Active Generation of Pareto Sets
by: Steinberg, Daniel M., et al.
Published: (2025)
by: Steinberg, Daniel M., et al.
Published: (2025)
Identification of Energy Management Configuration Concepts from a Set of Pareto-optimal Solutions
by: Lanfermann, Felix, et al.
Published: (2023)
by: Lanfermann, Felix, et al.
Published: (2023)
Causal Bandits: The Pareto Optimal Frontier of Adaptivity, a Reduction to Linear Bandits, and Limitations around Unknown Marginals
by: Liu, Ziyi, et al.
Published: (2024)
by: Liu, Ziyi, et al.
Published: (2024)
Nearest Neighbour with Bandit Feedback
by: Pasteris, Stephen, et al.
Published: (2023)
by: Pasteris, Stephen, et al.
Published: (2023)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
by: Goyal, Tanmay, et al.
Published: (2025)
by: Goyal, Tanmay, et al.
Published: (2025)
PAC Learning with Bandit Feedback: Sharp Sample Complexity in the Realizable Setting
by: Hanneke, Steve, et al.
Published: (2026)
by: Hanneke, Steve, et al.
Published: (2026)
Pareto Front Shape-Agnostic Pareto Set Learning in Multi-Objective Optimization
by: Ye, Rongguang, et al.
Published: (2024)
by: Ye, Rongguang, et al.
Published: (2024)
Lipschitz Bandits with Stochastic Delayed Feedback
by: Liu, Zhongxuan, et al.
Published: (2025)
by: Liu, Zhongxuan, et al.
Published: (2025)
Queueing Matching Bandits with Preference Feedback
by: Kim, Jung-hun, et al.
Published: (2024)
by: Kim, Jung-hun, et al.
Published: (2024)
Graph Feedback Bandits with Similar Arms
by: Qi, Han, et al.
Published: (2024)
by: Qi, Han, et al.
Published: (2024)
Nonparametric Kernel Clustering with Bandit Feedback
by: Thuot, Victor, et al.
Published: (2026)
by: Thuot, Victor, et al.
Published: (2026)
Cascading Bandits With Feedback
by: Prakash, R Sri, et al.
Published: (2025)
by: Prakash, R Sri, et al.
Published: (2025)
Optimal Clustering with Bandit Feedback
by: Yang, Junwen, et al.
Published: (2022)
by: Yang, Junwen, et al.
Published: (2022)
DP-SPRT: Differentially Private Sequential Probability Ratio Tests
by: Michel, Thomas, et al.
Published: (2025)
by: Michel, Thomas, et al.
Published: (2025)
Sequential Membership Inference Attacks
by: Michel, Thomas, et al.
Published: (2026)
by: Michel, Thomas, et al.
Published: (2026)
Pareto Set Learning for Multi-Objective Reinforcement Learning
by: Liu, Erlong, et al.
Published: (2025)
by: Liu, Erlong, et al.
Published: (2025)
Stochastic $k$-Submodular Bandits with Full Bandit Feedback
by: Nie, Guanyu, et al.
Published: (2024)
by: Nie, Guanyu, et al.
Published: (2024)
Constrained Contextual Bandits with Adversarial Contexts
by: Sarkar, Dhruv, et al.
Published: (2026)
by: Sarkar, Dhruv, et al.
Published: (2026)
Does Feedback Help in Bandits with Arm Erasures?
by: Karakas, Merve, et al.
Published: (2025)
by: Karakas, Merve, et al.
Published: (2025)
Bandit and Delayed Feedback in Online Structured Prediction
by: Shibukawa, Yuki, et al.
Published: (2025)
by: Shibukawa, Yuki, et al.
Published: (2025)
Learning Equilibria in Matching Games with Bandit Feedback
by: Athanasopoulos, Andreas, et al.
Published: (2025)
by: Athanasopoulos, Andreas, et al.
Published: (2025)
Similar Items
-
Bandit Pareto Set Identification: the Fixed Budget Setting
by: Kone, Cyrille, et al.
Published: (2023) -
Bandit Pareto Set Identification in a Multi-Output Linear Model
by: Kone, Cyrille, et al.
Published: (2025) -
Pareto Set Identification With Posterior Sampling
by: Kone, Cyrille, et al.
Published: (2024) -
Robust Pareto Set Identification with Contaminated Bandit Feedback
by: Korkmaz, İlter Onat, et al.
Published: (2022) -
Optimal Posterior Sampling for Policy Identification in Tabular Markov Decision Processes
by: Kone, Cyrille, et al.
Published: (2026)