A Finite Time Analysis of Thompson Sampling for Bayesian Optimization with Preferential Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Lazzaro, Joseph, Buffelli, Davide, Shiu, Da-shan, Vakili, Sattar |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds
by: Kayal, Aya, et al.
Published: (2025)
by: Kayal, Aya, et al.
Published: (2025)
Learning Kernel-Based MDPs from Episodic Preferential Feedback
by: Pavlovic, Nikola, et al.
Published: (2026)
by: Pavlovic, Nikola, et al.
Published: (2026)
Towards a Foundation Model for Communication Systems
by: Buffelli, Davide, et al.
Published: (2025)
by: Buffelli, Davide, et al.
Published: (2025)
No-Regret Thompson Sampling for Finite-Horizon Markov Decision Processes with Gaussian Processes
by: Bayrooti, Jasmine, et al.
Published: (2025)
by: Bayrooti, Jasmine, et al.
Published: (2025)
Exact, Tractable Gauss-Newton Optimization in Deep Reversible Architectures Reveal Poor Generalization
by: Buffelli, Davide, et al.
Published: (2024)
by: Buffelli, Davide, et al.
Published: (2024)
Random Exploration in Bayesian Optimization: Order-Optimal Regret and Computational Efficiency
by: Salgia, Sudeep, et al.
Published: (2023)
by: Salgia, Sudeep, et al.
Published: (2023)
Open Problem: Order Optimal Regret Bounds for Kernel-Based Reinforcement Learning
by: Vakili, Sattar
Published: (2024)
by: Vakili, Sattar
Published: (2024)
Group Think: Multiple Concurrent Reasoning Agents Collaborating at Token Level Granularity
by: Hsu, Chan-Jan, et al.
Published: (2025)
by: Hsu, Chan-Jan, et al.
Published: (2025)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
by: Vakili, Sattar, et al.
Published: (2024)
by: Vakili, Sattar, et al.
Published: (2024)
Kernelized Reinforcement Learning with Order Optimal Regret Bounds
by: Vakili, Sattar, et al.
Published: (2023)
by: Vakili, Sattar, et al.
Published: (2023)
Near-Optimal Sample Complexity in Reward-Free Kernel-Based Reinforcement Learning
by: Kayal, Aya, et al.
Published: (2025)
by: Kayal, Aya, et al.
Published: (2025)
On Regret Bounds of Thompson Sampling for Bayesian Optimization
by: Takeno, Shion, et al.
Published: (2026)
by: Takeno, Shion, et al.
Published: (2026)
Fast, Precise Thompson Sampling for Bayesian Optimization
by: Sweet, David
Published: (2024)
by: Sweet, David
Published: (2024)
Local Preferential Bayesian Optimization
by: Menn, Johanna, et al.
Published: (2026)
by: Menn, Johanna, et al.
Published: (2026)
Principled Preferential Bayesian Optimization
by: Xu, Wenjie, et al.
Published: (2024)
by: Xu, Wenjie, et al.
Published: (2024)
Consecutive Preferential Bayesian Optimization
by: Erarslan, Aras, et al.
Published: (2025)
by: Erarslan, Aras, et al.
Published: (2025)
On Thompson Sampling and Bilateral Uncertainty in Additive Bayesian Optimization
by: Wycoff, Nathan
Published: (2025)
by: Wycoff, Nathan
Published: (2025)
Epsilon-Greedy Thompson Sampling to Bayesian Optimization
by: Do, Bach, et al.
Published: (2024)
by: Do, Bach, et al.
Published: (2024)
The Deep Equilibrium Algorithmic Reasoner
by: Georgiev, Dobrik, et al.
Published: (2024)
by: Georgiev, Dobrik, et al.
Published: (2024)
Fixed-Confidence Multiple Change Point Identification under Bandit Feedback
by: Lazzaro, Joseph, et al.
Published: (2025)
by: Lazzaro, Joseph, et al.
Published: (2025)
CliquePH: Higher-Order Information for Graph Neural Networks through Persistent Homology on Clique Graphs
by: Buffelli, Davide, et al.
Published: (2024)
by: Buffelli, Davide, et al.
Published: (2024)
Preferential Multi-Objective Bayesian Optimization
by: Astudillo, Raul, et al.
Published: (2024)
by: Astudillo, Raul, et al.
Published: (2024)
Anchor-Based Heteroscedastic Noise for Preferential Bayesian Optimization
by: Sinaga, Marshal Arijona, et al.
Published: (2024)
by: Sinaga, Marshal Arijona, et al.
Published: (2024)
Deep Equilibrium Algorithmic Reasoning
by: Georgiev, Dobrik, et al.
Published: (2024)
by: Georgiev, Dobrik, et al.
Published: (2024)
Enhanced Bayesian Optimization via Preferential Modeling of Abstract Properties
by: A V, Arun Kumar, et al.
Published: (2024)
by: A V, Arun Kumar, et al.
Published: (2024)
Rethinking the shape convention of an MLP
by: Chen, Meng-Hsi, et al.
Published: (2025)
by: Chen, Meng-Hsi, et al.
Published: (2025)
BFTS: Thompson Sampling with Bayesian Additive Regression Trees
by: Deng, Ruizhe, et al.
Published: (2026)
by: Deng, Ruizhe, et al.
Published: (2026)
Reinforcement Learning Using known Invariances
by: Cioba, Alexandru, et al.
Published: (2025)
by: Cioba, Alexandru, et al.
Published: (2025)
Revisiting the Shape Convention of Transformer Language Models
by: Liao, Feng-Ting, et al.
Published: (2026)
by: Liao, Feng-Ting, et al.
Published: (2026)
Preferential Multi-Objective Bayesian Optimization for Drug Discovery
by: Dang, Tai, et al.
Published: (2025)
by: Dang, Tai, et al.
Published: (2025)
Adaptive Candidate Point Thompson Sampling for High-Dimensional Bayesian Optimization
by: Fan, Donney, et al.
Published: (2026)
by: Fan, Donney, et al.
Published: (2026)
LGDC: Latent Graph Diffusion via Spectrum-Preserving Coarsening
by: Osman, Nagham, et al.
Published: (2025)
by: Osman, Nagham, et al.
Published: (2025)
BOTS: Batch Bayesian Optimization of Extended Thompson Sampling for Severely Episode-Limited RL Settings
by: Karine, Karine, et al.
Published: (2024)
by: Karine, Karine, et al.
Published: (2024)
Improved Bayesian Regret Bounds for Thompson Sampling in Reinforcement Learning
by: Moradipari, Ahmadreza, et al.
Published: (2023)
by: Moradipari, Ahmadreza, et al.
Published: (2023)
Bayesian Collaborative Bandits with Thompson Sampling for Improved Outreach in Maternal Health Program
by: Dasgupta, Arpan, et al.
Published: (2024)
by: Dasgupta, Arpan, et al.
Published: (2024)
An Information-Theoretic Analysis of Thompson Sampling for Logistic Bandits
by: Gouverneur, Amaury, et al.
Published: (2024)
by: Gouverneur, Amaury, et al.
Published: (2024)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
A Broader View of Thompson Sampling
by: Qu, Yanlin, et al.
Published: (2025)
by: Qu, Yanlin, et al.
Published: (2025)
Prior-Aligned Meta-RL: Thompson Sampling with Learned Priors and Guarantees in Finite-Horizon MDPs
by: Zhou, Runlin, et al.
Published: (2025)
by: Zhou, Runlin, et al.
Published: (2025)
Thompson Sampling for Repeated Newsvendor
by: Chen, Li, et al.
Published: (2025)
by: Chen, Li, et al.
Published: (2025)
Similar Items
-
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds
by: Kayal, Aya, et al.
Published: (2025) -
Learning Kernel-Based MDPs from Episodic Preferential Feedback
by: Pavlovic, Nikola, et al.
Published: (2026) -
Towards a Foundation Model for Communication Systems
by: Buffelli, Davide, et al.
Published: (2025) -
No-Regret Thompson Sampling for Finite-Horizon Markov Decision Processes with Gaussian Processes
by: Bayrooti, Jasmine, et al.
Published: (2025) -
Exact, Tractable Gauss-Newton Optimization in Deep Reversible Architectures Reveal Poor Generalization
by: Buffelli, Davide, et al.
Published: (2024)