Regenerative Particle Thompson Sampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Zeyu, Hajek, Bruce, Choi, Nakjung, Walid, Anwar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Is Thompson Sampling Susceptible to Algorithmic Collusion?
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
Learning-augmented Online Algorithm for Two-level Ski-rental Problem
von: Zhang, Keyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Keyuan, et al.
Veröffentlicht: (2024)
Thompson Sampling for Repeated Newsvendor
von: Chen, Li, et al.
Veröffentlicht: (2025)
von: Chen, Li, et al.
Veröffentlicht: (2025)
Constrained Linear Thompson Sampling
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025)
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
von: Namkoong, Hongseok, et al.
Veröffentlicht: (2020)
von: Namkoong, Hongseok, et al.
Veröffentlicht: (2020)
Adaptive Data Augmentation for Thompson Sampling
von: Kim, Wonyoung
Veröffentlicht: (2025)
von: Kim, Wonyoung
Veröffentlicht: (2025)
A Broader View of Thompson Sampling
von: Qu, Yanlin, et al.
Veröffentlicht: (2025)
von: Qu, Yanlin, et al.
Veröffentlicht: (2025)
Graph Neural Thompson Sampling
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
Fast, Precise Thompson Sampling for Bayesian Optimization
von: Sweet, David
Veröffentlicht: (2024)
von: Sweet, David
Veröffentlicht: (2024)
Thompson Sampling in Partially Observable Contextual Bandits
von: Park, Hongju, et al.
Veröffentlicht: (2024)
von: Park, Hongju, et al.
Veröffentlicht: (2024)
Online Learning of Decision Trees with Thompson Sampling
von: Chaouki, Ayman, et al.
Veröffentlicht: (2024)
von: Chaouki, Ayman, et al.
Veröffentlicht: (2024)
On Regret Bounds of Thompson Sampling for Bayesian Optimization
von: Takeno, Shion, et al.
Veröffentlicht: (2026)
von: Takeno, Shion, et al.
Veröffentlicht: (2026)
Thompson Sampling For Combinatorial Bandits: Polynomial Regret and Mismatched Sampling Paradox
von: Zhang, Raymond, et al.
Veröffentlicht: (2024)
von: Zhang, Raymond, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Automated Stock Trading: An Ensemble Strategy
von: Yang, Hongyang, et al.
Veröffentlicht: (2025)
von: Yang, Hongyang, et al.
Veröffentlicht: (2025)
Thompson Sampling for Multi-Objective Linear Contextual Bandit
von: Park, Somangchan, et al.
Veröffentlicht: (2025)
von: Park, Somangchan, et al.
Veröffentlicht: (2025)
An Information-Theoretic Analysis of Thompson Sampling for Logistic Bandits
von: Gouverneur, Amaury, et al.
Veröffentlicht: (2024)
von: Gouverneur, Amaury, et al.
Veröffentlicht: (2024)
On Thompson Sampling and Bilateral Uncertainty in Additive Bayesian Optimization
von: Wycoff, Nathan
Veröffentlicht: (2025)
von: Wycoff, Nathan
Veröffentlicht: (2025)
Thompson Sampling in Online RLHF with General Function Approximation
von: Feng, Songtao, et al.
Veröffentlicht: (2025)
von: Feng, Songtao, et al.
Veröffentlicht: (2025)
Thompson Sampling-like Algorithms for Stochastic Rising Bandits
von: Fiandri, Marco, et al.
Veröffentlicht: (2025)
von: Fiandri, Marco, et al.
Veröffentlicht: (2025)
Sliding-Window Thompson Sampling for Non-Stationary Settings
von: Fiandri, Marco, et al.
Veröffentlicht: (2024)
von: Fiandri, Marco, et al.
Veröffentlicht: (2024)
Stochastically Constrained Best Arm Identification with Thompson Sampling
von: Yang, Le, et al.
Veröffentlicht: (2025)
von: Yang, Le, et al.
Veröffentlicht: (2025)
VITS : Variational Inference Thompson Sampling for contextual bandits
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
Thompson Sampling in Function Spaces via Neural Operators
von: Oliveira, Rafael, et al.
Veröffentlicht: (2025)
von: Oliveira, Rafael, et al.
Veröffentlicht: (2025)
MINTS: Minimalist Thompson Sampling
von: Wang, Kaizheng
Veröffentlicht: (2026)
von: Wang, Kaizheng
Veröffentlicht: (2026)
Prior-Aligned Meta-RL: Thompson Sampling with Learned Priors and Guarantees in Finite-Horizon MDPs
von: Zhou, Runlin, et al.
Veröffentlicht: (2025)
von: Zhou, Runlin, et al.
Veröffentlicht: (2025)
Sample-Mean Anchored Thompson Sampling for Offline-to-Online Learning with Distribution Shift
von: Li, Bochao, et al.
Veröffentlicht: (2026)
von: Li, Bochao, et al.
Veröffentlicht: (2026)
FinGPT-HPC: Efficient Pretraining and Finetuning Large Language Models for Financial Applications with High-Performance Computing
von: Liu, Xiao-Yang, et al.
Veröffentlicht: (2024)
von: Liu, Xiao-Yang, et al.
Veröffentlicht: (2024)
Thompson Sampling Itself is Differentially Private
von: Ou, Tingting, et al.
Veröffentlicht: (2024)
von: Ou, Tingting, et al.
Veröffentlicht: (2024)
Counterfactual Inference under Thompson Sampling
von: Jeunen, Olivier
Veröffentlicht: (2025)
von: Jeunen, Olivier
Veröffentlicht: (2025)
Gated Graph Attention Networks for Predicting Duration of Large Scale Power Outages Induced by Natural Disasters
von: Duan, Chenghao, et al.
Veröffentlicht: (2026)
von: Duan, Chenghao, et al.
Veröffentlicht: (2026)
Diffusion Models for Solving Inverse Problems via Posterior Sampling with Piecewise Guidance
von: Mohseni-Sehdeh, Saeed, et al.
Veröffentlicht: (2025)
von: Mohseni-Sehdeh, Saeed, et al.
Veröffentlicht: (2025)
Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models
von: Fang, Zeyu, et al.
Veröffentlicht: (2024)
von: Fang, Zeyu, et al.
Veröffentlicht: (2024)
Thompson Sampling via Fine-Tuning of LLMs
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
Gaussian Process Thompson Sampling via Rootfinding
von: Adebiyi, Taiwo A., et al.
Veröffentlicht: (2024)
von: Adebiyi, Taiwo A., et al.
Veröffentlicht: (2024)
Epsilon-Greedy Thompson Sampling to Bayesian Optimization
von: Do, Bach, et al.
Veröffentlicht: (2024)
von: Do, Bach, et al.
Veröffentlicht: (2024)
Reinterpreting 'the Company a Word Keeps': Towards Explainable and Ontologically Grounded Language Models
von: Saba, Walid S.
Veröffentlicht: (2024)
von: Saba, Walid S.
Veröffentlicht: (2024)
Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks
von: Xu, Yinglun, et al.
Veröffentlicht: (2024)
von: Xu, Yinglun, et al.
Veröffentlicht: (2024)
Thompson Sampling-Based Learning and Control for Unknown Dynamic Systems
von: Zheng, Kaikai, et al.
Veröffentlicht: (2025)
von: Zheng, Kaikai, et al.
Veröffentlicht: (2025)
Adaptive Prior Selection in Gaussian Process Bandits with Thompson Sampling
von: Sandberg, Jack, et al.
Veröffentlicht: (2025)
von: Sandberg, Jack, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Is Thompson Sampling Susceptible to Algorithmic Collusion?
von: Xiong, Yi, et al.
Veröffentlicht: (2024) -
Learning-augmented Online Algorithm for Two-level Ski-rental Problem
von: Zhang, Keyuan, et al.
Veröffentlicht: (2024) -
Thompson Sampling for Repeated Newsvendor
von: Chen, Li, et al.
Veröffentlicht: (2025) -
Constrained Linear Thompson Sampling
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025) -
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
von: Li, Yingru, et al.
Veröffentlicht: (2024)