Variance-Aware Linear UCB with Deep Representation for Neural Contextual Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bui, Ha Manh, Mallada, Enrique, Liu, Anqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Density-Regression: Efficient and Distance-Aware Deep Regressor for Uncertainty Estimation under Distribution Shifts
von: Bui, Ha Manh, et al.
Veröffentlicht: (2024)
von: Bui, Ha Manh, et al.
Veröffentlicht: (2024)
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts
von: Bui, Ha Manh, et al.
Veröffentlicht: (2023)
von: Bui, Ha Manh, et al.
Veröffentlicht: (2023)
Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2026)
von: Bui, Ha Manh, et al.
Veröffentlicht: (2026)
Q-Learning with Shift-Aware Upper Confidence Bound in Non-Stationary Reinforcement Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2025)
von: Bui, Ha Manh, et al.
Veröffentlicht: (2025)
Calibrated Uncertainty Sampling for Active Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2025)
von: Bui, Ha Manh, et al.
Veröffentlicht: (2025)
Truncated LinUCB for Stochastic Linear Bandits
von: Song, Yanglei, et al.
Veröffentlicht: (2022)
von: Song, Yanglei, et al.
Veröffentlicht: (2022)
PAK-UCB Contextual Bandit: An Online Learning Approach to Prompt-Aware Selection of Generative Models and LLMs
von: Hu, Xiaoyan, et al.
Veröffentlicht: (2024)
von: Hu, Xiaoyan, et al.
Veröffentlicht: (2024)
Precise Asymptotics and Refined Regret of Variance-Aware UCB
von: Fan, Yingying, et al.
Veröffentlicht: (2024)
von: Fan, Yingying, et al.
Veröffentlicht: (2024)
Neural Variance-aware Dueling Bandits with Deep Representation and Shallow Exploration
von: Oh, Youngmin, et al.
Veröffentlicht: (2025)
von: Oh, Youngmin, et al.
Veröffentlicht: (2025)
Variance-Aware Feel-Good Thompson Sampling for Contextual Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
Variance-Aware Regret Bounds for Stochastic Contextual Dueling Bandits
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
Replicable Bandits with UCB based Exploration
von: Deb, Rohan, et al.
Veröffentlicht: (2026)
von: Deb, Rohan, et al.
Veröffentlicht: (2026)
Conservative Contextual Bandits: Beyond Linear Representations
von: Deb, Rohan, et al.
Veröffentlicht: (2024)
von: Deb, Rohan, et al.
Veröffentlicht: (2024)
Group-Sensitive Offline Contextual Bandits
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
How Does Variance Shape the Regret in Contextual Bandits?
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
Direction-Aware Offline-to-Online Learning in Linear Contextual Bandits
von: Han, Zean, et al.
Veröffentlicht: (2026)
von: Han, Zean, et al.
Veröffentlicht: (2026)
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
Extended UCB Policies for Multi-armed Bandit Problems
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
von: Liu, Keqin, et al.
Veröffentlicht: (2011)
Linear Contextual Bandits with Interference
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
von: He, Jiafan, et al.
Veröffentlicht: (2025)
von: He, Jiafan, et al.
Veröffentlicht: (2025)
Contextual Linear Bandits with Delay as Payoff
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
Clus-UCB: A Near-Optimal Algorithm for Clustered Bandits
von: Gore, Aakash, et al.
Veröffentlicht: (2025)
von: Gore, Aakash, et al.
Veröffentlicht: (2025)
Risk-Aware Continuous Control with Neural Contextual Bandits
von: Ayala-Romero, Jose A., et al.
Veröffentlicht: (2023)
von: Ayala-Romero, Jose A., et al.
Veröffentlicht: (2023)
Scaling Federated Linear Contextual Bandits via Sketching
von: Yang, Hantao, et al.
Veröffentlicht: (2026)
von: Yang, Hantao, et al.
Veröffentlicht: (2026)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
von: Kang, Yue, et al.
Veröffentlicht: (2025)
von: Kang, Yue, et al.
Veröffentlicht: (2025)
Revisiting Social Welfare in Bandits: UCB is (Nearly) All You Need
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2025)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2025)
Partially Observable Contextual Bandits with Linear Payoffs
von: Zeng, Sihan, et al.
Veröffentlicht: (2024)
von: Zeng, Sihan, et al.
Veröffentlicht: (2024)
Active Learning for Stochastic Contextual Linear Bandits
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
Strategic Linear Contextual Bandits
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
A UCB Bandit Algorithm for General ML-Based Estimators
von: Liu, Yajing, et al.
Veröffentlicht: (2026)
von: Liu, Yajing, et al.
Veröffentlicht: (2026)
Uncertainty of Joint Neural Contextual Bandit
von: Guo, Hongbo, et al.
Veröffentlicht: (2024)
von: Guo, Hongbo, et al.
Veröffentlicht: (2024)
Neural Exploitation and Exploration of Contextual Bandits
von: Ban, Yikun, et al.
Veröffentlicht: (2023)
von: Ban, Yikun, et al.
Veröffentlicht: (2023)
From Theory to Practice with RAVEN-UCB: Addressing Non-Stationarity in Multi-Armed Bandits through Variance Adaptation
von: Fang, Junyi, et al.
Veröffentlicht: (2025)
von: Fang, Junyi, et al.
Veröffentlicht: (2025)
Thompson Sampling for Multi-Objective Linear Contextual Bandit
von: Park, Somangchan, et al.
Veröffentlicht: (2025)
von: Park, Somangchan, et al.
Veröffentlicht: (2025)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
A Reduction Algorithm for Markovian Contextual Linear Bandits
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026)
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026)
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
von: Paschalidis, Phevos, et al.
Veröffentlicht: (2024)
von: Paschalidis, Phevos, et al.
Veröffentlicht: (2024)
Federated Linear Contextual Bandits with Heterogeneous Clients
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
Linear Contextual Bandits with Hybrid Payoff: Revisited
von: Das, Nirjhar, et al.
Veröffentlicht: (2024)
von: Das, Nirjhar, et al.
Veröffentlicht: (2024)
Neural Risk-sensitive Satisficing in Contextual Bandits
von: Ito, Shogo, et al.
Veröffentlicht: (2025)
von: Ito, Shogo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Density-Regression: Efficient and Distance-Aware Deep Regressor for Uncertainty Estimation under Distribution Shifts
von: Bui, Ha Manh, et al.
Veröffentlicht: (2024) -
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts
von: Bui, Ha Manh, et al.
Veröffentlicht: (2023) -
Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2026) -
Q-Learning with Shift-Aware Upper Confidence Bound in Non-Stationary Reinforcement Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2025) -
Calibrated Uncertainty Sampling for Active Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2025)