An Improved Algorithm for Adversarial Linear Contextual Bandits via Reduction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | van Erven, Tim, Mayo, Jack, Olkhovskaya, Julia, Wei, Chen-Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sparse Nonparametric Contextual Bandits
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
Improved Regret Bounds for Bandits with Expert Advice
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024)
A Reduction Algorithm for Markovian Contextual Linear Bandits
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026)
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
Nearly Minimax Discrete Distribution Estimation in Kullback-Leibler Divergence with High Probability
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2025)
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2025)
Accelerated Rates between Stochastic and Adversarial Online Convex Optimization
von: Sachs, Sarah, et al.
Veröffentlicht: (2023)
von: Sachs, Sarah, et al.
Veröffentlicht: (2023)
Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates
von: Yu, Sanghoon, et al.
Veröffentlicht: (2026)
von: Yu, Sanghoon, et al.
Veröffentlicht: (2026)
Kernelized Reinforcement Learning with Order Optimal Regret Bounds
von: Vakili, Sattar, et al.
Veröffentlicht: (2023)
von: Vakili, Sattar, et al.
Veröffentlicht: (2023)
A Simple Reduction Scheme for Constrained Contextual Bandits with Adversarial Contexts via Regression
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
von: Kuroki, Yuko, et al.
Veröffentlicht: (2023)
Improved Algorithm for Adversarial Linear Mixture MDPs with Bandit Feedback and Unknown Transition
von: Li, Long-Fei, et al.
Veröffentlicht: (2024)
von: Li, Long-Fei, et al.
Veröffentlicht: (2024)
Sample-efficient Learning of Concepts with Theoretical Guarantees: from Data to Concepts without Interventions
von: Fokkema, Hidde, et al.
Veröffentlicht: (2025)
von: Fokkema, Hidde, et al.
Veröffentlicht: (2025)
Nearly Optimal Algorithms for Contextual Dueling Bandits from Adversarial Feedback
von: Di, Qiwei, et al.
Veröffentlicht: (2024)
von: Di, Qiwei, et al.
Veröffentlicht: (2024)
Constrained Contextual Bandits with Adversarial Contexts
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
Scaling Federated Linear Contextual Bandits via Sketching
von: Yang, Hantao, et al.
Veröffentlicht: (2026)
von: Yang, Hantao, et al.
Veröffentlicht: (2026)
The Risks of Recourse in Binary Classification
von: Fokkema, Hidde, et al.
Veröffentlicht: (2023)
von: Fokkema, Hidde, et al.
Veröffentlicht: (2023)
Linear Contextual Bandits with Interference
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Improved Algorithms for Nash Welfare in Linear Bandits
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
Contextual Linear Bandits with Delay as Payoff
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
Robust and Computationally Efficient Linear Contextual Bandits under Adversarial Corruption and Heavy-Tailed Noise
von: Tani, Naoto, et al.
Veröffentlicht: (2026)
von: Tani, Naoto, et al.
Veröffentlicht: (2026)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
An Online Feasible Point Method for Benign Generalized Nash Equilibrium Problems
von: Sachs, Sarah, et al.
Veröffentlicht: (2024)
von: Sachs, Sarah, et al.
Veröffentlicht: (2024)
Active Learning for Stochastic Contextual Linear Bandits
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
Partially Observable Contextual Bandits with Linear Payoffs
von: Zeng, Sihan, et al.
Veröffentlicht: (2024)
von: Zeng, Sihan, et al.
Veröffentlicht: (2024)
Online Newton Method for Bandit Convex Optimisation
von: Fokkema, Hidde, et al.
Veröffentlicht: (2024)
von: Fokkema, Hidde, et al.
Veröffentlicht: (2024)
Strategic Linear Contextual Bandits
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
Navigating Sparsities in High-Dimensional Linear Contextual Bandits
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
von: Kang, Yue, et al.
Veröffentlicht: (2025)
von: Kang, Yue, et al.
Veröffentlicht: (2025)
Optimal and Practical Batched Linear Bandit Algorithm
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
How Does Variance Shape the Regret in Contextual Bandits?
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
Thompson Sampling for Multi-Objective Linear Contextual Bandit
von: Park, Somangchan, et al.
Veröffentlicht: (2025)
von: Park, Somangchan, et al.
Veröffentlicht: (2025)
Federated Linear Contextual Bandits with Heterogeneous Clients
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
Conservative Contextual Bandits: Beyond Linear Representations
von: Deb, Rohan, et al.
Veröffentlicht: (2024)
von: Deb, Rohan, et al.
Veröffentlicht: (2024)
Linear Contextual Bandits with Hybrid Payoff: Revisited
von: Das, Nirjhar, et al.
Veröffentlicht: (2024)
von: Das, Nirjhar, et al.
Veröffentlicht: (2024)
A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026)
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026)
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
Slowly Changing Adversarial Bandit Algorithms are Efficient for Discounted MDPs
von: Kash, Ian A., et al.
Veröffentlicht: (2022)
von: Kash, Ian A., et al.
Veröffentlicht: (2022)
Learning with Incomplete Context: Linear Contextual Bandits with Pretrained Imputation
von: Yan, Hao, et al.
Veröffentlicht: (2025)
von: Yan, Hao, et al.
Veröffentlicht: (2025)
Direction-Aware Offline-to-Online Learning in Linear Contextual Bandits
von: Han, Zean, et al.
Veröffentlicht: (2026)
von: Han, Zean, et al.
Veröffentlicht: (2026)
Online Continuous Hyperparameter Optimization for Generalized Linear Contextual Bandits
von: Kang, Yue, et al.
Veröffentlicht: (2023)
von: Kang, Yue, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Sparse Nonparametric Contextual Bandits
von: Flynn, Hamish, et al.
Veröffentlicht: (2025) -
Improved Regret Bounds for Bandits with Expert Advice
von: Cesa-Bianchi, Nicolò, et al.
Veröffentlicht: (2024) -
A Reduction Algorithm for Markovian Contextual Linear Bandits
von: Buyukkalayci, Kaan, et al.
Veröffentlicht: (2026) -
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
von: Vakili, Sattar, et al.
Veröffentlicht: (2024) -
Nearly Minimax Discrete Distribution Estimation in Kullback-Leibler Divergence with High Probability
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2025)