Partially Observable Contextual Bandits with Linear Payoffs
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Sihan, Bhatt, Sujay, Koppel, Alec, Ganesh, Sumitra |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
by: Zeng, Sihan, et al.
Published: (2026)
by: Zeng, Sihan, et al.
Published: (2026)
Regularized Proportional Fairness Mechanism for Resource Allocation Without Money
by: Zeng, Sihan, et al.
Published: (2025)
by: Zeng, Sihan, et al.
Published: (2025)
Rethinking Neural Network Learning Rates: A Stackelberg Perspective
by: Zeng, Sihan, et al.
Published: (2026)
by: Zeng, Sihan, et al.
Published: (2026)
Learning Payment-Free Resource Allocation Mechanisms
by: Zeng, Sihan, et al.
Published: (2023)
by: Zeng, Sihan, et al.
Published: (2023)
Learning in Stackelberg Mean Field Games: A Non-Asymptotic Analysis
by: Zeng, Sihan, et al.
Published: (2025)
by: Zeng, Sihan, et al.
Published: (2025)
Learning in Herding Mean Field Games: Single-Loop Algorithm with Finite-Time Convergence Analysis
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Approximate Equivariance in Reinforcement Learning
by: Park, Jung Yeon, et al.
Published: (2024)
by: Park, Jung Yeon, et al.
Published: (2024)
No One Size Fits All: QueryBandits for Hallucination Mitigation
by: Cho, Nicole, et al.
Published: (2026)
by: Cho, Nicole, et al.
Published: (2026)
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
by: Cho, Nicole, et al.
Published: (2025)
by: Cho, Nicole, et al.
Published: (2025)
Contextual Linear Bandits with Delay as Payoff
by: Zhang, Mengxiao, et al.
Published: (2025)
by: Zhang, Mengxiao, et al.
Published: (2025)
Linear Contextual Bandits with Hybrid Payoff: Revisited
by: Das, Nirjhar, et al.
Published: (2024)
by: Das, Nirjhar, et al.
Published: (2024)
Thompson Sampling in Partially Observable Contextual Bandits
by: Park, Hongju, et al.
Published: (2024)
by: Park, Hongju, et al.
Published: (2024)
Linear Bandits with Partially Observable Features
by: Kim, Wonyoung, et al.
Published: (2025)
by: Kim, Wonyoung, et al.
Published: (2025)
Efficient Inverse Multiagent Learning
by: Goktas, Denizalp, et al.
Published: (2025)
by: Goktas, Denizalp, et al.
Published: (2025)
Byzantine-Resilient Decentralized Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2023)
by: Zhu, Jingxuan, et al.
Published: (2023)
ADAGE: A generic two-layer framework for adaptive agent based modelling
by: Evans, Benjamin Patrick, et al.
Published: (2025)
by: Evans, Benjamin Patrick, et al.
Published: (2025)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
Catoni-Style Change Point Detection for Regret Minimization in Non-Stationary Heavy-Tailed Bandits
by: Genalti, Gianmarco, et al.
Published: (2025)
by: Genalti, Gianmarco, et al.
Published: (2025)
Linear Contextual Bandits with Interference
by: Xu, Yang, et al.
Published: (2024)
by: Xu, Yang, et al.
Published: (2024)
Decentralized Convergence to Equilibrium Prices in Trading Networks
by: Lock, Edwin, et al.
Published: (2024)
by: Lock, Edwin, et al.
Published: (2024)
Active Learning for Stochastic Contextual Linear Bandits
by: Brunskill, Emma, et al.
Published: (2026)
by: Brunskill, Emma, et al.
Published: (2026)
Strategic Linear Contextual Bandits
by: Buening, Thomas Kleine, et al.
Published: (2024)
by: Buening, Thomas Kleine, et al.
Published: (2024)
Scalable Representation Learning for Multimodal Tabular Transactions
by: Raman, Natraj, et al.
Published: (2024)
by: Raman, Natraj, et al.
Published: (2024)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
by: Kang, Yue, et al.
Published: (2025)
by: Kang, Yue, et al.
Published: (2025)
Scaling Federated Linear Contextual Bandits via Sketching
by: Yang, Hantao, et al.
Published: (2026)
by: Yang, Hantao, et al.
Published: (2026)
Thompson Sampling for Multi-Objective Linear Contextual Bandit
by: Park, Somangchan, et al.
Published: (2025)
by: Park, Somangchan, et al.
Published: (2025)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
by: Kuroki, Yuko, et al.
Published: (2023)
by: Kuroki, Yuko, et al.
Published: (2023)
A Reduction Algorithm for Markovian Contextual Linear Bandits
by: Buyukkalayci, Kaan, et al.
Published: (2026)
by: Buyukkalayci, Kaan, et al.
Published: (2026)
Federated Linear Contextual Bandits with Heterogeneous Clients
by: Blaser, Ethan, et al.
Published: (2024)
by: Blaser, Ethan, et al.
Published: (2024)
Conservative Contextual Bandits: Beyond Linear Representations
by: Deb, Rohan, et al.
Published: (2024)
by: Deb, Rohan, et al.
Published: (2024)
Blocking Bandits
by: Basu, Soumya, et al.
Published: (2019)
by: Basu, Soumya, et al.
Published: (2019)
Contextual Linear Optimization with Partial Feedback
by: Hu, Yichun, et al.
Published: (2024)
by: Hu, Yichun, et al.
Published: (2024)
Direction-Aware Offline-to-Online Learning in Linear Contextual Bandits
by: Han, Zean, et al.
Published: (2026)
by: Han, Zean, et al.
Published: (2026)
Online Continuous Hyperparameter Optimization for Generalized Linear Contextual Bandits
by: Kang, Yue, et al.
Published: (2023)
by: Kang, Yue, et al.
Published: (2023)
An Improved Algorithm for Adversarial Linear Contextual Bandits via Reduction
by: van Erven, Tim, et al.
Published: (2025)
by: van Erven, Tim, et al.
Published: (2025)
Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities
by: Hwang, Taehyun, et al.
Published: (2026)
by: Hwang, Taehyun, et al.
Published: (2026)
Learning with Incomplete Context: Linear Contextual Bandits with Pretrained Imputation
by: Yan, Hao, et al.
Published: (2025)
by: Yan, Hao, et al.
Published: (2025)
Shuffle and Joint Differential Privacy for Generalized Linear Contextual Bandits
by: Sarmasarkar, Sahasrajit
Published: (2026)
by: Sarmasarkar, Sahasrajit
Published: (2026)
Leveraging Offline Data in Linear Latent Contextual Bandits
by: Kausik, Chinmaya, et al.
Published: (2024)
by: Kausik, Chinmaya, et al.
Published: (2024)
On the Optimal Regret of Locally Private Linear Contextual Bandit
by: Li, Jiachun, et al.
Published: (2024)
by: Li, Jiachun, et al.
Published: (2024)
Similar Items
-
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
by: Zeng, Sihan, et al.
Published: (2026) -
Regularized Proportional Fairness Mechanism for Resource Allocation Without Money
by: Zeng, Sihan, et al.
Published: (2025) -
Rethinking Neural Network Learning Rates: A Stackelberg Perspective
by: Zeng, Sihan, et al.
Published: (2026) -
Learning Payment-Free Resource Allocation Mechanisms
by: Zeng, Sihan, et al.
Published: (2023) -
Learning in Stackelberg Mean Field Games: A Non-Asymptotic Analysis
by: Zeng, Sihan, et al.
Published: (2025)