Contextual Bandits for Unbounded Context Distributions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Puning, Fan, Rongfei, Wang, Shaowei, Shen, Li, Zhang, Qixin, Ke, Zong, Zheng, Tianhang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On Theoretical Limits of Learning with Label Differential Privacy
von: Zhao, Puning, et al.
Veröffentlicht: (2025)
von: Zhao, Puning, et al.
Veröffentlicht: (2025)
Consistent Estimation of Numerical Distributions under Local Differential Privacy by Wavelet Expansion
von: Zhao, Puning, et al.
Veröffentlicht: (2025)
von: Zhao, Puning, et al.
Veröffentlicht: (2025)
Learning with User-Level Local Differential Privacy
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
Sequential Federated Learning in Hierarchical Architecture on Non-IID Datasets
von: Yan, Xingrun, et al.
Veröffentlicht: (2024)
von: Yan, Xingrun, et al.
Veröffentlicht: (2024)
Enhancing Learning with Label Differential Privacy by Vector Approximation
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
H+: An Efficient Similarity-Aware Aggregation for Byzantine Resilient Federated Learning
von: Zuo, Shiyuan, et al.
Veröffentlicht: (2025)
von: Zuo, Shiyuan, et al.
Veröffentlicht: (2025)
Differential Private Stochastic Optimization with Heavy-tailed Data: Towards Optimal Rates
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
Efficient Federated Learning against Byzantine Attacks and Data Heterogeneity via Aggregating Normalized Gradients
von: Zuo, Shiyuan, et al.
Veröffentlicht: (2024)
von: Zuo, Shiyuan, et al.
Veröffentlicht: (2024)
High Dimensional Distributed Gradient Descent with Arbitrary Number of Byzantine Attackers
von: Liu, Wenyu, et al.
Veröffentlicht: (2023)
von: Liu, Wenyu, et al.
Veröffentlicht: (2023)
Constrained Contextual Bandits with Adversarial Contexts
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
von: Shen, Yi, et al.
Veröffentlicht: (2023)
von: Shen, Yi, et al.
Veröffentlicht: (2023)
High Probability Bound for Cross-Learning Contextual Bandits with Unknown Context Distributions
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024)
Causal Contextual Bandits with Adaptive Context
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
Learning with Incomplete Context: Linear Contextual Bandits with Pretrained Imputation
von: Yan, Hao, et al.
Veröffentlicht: (2025)
von: Yan, Hao, et al.
Veröffentlicht: (2025)
Federated Learning Resilient to Byzantine Attacks and Data Heterogeneity
von: Zuo, Shiyuan, et al.
Veröffentlicht: (2024)
von: Zuo, Shiyuan, et al.
Veröffentlicht: (2024)
Minimax Optimal Q Learning with Nearest Neighbors
von: Zhao, Puning, et al.
Veröffentlicht: (2023)
von: Zhao, Puning, et al.
Veröffentlicht: (2023)
Effective Policy Learning for Multi-Agent Online Coordination Beyond Submodular Objectives
von: Zhang, Qixin, et al.
Veröffentlicht: (2025)
von: Zhang, Qixin, et al.
Veröffentlicht: (2025)
Contextual Linear Bandits with Delay as Payoff
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
Active Context Selection Improves Simple Regret in Contextual Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
Navigating Sparsities in High-Dimensional Linear Contextual Bandits
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
von: Zhao, Rui, et al.
Veröffentlicht: (2025)
Nearly Optimal Differentially Private ReLU Regression
von: Ding, Meng, et al.
Veröffentlicht: (2025)
von: Ding, Meng, et al.
Veröffentlicht: (2025)
Differentially Private Kernelized Contextual Bandits
von: Pavlovic, Nikola, et al.
Veröffentlicht: (2025)
von: Pavlovic, Nikola, et al.
Veröffentlicht: (2025)
Sharp Analysis for KL-Regularized Contextual Bandits and RLHF
von: Zhao, Heyang, et al.
Veröffentlicht: (2024)
von: Zhao, Heyang, et al.
Veröffentlicht: (2024)
Context-Action Embedding Learning for Off-Policy Evaluation in Contextual Bandits
von: Chandak, Kushagra, et al.
Veröffentlicht: (2025)
von: Chandak, Kushagra, et al.
Veröffentlicht: (2025)
IBCB: Efficient Inverse Batched Contextual Bandit for Behavioral Evolution History
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Revisiting Follow-the-Perturbed-Leader with Unbounded Perturbations in Bandit Problems
von: Lee, Jongyeong, et al.
Veröffentlicht: (2025)
von: Lee, Jongyeong, et al.
Veröffentlicht: (2025)
Episodic Contextual Bandits with Knapsacks under Conversion Models
von: Cheung, Wang Chi, et al.
Veröffentlicht: (2025)
von: Cheung, Wang Chi, et al.
Veröffentlicht: (2025)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
von: Kang, Yue, et al.
Veröffentlicht: (2025)
von: Kang, Yue, et al.
Veröffentlicht: (2025)
Feel-Good Thompson Sampling for Contextual Dueling Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
Efficient Contextual Bandits with Uninformed Feedback Graphs
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2024)
Taming the Monster Every Context: Complexity Measure and Unified Framework for Offline-Oracle Efficient Contextual Bandits
von: Qin, Hao, et al.
Veröffentlicht: (2026)
von: Qin, Hao, et al.
Veröffentlicht: (2026)
Sparse Nonparametric Contextual Bandits
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
Federated Linear Contextual Bandits with Heterogeneous Clients
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
PAC Off-Policy Prediction of Contextual Bandits
von: Wan, Yilong, et al.
Veröffentlicht: (2025)
von: Wan, Yilong, et al.
Veröffentlicht: (2025)
Active Learning for Stochastic Contextual Linear Bandits
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
Online Multi-LLM Selection via Contextual Bandits under Unstructured Context Evolution
von: Poon, Manhin, et al.
Veröffentlicht: (2025)
von: Poon, Manhin, et al.
Veröffentlicht: (2025)
A Simple Reduction Scheme for Constrained Contextual Bandits with Adversarial Contexts via Regression
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
Soft Label PU Learning
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
von: Zhao, Puning, et al.
Veröffentlicht: (2024)
Attack-Resistant Uniform Fairness for Linear and Smooth Contextual Bandits
von: Zhang, Qingwen, et al.
Veröffentlicht: (2026)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2026)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
On Theoretical Limits of Learning with Label Differential Privacy
von: Zhao, Puning, et al.
Veröffentlicht: (2025) -
Consistent Estimation of Numerical Distributions under Local Differential Privacy by Wavelet Expansion
von: Zhao, Puning, et al.
Veröffentlicht: (2025) -
Learning with User-Level Local Differential Privacy
von: Zhao, Puning, et al.
Veröffentlicht: (2024) -
Sequential Federated Learning in Hierarchical Architecture on Non-IID Datasets
von: Yan, Xingrun, et al.
Veröffentlicht: (2024) -
Enhancing Learning with Label Differential Privacy by Vector Approximation
von: Zhao, Puning, et al.
Veröffentlicht: (2024)