A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Sanghwa, Lee, Junghyun, Yun, Se-Young |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Unified Confidence Sequence for Generalized Linear Models, with Applications to Bandits
di: Lee, Junghyun, et al.
Pubblicazione: (2024)
di: Lee, Junghyun, et al.
Pubblicazione: (2024)
Instance-Optimal Estimation with Multiple LLM Judges on a Budget
di: Lee, Junghyun, et al.
Pubblicazione: (2026)
di: Lee, Junghyun, et al.
Pubblicazione: (2026)
Flooding with Absorption: An Efficient Protocol for Heterogeneous Bandits over Complex Networks
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
Adversarial Bandits against Arbitrary Strategies
di: Kim, Jung-hun, et al.
Pubblicazione: (2022)
di: Kim, Jung-hun, et al.
Pubblicazione: (2022)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
Near-Optimal Clustering in Mixture of Markov Chains
di: Lee, Junghyun, et al.
Pubblicazione: (2025)
di: Lee, Junghyun, et al.
Pubblicazione: (2025)
GL-LowPopArt: A Nearly Instance-Wise Minimax-Optimal Estimator for Generalized Low-Rank Trace Regression
di: Lee, Junghyun, et al.
Pubblicazione: (2025)
di: Lee, Junghyun, et al.
Pubblicazione: (2025)
Efficient and Optimal Policy Gradient Algorithm for Corrupted Multi-armed Bandits
di: Liu, Jiayuan, et al.
Pubblicazione: (2025)
di: Liu, Jiayuan, et al.
Pubblicazione: (2025)
An Adaptive Approach for Infinitely Many-armed Bandits under Generalized Rotting Constraints
di: Kim, Jung-hun, et al.
Pubblicazione: (2024)
di: Kim, Jung-hun, et al.
Pubblicazione: (2024)
Robust and Computationally Efficient Linear Contextual Bandits under Adversarial Corruption and Heavy-Tailed Noise
di: Tani, Naoto, et al.
Pubblicazione: (2026)
di: Tani, Naoto, et al.
Pubblicazione: (2026)
Probability-Flow ODE in Infinite-Dimensional Function Spaces
di: Na, Kunwoo, et al.
Pubblicazione: (2025)
di: Na, Kunwoo, et al.
Pubblicazione: (2025)
Cascading Bandits Robust to Adversarial Corruptions
di: Xie, Jize, et al.
Pubblicazione: (2025)
di: Xie, Jize, et al.
Pubblicazione: (2025)
Corruption-Robust Linear Bandits: Minimax Optimality and Gap-Dependent Misspecification
di: Liu, Haolin, et al.
Pubblicazione: (2024)
di: Liu, Haolin, et al.
Pubblicazione: (2024)
Regularized Online RLHF with Generalized Bilinear Preferences
di: Lee, Junghyun, et al.
Pubblicazione: (2026)
di: Lee, Junghyun, et al.
Pubblicazione: (2026)
Nearly-Optimal Algorithm for Adversarial Kernelized Bandits
di: Iwazaki, Shogo
Pubblicazione: (2026)
di: Iwazaki, Shogo
Pubblicazione: (2026)
On the Optimality of Tracking Fisher Information in Adaptive Testing with Stochastic Binary Responses
di: Kim, Sanghwa, et al.
Pubblicazione: (2025)
di: Kim, Sanghwa, et al.
Pubblicazione: (2025)
Contextual Linear Bandits under Noisy Features: Towards Bayesian Oracles
di: Kim, Jung-hun, et al.
Pubblicazione: (2017)
di: Kim, Jung-hun, et al.
Pubblicazione: (2017)
Multi-Agent Stochastic Bandits Robust to Adversarial Corruptions
di: Ghaffari, Fatemeh, et al.
Pubblicazione: (2024)
di: Ghaffari, Fatemeh, et al.
Pubblicazione: (2024)
Robust Linear Dueling Bandits with Post-serving Context under Unknown Delays and Adversarial Corruptions
di: Oh, Youngmin
Pubblicazione: (2026)
di: Oh, Youngmin
Pubblicazione: (2026)
Optimal and Practical Batched Linear Bandit Algorithm
di: Yu, Sanghoon, et al.
Pubblicazione: (2025)
di: Yu, Sanghoon, et al.
Pubblicazione: (2025)
Linear Bandits on Ellipsoids: Minimax Optimal Algorithms
di: Zhang, Raymond, et al.
Pubblicazione: (2025)
di: Zhang, Raymond, et al.
Pubblicazione: (2025)
An Improved Algorithm for Adversarial Linear Contextual Bandits via Reduction
di: van Erven, Tim, et al.
Pubblicazione: (2025)
di: van Erven, Tim, et al.
Pubblicazione: (2025)
LinearAPT: An Adaptive Algorithm for the Fixed-Budget Thresholding Linear Bandit Problem
di: Wu, Yun-Ang, et al.
Pubblicazione: (2024)
di: Wu, Yun-Ang, et al.
Pubblicazione: (2024)
FlickerFusion: Intra-trajectory Domain Generalizing Multi-Agent RL
di: Koh, Woosung, et al.
Pubblicazione: (2024)
di: Koh, Woosung, et al.
Pubblicazione: (2024)
Constructing Adversarial Examples for Vertical Federated Learning: Optimal Client Corruption through Multi-Armed Bandit
di: Yao, Duanyi, et al.
Pubblicazione: (2024)
di: Yao, Duanyi, et al.
Pubblicazione: (2024)
Multi-Task Corrupted Prediction for Learning Robust Audio-Visual Speech Representation
di: Kim, Sungnyun, et al.
Pubblicazione: (2025)
di: Kim, Sungnyun, et al.
Pubblicazione: (2025)
Nearly Optimal Algorithms for Contextual Dueling Bandits from Adversarial Feedback
di: Di, Qiwei, et al.
Pubblicazione: (2024)
di: Di, Qiwei, et al.
Pubblicazione: (2024)
Slowly Changing Adversarial Bandit Algorithms are Efficient for Discounted MDPs
di: Kash, Ian A., et al.
Pubblicazione: (2022)
di: Kash, Ian A., et al.
Pubblicazione: (2022)
Improved Algorithm for Adversarial Linear Mixture MDPs with Bandit Feedback and Unknown Transition
di: Li, Long-Fei, et al.
Pubblicazione: (2024)
di: Li, Long-Fei, et al.
Pubblicazione: (2024)
Shuffle and Joint Differential Privacy for Generalized Linear Contextual Bandits
di: Sarmasarkar, Sahasrajit
Pubblicazione: (2026)
di: Sarmasarkar, Sahasrajit
Pubblicazione: (2026)
Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates
di: Yu, Sanghoon, et al.
Pubblicazione: (2026)
di: Yu, Sanghoon, et al.
Pubblicazione: (2026)
Optimal Thresholding Linear Bandit
di: Rivera, Eduardo Ochoa, et al.
Pubblicazione: (2024)
di: Rivera, Eduardo Ochoa, et al.
Pubblicazione: (2024)
A Near-optimal, Scalable and Parallelizable Framework for Stochastic Bandits Robust to Adversarial Corruptions and Beyond
di: Hu, Zicheng, et al.
Pubblicazione: (2025)
di: Hu, Zicheng, et al.
Pubblicazione: (2025)
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
di: Jin, Tianyuan, et al.
Pubblicazione: (2024)
di: Jin, Tianyuan, et al.
Pubblicazione: (2024)
Near-Optimal Regret in Adversarial Kernel Bandits
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2026)
di: Zhang, Yu-Jie, et al.
Pubblicazione: (2026)
Bayesian Bandit Algorithms with Approximate Inference in Stochastic Linear Bandits
di: Huang, Ziyi, et al.
Pubblicazione: (2024)
di: Huang, Ziyi, et al.
Pubblicazione: (2024)
Minimax Rate-Optimal Algorithms for High-Dimensional Stochastic Linear Bandits
di: Liu, Jingyu, et al.
Pubblicazione: (2025)
di: Liu, Jingyu, et al.
Pubblicazione: (2025)
Optimal Clustering from Noisy Binary Feedback
di: Ariu, Kaito, et al.
Pubblicazione: (2019)
di: Ariu, Kaito, et al.
Pubblicazione: (2019)
Optimal Batched Linear Bandits
di: Ren, Xuanfei, et al.
Pubblicazione: (2024)
di: Ren, Xuanfei, et al.
Pubblicazione: (2024)
Corruption-Robust Algorithms with Uncertainty Weighting for Nonlinear Contextual Bandits and Markov Decision Processes
di: Ye, Chenlu, et al.
Pubblicazione: (2022)
di: Ye, Chenlu, et al.
Pubblicazione: (2022)
Documenti analoghi
-
A Unified Confidence Sequence for Generalized Linear Models, with Applications to Bandits
di: Lee, Junghyun, et al.
Pubblicazione: (2024) -
Instance-Optimal Estimation with Multiple LLM Judges on a Budget
di: Lee, Junghyun, et al.
Pubblicazione: (2026) -
Flooding with Absorption: An Efficient Protocol for Heterogeneous Bandits over Complex Networks
di: Lee, Junghyun, et al.
Pubblicazione: (2023) -
Adversarial Bandits against Arbitrary Strategies
di: Kim, Jung-hun, et al.
Pubblicazione: (2022) -
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
di: Lee, Junghyun, et al.
Pubblicazione: (2023)