Adversarial Bandits against Arbitrary Strategies
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jung-hun, Yun, Se-Young |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Adaptive Approach for Infinitely Many-armed Bandits under Generalized Rotting Constraints
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
Tracking Most Significant Shifts in Infinite-Armed Bandits
von: Suk, Joe, et al.
Veröffentlicht: (2025)
von: Suk, Joe, et al.
Veröffentlicht: (2025)
Contextual Linear Bandits under Noisy Features: Towards Bayesian Oracles
von: Kim, Jung-hun, et al.
Veröffentlicht: (2017)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2017)
Queueing Matching Bandits with Preference Feedback
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
Stochastic Matching Bandits with Rare Optimization Updates
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026)
von: Kim, Sanghwa, et al.
Veröffentlicht: (2026)
Oracle-Efficient Combinatorial Semi-Bandits
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
Learning in Prophet Inequalities with Noisy Observations
von: Kim, Jung-hun, et al.
Veröffentlicht: (2026)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2026)
Dynamic Assortment Selection and Pricing with Censored Preference Feedback
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)
A Unified Confidence Sequence for Generalized Linear Models, with Applications to Bandits
von: Lee, Junghyun, et al.
Veröffentlicht: (2024)
von: Lee, Junghyun, et al.
Veröffentlicht: (2024)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
Learning to Schedule in Parallel-Server Queues with Stochastic Bilinear Rewards
von: Kim, Jung-hun, et al.
Veröffentlicht: (2021)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2021)
Flooding with Absorption: An Efficient Protocol for Heterogeneous Bandits over Complex Networks
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
KLASS: KL-Guided Fast Inference in Masked Diffusion Models
von: Kim, Seo Hyun, et al.
Veröffentlicht: (2025)
von: Kim, Seo Hyun, et al.
Veröffentlicht: (2025)
Revisiting Early-Learning Regularization When Federated Learning Meets Noisy Labels
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
Online Conformal Abstention for Factuality Control Under Adversarial Bandit Feedback
von: Lee, Minjae, et al.
Veröffentlicht: (2025)
von: Lee, Minjae, et al.
Veröffentlicht: (2025)
Temporal Alignment Guidance: On-Manifold Sampling in Diffusion Models
von: Park, Youngrok, et al.
Veröffentlicht: (2025)
von: Park, Youngrok, et al.
Veröffentlicht: (2025)
What is the Alignment Objective of GRPO?
von: Vojnovic, Milan, et al.
Veröffentlicht: (2025)
von: Vojnovic, Milan, et al.
Veröffentlicht: (2025)
Diffusion-based Episodes Augmentation for Offline Multi-Agent Reinforcement Learning
von: Oh, Jihwan, et al.
Veröffentlicht: (2024)
von: Oh, Jihwan, et al.
Veröffentlicht: (2024)
Adversarial Multi-dueling Bandits
von: Gajane, Pratik
Veröffentlicht: (2024)
von: Gajane, Pratik
Veröffentlicht: (2024)
Adversarial Bandit over Bandits: Hierarchical Bandits for Online Configuration Management
von: Avin, Chen, et al.
Veröffentlicht: (2025)
von: Avin, Chen, et al.
Veröffentlicht: (2025)
On Characterizing Learnability for Adversarial Noisy Bandits
von: Hanneke, Steve, et al.
Veröffentlicht: (2026)
von: Hanneke, Steve, et al.
Veröffentlicht: (2026)
Adversarial Combinatorial Bandits with Switching Costs
von: Dong, Yanyan, et al.
Veröffentlicht: (2024)
von: Dong, Yanyan, et al.
Veröffentlicht: (2024)
Faster Rates for Private Adversarial Bandits
von: Asi, Hilal, et al.
Veröffentlicht: (2025)
von: Asi, Hilal, et al.
Veröffentlicht: (2025)
Cascading Bandits Robust to Adversarial Corruptions
von: Xie, Jize, et al.
Veröffentlicht: (2025)
von: Xie, Jize, et al.
Veröffentlicht: (2025)
Stochastic Bandits Robust to Adversarial Attacks
von: Wang, Xuchuang, et al.
Veröffentlicht: (2024)
von: Wang, Xuchuang, et al.
Veröffentlicht: (2024)
Constrained Contextual Bandits with Adversarial Contexts
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
von: Sarkar, Dhruv, et al.
Veröffentlicht: (2026)
Weisfeiler and Lehman Go Categorical
von: Choi, Seongjin, et al.
Veröffentlicht: (2026)
von: Choi, Seongjin, et al.
Veröffentlicht: (2026)
Automated Filtering of Human Feedback Data for Aligning Text-to-Image Diffusion Models
von: Yang, Yongjin, et al.
Veröffentlicht: (2024)
von: Yang, Yongjin, et al.
Veröffentlicht: (2024)
Hypernetwork-Driven Model Fusion for Federated Domain Generalization
von: Bartholet, Marc, et al.
Veröffentlicht: (2024)
von: Bartholet, Marc, et al.
Veröffentlicht: (2024)
BAPO: Base-Anchored Preference Optimization for Overcoming Forgetting in Large Language Models Personalization
von: Lee, Gihun, et al.
Veröffentlicht: (2024)
von: Lee, Gihun, et al.
Veröffentlicht: (2024)
DistiLLM: Towards Streamlined Distillation for Large Language Models
von: Ko, Jongwoo, et al.
Veröffentlicht: (2024)
von: Ko, Jongwoo, et al.
Veröffentlicht: (2024)
Deterministic Certification of Graph Neural Networks against Graph Poisoning Attacks with Arbitrary Perturbations
von: Li, Jiate, et al.
Veröffentlicht: (2025)
von: Li, Jiate, et al.
Veröffentlicht: (2025)
Nearly-Optimal Algorithm for Adversarial Kernelized Bandits
von: Iwazaki, Shogo
Veröffentlicht: (2026)
von: Iwazaki, Shogo
Veröffentlicht: (2026)
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
Near-Optimal Regret in Adversarial Kernel Bandits
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2026)
Bandits in Flux: Adversarial Constraints in Dynamic Environments
von: Salem, Tareq Si
Veröffentlicht: (2026)
von: Salem, Tareq Si
Veröffentlicht: (2026)
Optimal Clustering from Noisy Binary Feedback
von: Ariu, Kaito, et al.
Veröffentlicht: (2019)
von: Ariu, Kaito, et al.
Veröffentlicht: (2019)
Probability-Flow ODE in Infinite-Dimensional Function Spaces
von: Na, Kunwoo, et al.
Veröffentlicht: (2025)
von: Na, Kunwoo, et al.
Veröffentlicht: (2025)
Revisiting Instance-Optimal Cluster Recovery in the Labeled Stochastic Block Model
von: Ariu, Kaito, et al.
Veröffentlicht: (2023)
von: Ariu, Kaito, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
An Adaptive Approach for Infinitely Many-armed Bandits under Generalized Rotting Constraints
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024) -
Tracking Most Significant Shifts in Infinite-Armed Bandits
von: Suk, Joe, et al.
Veröffentlicht: (2025) -
Contextual Linear Bandits under Noisy Features: Towards Bayesian Oracles
von: Kim, Jung-hun, et al.
Veröffentlicht: (2017) -
Queueing Matching Bandits with Preference Feedback
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024) -
Stochastic Matching Bandits with Rare Optimization Updates
von: Kim, Jung-hun, et al.
Veröffentlicht: (2025)