Adversarial bandit optimization for approximately linear functions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Zhuoyu, Hatano, Kohei, Takimoto, Eiji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial Bandit Optimization with Globally Bounded Perturbations to Linear Losses
von: Cheng, Zhuoyu, et al.
Veröffentlicht: (2026)
von: Cheng, Zhuoyu, et al.
Veröffentlicht: (2026)
Multi-thresholding Good Arm Identification with Bandit Feedback
von: Jiang, Xuanke, et al.
Veröffentlicht: (2025)
von: Jiang, Xuanke, et al.
Veröffentlicht: (2025)
Spectral bandits
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Linear bandits with polylogarithmic minimax regret
von: Lumbreras, Josep, et al.
Veröffentlicht: (2024)
von: Lumbreras, Josep, et al.
Veröffentlicht: (2024)
Functional multi-armed bandit and the best function identification problems
von: Dorn, Yuriy, et al.
Veröffentlicht: (2025)
von: Dorn, Yuriy, et al.
Veröffentlicht: (2025)
Prior-informed optimization of treatment recommendation via bandit algorithms trained on large language model-processed historical records
von: Nessari, Saman, et al.
Veröffentlicht: (2025)
von: Nessari, Saman, et al.
Veröffentlicht: (2025)
Reinforcement learning with combinatorial actions for coupled restless bandits
von: Xu, Lily, et al.
Veröffentlicht: (2025)
von: Xu, Lily, et al.
Veröffentlicht: (2025)
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2023)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2023)
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
von: Patil, Gandharv, et al.
Veröffentlicht: (2022)
von: Patil, Gandharv, et al.
Veröffentlicht: (2022)
Softmax gradient policy for variance minimization and risk-averse multi armed bandits
von: Turinici, Gabriel
Veröffentlicht: (2026)
von: Turinici, Gabriel
Veröffentlicht: (2026)
On the optimal regret of collaborative personalized linear bandits
von: Huang, Bruce, et al.
Veröffentlicht: (2025)
von: Huang, Bruce, et al.
Veröffentlicht: (2025)
On the Rate of Convergence of GD in Non-linear Neural Networks: An Adversarial Robustness Perspective
von: Smorodinsky, Guy, et al.
Veröffentlicht: (2026)
von: Smorodinsky, Guy, et al.
Veröffentlicht: (2026)
Understanding the theoretical properties of projected Bellman equation, linear Q-learning, and approximate value iteration
von: Lim, Han-Dong, et al.
Veröffentlicht: (2025)
von: Lim, Han-Dong, et al.
Veröffentlicht: (2025)
Reward-Punishment Reinforcement Learning with Maximum Entropy
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
Adversarial Graph Disentanglement
von: Zheng, Shuai, et al.
Veröffentlicht: (2021)
von: Zheng, Shuai, et al.
Veröffentlicht: (2021)
LLM-ABBA: Understanding time series via symbolic approximation
von: Chen, Xinye, et al.
Veröffentlicht: (2024)
von: Chen, Xinye, et al.
Veröffentlicht: (2024)
Adversarial Robustness Overestimation and Instability in TRADES
von: Li, Jonathan Weiping, et al.
Veröffentlicht: (2024)
von: Li, Jonathan Weiping, et al.
Veröffentlicht: (2024)
Frugality in second-order optimization: floating-point approximations for Newton's method
von: Carrino, Giuseppe, et al.
Veröffentlicht: (2025)
von: Carrino, Giuseppe, et al.
Veröffentlicht: (2025)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
von: Belaire, Roman, et al.
Veröffentlicht: (2024)
von: Belaire, Roman, et al.
Veröffentlicht: (2024)
V2X-VLM: End-to-End V2X Cooperative Autonomous Driving Through Large Vision-Language Models
von: You, Junwei, et al.
Veröffentlicht: (2024)
von: You, Junwei, et al.
Veröffentlicht: (2024)
MemLoss: Enhancing Adversarial Training with Recycling Adversarial Examples
von: Mahdi, Soroush, et al.
Veröffentlicht: (2025)
von: Mahdi, Soroush, et al.
Veröffentlicht: (2025)
Towards Interpretable Adversarial Examples via Sparse Adversarial Attack
von: Lin, Fudong, et al.
Veröffentlicht: (2025)
von: Lin, Fudong, et al.
Veröffentlicht: (2025)
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
von: Liu, Qihao, et al.
Veröffentlicht: (2025)
von: Liu, Qihao, et al.
Veröffentlicht: (2025)
Adversarial Examples Might be Avoidable: The Role of Data Concentration in Adversarial Robustness
von: Pal, Ambar, et al.
Veröffentlicht: (2023)
von: Pal, Ambar, et al.
Veröffentlicht: (2023)
How Worst-Case Are Adversarial Attacks? Linking Adversarial and Perturbation Robustness
von: Rossolini, Giulio
Veröffentlicht: (2026)
von: Rossolini, Giulio
Veröffentlicht: (2026)
Discriminative Adversarial Unlearning
von: Sharma, Rohan, et al.
Veröffentlicht: (2024)
von: Sharma, Rohan, et al.
Veröffentlicht: (2024)
Adversarial Reasoning at Jailbreaking Time
von: Sabbaghi, Mahdi, et al.
Veröffentlicht: (2025)
von: Sabbaghi, Mahdi, et al.
Veröffentlicht: (2025)
Conflict-Aware Adversarial Training
von: Xue, Zhiyu, et al.
Veröffentlicht: (2024)
von: Xue, Zhiyu, et al.
Veröffentlicht: (2024)
Adversarially Robust Decision Transformer
von: Tang, Xiaohang, et al.
Veröffentlicht: (2024)
von: Tang, Xiaohang, et al.
Veröffentlicht: (2024)
Adversarial Training: A Survey
von: Zhao, Mengnan, et al.
Veröffentlicht: (2024)
von: Zhao, Mengnan, et al.
Veröffentlicht: (2024)
Adversarial Attacks on Hyperbolic Networks
von: van Spengler, Max, et al.
Veröffentlicht: (2024)
von: van Spengler, Max, et al.
Veröffentlicht: (2024)
A Closer Look at Adversarial Suffix Learning for Jailbreaking LLMs: Augmented Adversarial Trigger Learning
von: Wang, Zhe, et al.
Veröffentlicht: (2025)
von: Wang, Zhe, et al.
Veröffentlicht: (2025)
FairSTG: Countering performance heterogeneity via collaborative sample-level optimization
von: Lin, Gengyu, et al.
Veröffentlicht: (2024)
von: Lin, Gengyu, et al.
Veröffentlicht: (2024)
Adversarial Diffusion for Robust Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
Adversarial Training for Process Reward Models
von: Juneja, Gurusha, et al.
Veröffentlicht: (2025)
von: Juneja, Gurusha, et al.
Veröffentlicht: (2025)
Algorithms for Adversarially Robust Deep Learning
von: Robey, Alexander
Veröffentlicht: (2025)
von: Robey, Alexander
Veröffentlicht: (2025)
Single Domain Generalization with Adversarial Memory
von: Yan, Hao, et al.
Veröffentlicht: (2025)
von: Yan, Hao, et al.
Veröffentlicht: (2025)
Maintaining Adversarial Robustness in Continuous Learning
von: Ru, Xiaolei, et al.
Veröffentlicht: (2024)
von: Ru, Xiaolei, et al.
Veröffentlicht: (2024)
Adversarial Imitation Learning via Boosting
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)
Decentralized Adversarial Training over Graphs
von: Cao, Ying, et al.
Veröffentlicht: (2023)
von: Cao, Ying, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Adversarial Bandit Optimization with Globally Bounded Perturbations to Linear Losses
von: Cheng, Zhuoyu, et al.
Veröffentlicht: (2026) -
Multi-thresholding Good Arm Identification with Bandit Feedback
von: Jiang, Xuanke, et al.
Veröffentlicht: (2025) -
Spectral bandits
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026) -
Linear bandits with polylogarithmic minimax regret
von: Lumbreras, Josep, et al.
Veröffentlicht: (2024) -
Functional multi-armed bandit and the best function identification problems
von: Dorn, Yuriy, et al.
Veröffentlicht: (2025)