Directional Optimism for Safe Linear Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Hutchinson, Spencer, Turan, Berkay, Alizadeh, Mahnoosh |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safe Online Convex Optimization with Multi-Point Feedback
by: Hutchinson, Spencer, et al.
Published: (2024)
by: Hutchinson, Spencer, et al.
Published: (2024)
A Safe First-Order Method for Pricing-Based Resource Allocation in Safety-Critical Networks
by: Turan, Berkay, et al.
Published: (2023)
by: Turan, Berkay, et al.
Published: (2023)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
by: Hutchinson, Spencer, et al.
Published: (2026)
by: Hutchinson, Spencer, et al.
Published: (2026)
The Safety-Privacy Tradeoff in Linear Bandits
by: Zibaie, Arghavan, et al.
Published: (2025)
by: Zibaie, Arghavan, et al.
Published: (2025)
Constrained Online Convex Optimization with Polyak Feasibility Steps
by: Hutchinson, Spencer, et al.
Published: (2025)
by: Hutchinson, Spencer, et al.
Published: (2025)
Optimistic Safety for Online Convex Optimization with Unknown Linear Constraints
by: Hutchinson, Spencer, et al.
Published: (2024)
by: Hutchinson, Spencer, et al.
Published: (2024)
Online Nonstochastic Control with Convex Safety Constraints
by: Jiang, Nanfei, et al.
Published: (2025)
by: Jiang, Nanfei, et al.
Published: (2025)
Decentralized Low-Rank Fine-Tuning of Large Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2025)
by: Ghiasvand, Sajjad, et al.
Published: (2025)
Stochastic Gradient Descent with Strategic Querying
by: Jiang, Nanfei, et al.
Published: (2025)
by: Jiang, Nanfei, et al.
Published: (2025)
Optimism in the Face of Ambiguity Principle for Multi-Armed Bandits
by: Li, Mengmeng, et al.
Published: (2024)
by: Li, Mengmeng, et al.
Published: (2024)
pFedMMA: Personalized Federated Fine-Tuning with Multi-Modal Adapter for Vision-Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2025)
by: Ghiasvand, Sajjad, et al.
Published: (2025)
REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations
by: Ghiasvand, Sajjad, et al.
Published: (2026)
by: Ghiasvand, Sajjad, et al.
Published: (2026)
Safe Linear Bandits over Unknown Polytopes
by: Gangrade, Aditya, et al.
Published: (2022)
by: Gangrade, Aditya, et al.
Published: (2022)
Few-Shot Adversarial Low-Rank Fine-Tuning of Vision-Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2025)
by: Ghiasvand, Sajjad, et al.
Published: (2025)
Steady-state Based Approach to Online Non-stochastic Control
by: Hebbar, Vijeth, et al.
Published: (2026)
by: Hebbar, Vijeth, et al.
Published: (2026)
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
by: Praharaj, Samya, et al.
Published: (2025)
by: Praharaj, Samya, et al.
Published: (2025)
Robust Decentralized Learning with Local Updates and Gradient Tracking
by: Ghiasvand, Sajjad, et al.
Published: (2024)
by: Ghiasvand, Sajjad, et al.
Published: (2024)
Upper Counterfactual Confidence Bounds: a New Optimism Principle for Contextual Bandits
by: Xu, Yunbei, et al.
Published: (2020)
by: Xu, Yunbei, et al.
Published: (2020)
MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation
by: Ghiasvand, Sajjad, et al.
Published: (2026)
by: Ghiasvand, Sajjad, et al.
Published: (2026)
Direction-Aware Offline-to-Online Learning in Linear Contextual Bandits
by: Han, Zean, et al.
Published: (2026)
by: Han, Zean, et al.
Published: (2026)
Communication-Efficient and Tensorized Federated Fine-Tuning of Large Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2024)
by: Ghiasvand, Sajjad, et al.
Published: (2024)
Why Most Optimism Bandit Algorithms Have the Same Regret Analysis: A Simple Unifying Theorem
by: Krishnamurthy, Vikram
Published: (2025)
by: Krishnamurthy, Vikram
Published: (2025)
Improved Regret Bound for Safe Reinforcement Learning via Tighter Cost Pessimism and Reward Optimism
by: Yu, Kihyun, et al.
Published: (2024)
by: Yu, Kihyun, et al.
Published: (2024)
Randomised Optimism via Competitive Co-Evolution for Matrix Games with Bandit Feedback
by: Lin, Shishen
Published: (2025)
by: Lin, Shishen
Published: (2025)
Data-Driven Policy Mapping for Safe RL-based Energy Management Systems
by: Zangato, Theo, et al.
Published: (2025)
by: Zangato, Theo, et al.
Published: (2025)
Fast State-Augmented Learning for Wireless Resource Allocation with Dual Variable Regression
by: Uslu, Yigit Berkay, et al.
Published: (2025)
by: Uslu, Yigit Berkay, et al.
Published: (2025)
Learning to Slice Wi-Fi Networks: A State-Augmented Primal-Dual Approach
by: Uslu, Yiğit Berkay, et al.
Published: (2024)
by: Uslu, Yiğit Berkay, et al.
Published: (2024)
HR-Bandit: Human-AI Collaborated Linear Recourse Bandit
by: Cao, Junyu, et al.
Published: (2024)
by: Cao, Junyu, et al.
Published: (2024)
Bayesian Bandit Algorithms with Approximate Inference in Stochastic Linear Bandits
by: Huang, Ziyi, et al.
Published: (2024)
by: Huang, Ziyi, et al.
Published: (2024)
Infrequent Exploration in Linear Bandits
by: Lee, Harin, et al.
Published: (2025)
by: Lee, Harin, et al.
Published: (2025)
Optimal Thresholding Linear Bandit
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
Federated Linear Dueling Bandits
by: Huang, Xuhan, et al.
Published: (2025)
by: Huang, Xuhan, et al.
Published: (2025)
Asymptotic Optimism of Random-Design Linear and Kernel Regression Models
by: Luo, Hengrui, et al.
Published: (2025)
by: Luo, Hengrui, et al.
Published: (2025)
Minimax Optimal Reinforcement Learning with Quasi-Optimism
by: Lee, Harin, et al.
Published: (2025)
by: Lee, Harin, et al.
Published: (2025)
Beyond Optimism: Exploration With Partially Observable Rewards
by: Parisi, Simone, et al.
Published: (2024)
by: Parisi, Simone, et al.
Published: (2024)
Restless Linear Bandits
by: Khaleghi, Azadeh
Published: (2024)
by: Khaleghi, Azadeh
Published: (2024)
Linear Contextual Bandits with Interference
by: Xu, Yang, et al.
Published: (2024)
by: Xu, Yang, et al.
Published: (2024)
Pure Exploration in Bandits with Linear Constraints
by: Carlsson, Emil, et al.
Published: (2023)
by: Carlsson, Emil, et al.
Published: (2023)
Robust Causal Bandits for Linear Models
by: Yan, Zirui, et al.
Published: (2023)
by: Yan, Zirui, et al.
Published: (2023)
Generalized Linear Bandits with Limited Adaptivity
by: Sawarni, Ayush, et al.
Published: (2024)
by: Sawarni, Ayush, et al.
Published: (2024)
Similar Items
-
Safe Online Convex Optimization with Multi-Point Feedback
by: Hutchinson, Spencer, et al.
Published: (2024) -
A Safe First-Order Method for Pricing-Based Resource Allocation in Safety-Critical Networks
by: Turan, Berkay, et al.
Published: (2023) -
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
by: Hutchinson, Spencer, et al.
Published: (2026) -
The Safety-Privacy Tradeoff in Linear Bandits
by: Zibaie, Arghavan, et al.
Published: (2025) -
Constrained Online Convex Optimization with Polyak Feasibility Steps
by: Hutchinson, Spencer, et al.
Published: (2025)