Safe Linear Bandits over Unknown Polytopes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gangrade, Aditya, Chen, Tianrui, Saligrama, Venkatesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Constrained Linear Thompson Sampling
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025)
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025)
Testing the Feasibility of Linear Programs with Bandit Feedback
von: Gangrade, Aditya, et al.
Veröffentlicht: (2024)
von: Gangrade, Aditya, et al.
Veröffentlicht: (2024)
Linear Transformers Implicitly Discover Unified Numerical Algorithms
von: Lutz, Patrick, et al.
Veröffentlicht: (2025)
von: Lutz, Patrick, et al.
Veröffentlicht: (2025)
Data Deletion Can Help in Adaptive RL
von: Budhraja, Param, et al.
Veröffentlicht: (2026)
von: Budhraja, Param, et al.
Veröffentlicht: (2026)
Symmetry Reveals Layerwise Dynamics: How Transformers Perform In-Context Classification
von: Lutz, Patrick, et al.
Veröffentlicht: (2026)
von: Lutz, Patrick, et al.
Veröffentlicht: (2026)
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
von: Zhu, Ruizhao, et al.
Veröffentlicht: (2024)
von: Zhu, Ruizhao, et al.
Veröffentlicht: (2024)
Directional Optimism for Safe Linear Bandits
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2023)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2023)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
von: Kang, Yue, et al.
Veröffentlicht: (2025)
von: Kang, Yue, et al.
Veröffentlicht: (2025)
Linear Causal Bandits: Unknown Graph and Soft Interventions
von: Yan, Zirui, et al.
Veröffentlicht: (2024)
von: Yan, Zirui, et al.
Veröffentlicht: (2024)
SPARC: Score Prompting and Adaptive Fusion for Zero-Shot Multi-Label Recognition in Vision-Language Models
von: Miller, Kevin, et al.
Veröffentlicht: (2025)
von: Miller, Kevin, et al.
Veröffentlicht: (2025)
PolytopeWalk: Sparse MCMC Sampling over Polytopes
von: Sun, Benny, et al.
Veröffentlicht: (2024)
von: Sun, Benny, et al.
Veröffentlicht: (2024)
Label Noise: Ignorance Is Bliss
von: Zhu, Yilun, et al.
Veröffentlicht: (2024)
von: Zhu, Yilun, et al.
Veröffentlicht: (2024)
Causal Bandits: The Pareto Optimal Frontier of Adaptivity, a Reduction to Linear Bandits, and Limitations around Unknown Marginals
von: Liu, Ziyi, et al.
Veröffentlicht: (2024)
von: Liu, Ziyi, et al.
Veröffentlicht: (2024)
Improved Algorithm for Adversarial Linear Mixture MDPs with Bandit Feedback and Unknown Transition
von: Li, Long-Fei, et al.
Veröffentlicht: (2024)
von: Li, Long-Fei, et al.
Veröffentlicht: (2024)
Linear Convergence of the Frank-Wolfe Algorithm over Product Polytopes
von: Iommazzo, Gabriele, et al.
Veröffentlicht: (2025)
von: Iommazzo, Gabriele, et al.
Veröffentlicht: (2025)
Learning to Explore with Lagrangians for Bandits under Unknown Linear Constraints
von: Das, Udvas, et al.
Veröffentlicht: (2024)
von: Das, Udvas, et al.
Veröffentlicht: (2024)
Universal Inference Meets Random Projections: A Scalable Test for Log-concavity
von: Dunn, Robin, et al.
Veröffentlicht: (2021)
von: Dunn, Robin, et al.
Veröffentlicht: (2021)
Domain Generalization Under Posterior Drift
von: Zhu, Yilun, et al.
Veröffentlicht: (2025)
von: Zhu, Yilun, et al.
Veröffentlicht: (2025)
Robust Linear Dueling Bandits with Post-serving Context under Unknown Delays and Adversarial Corruptions
von: Oh, Youngmin
Veröffentlicht: (2026)
von: Oh, Youngmin
Veröffentlicht: (2026)
Beating Adversarial Low-Rank MDPs with Unknown Transition and Bandit Feedback
von: Liu, Haolin, et al.
Veröffentlicht: (2024)
von: Liu, Haolin, et al.
Veröffentlicht: (2024)
Time-Varying Gaussian Process Bandits with Unknown Prior
von: Ziomek, Juliusz, et al.
Veröffentlicht: (2024)
von: Ziomek, Juliusz, et al.
Veröffentlicht: (2024)
Faster Sampling from Log-Concave Densities over Polytopes via Efficient Linear Solvers
von: Mangoubi, Oren, et al.
Veröffentlicht: (2024)
von: Mangoubi, Oren, et al.
Veröffentlicht: (2024)
Adversarial Bandit over Bandits: Hierarchical Bandits for Online Configuration Management
von: Avin, Chen, et al.
Veröffentlicht: (2025)
von: Avin, Chen, et al.
Veröffentlicht: (2025)
Transformers in the Dark: Navigating Unknown Search Spaces via Bandit Feedback
von: Kim, Jungtaek, et al.
Veröffentlicht: (2026)
von: Kim, Jungtaek, et al.
Veröffentlicht: (2026)
HR-Bandit: Human-AI Collaborated Linear Recourse Bandit
von: Cao, Junyu, et al.
Veröffentlicht: (2024)
von: Cao, Junyu, et al.
Veröffentlicht: (2024)
Bayesian Bandit Algorithms with Approximate Inference in Stochastic Linear Bandits
von: Huang, Ziyi, et al.
Veröffentlicht: (2024)
von: Huang, Ziyi, et al.
Veröffentlicht: (2024)
Infrequent Exploration in Linear Bandits
von: Lee, Harin, et al.
Veröffentlicht: (2025)
von: Lee, Harin, et al.
Veröffentlicht: (2025)
Optimal Thresholding Linear Bandit
von: Rivera, Eduardo Ochoa, et al.
Veröffentlicht: (2024)
von: Rivera, Eduardo Ochoa, et al.
Veröffentlicht: (2024)
Federated Linear Dueling Bandits
von: Huang, Xuhan, et al.
Veröffentlicht: (2025)
von: Huang, Xuhan, et al.
Veröffentlicht: (2025)
Introducing Graph Learning over Polytopic Uncertain Graph
von: Kishida, Masako, et al.
Veröffentlicht: (2024)
von: Kishida, Masako, et al.
Veröffentlicht: (2024)
Restless Linear Bandits
von: Khaleghi, Azadeh
Veröffentlicht: (2024)
von: Khaleghi, Azadeh
Veröffentlicht: (2024)
High Probability Bound for Cross-Learning Contextual Bandits with Unknown Context Distributions
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2024)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
Linear Contextual Bandits with Interference
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Generalized Linear Bandits with Limited Adaptivity
von: Sawarni, Ayush, et al.
Veröffentlicht: (2024)
von: Sawarni, Ayush, et al.
Veröffentlicht: (2024)
Symmetric Linear Bandits with Hidden Symmetry
von: Tran, Nam Phuong, et al.
Veröffentlicht: (2024)
von: Tran, Nam Phuong, et al.
Veröffentlicht: (2024)
Pure Exploration in Bandits with Linear Constraints
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
Linear Bandits with Partially Observable Features
von: Kim, Wonyoung, et al.
Veröffentlicht: (2025)
von: Kim, Wonyoung, et al.
Veröffentlicht: (2025)
Contextual Linear Bandits with Delay as Payoff
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
Sparse Linear Bandits with Blocking Constraints
von: Jain, Adit, et al.
Veröffentlicht: (2024)
von: Jain, Adit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Constrained Linear Thompson Sampling
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025) -
Testing the Feasibility of Linear Programs with Bandit Feedback
von: Gangrade, Aditya, et al.
Veröffentlicht: (2024) -
Linear Transformers Implicitly Discover Unified Numerical Algorithms
von: Lutz, Patrick, et al.
Veröffentlicht: (2025) -
Data Deletion Can Help in Adaptive RL
von: Budhraja, Param, et al.
Veröffentlicht: (2026) -
Symmetry Reveals Layerwise Dynamics: How Transformers Perform In-Context Classification
von: Lutz, Patrick, et al.
Veröffentlicht: (2026)