Identifying All ε-Best Arms in (Misspecified) Linear Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zhekai, Ma, Tianyi, Hua, Cheng, Zhu, Ruihao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kernel $ε$-Greedy for Multi-Armed Bandits with Covariates
von: Arya, Sakshi, et al.
Veröffentlicht: (2023)
von: Arya, Sakshi, et al.
Veröffentlicht: (2023)
Asymptotic Optimism of Random-Design Linear and Kernel Regression Models
von: Luo, Hengrui, et al.
Veröffentlicht: (2025)
von: Luo, Hengrui, et al.
Veröffentlicht: (2025)
Sparse joint shift in multinomial classification
von: Tasche, Dirk
Veröffentlicht: (2023)
von: Tasche, Dirk
Veröffentlicht: (2023)
A Frequency-Domain Analysis of the Multi-Armed Bandit Problem: A New Perspective on the Exploration-Exploitation Trade-off
von: Zhang, Di
Veröffentlicht: (2025)
von: Zhang, Di
Veröffentlicht: (2025)
Optimal In-context Adaptivity and Distributional Robustness of Transformers
von: Ma, Tianyi, et al.
Veröffentlicht: (2025)
von: Ma, Tianyi, et al.
Veröffentlicht: (2025)
Model Monitoring: A General Framework with an Application to Non-life Insurance Pricing
von: Brauer, Alexej, et al.
Veröffentlicht: (2025)
von: Brauer, Alexej, et al.
Veröffentlicht: (2025)
Batched Nonparametric Bandits via k-Nearest Neighbor UCB
von: Arya, Sakshi
Veröffentlicht: (2025)
von: Arya, Sakshi
Veröffentlicht: (2025)
Conformal e-prediction in the presence of confounding
von: Vovk, Vladimir, et al.
Veröffentlicht: (2026)
von: Vovk, Vladimir, et al.
Veröffentlicht: (2026)
Universality of conformal prediction under the assumption of randomness
von: Vovk, Vladimir
Veröffentlicht: (2025)
von: Vovk, Vladimir
Veröffentlicht: (2025)
Information-theoretic Limits of Learning and Estimation
von: Gamal, Abbas El, et al.
Veröffentlicht: (2026)
von: Gamal, Abbas El, et al.
Veröffentlicht: (2026)
Efficient and Minimax Optimal In-context Nonparametric Regression with Transformers
von: Ching, Michelle, et al.
Veröffentlicht: (2026)
von: Ching, Michelle, et al.
Veröffentlicht: (2026)
When Do Credal Sets Stabilize? Fixed-Point Theorems for Credal Set Updates
von: Caprio, Michele, et al.
Veröffentlicht: (2025)
von: Caprio, Michele, et al.
Veröffentlicht: (2025)
Realizable Bayes-Consistency for General Metric Losses
von: Cohen, Dan Tsir, et al.
Veröffentlicht: (2026)
von: Cohen, Dan Tsir, et al.
Veröffentlicht: (2026)
Automated Model Selection for Generalized Linear Models
von: Schwendinger, Benjamin, et al.
Veröffentlicht: (2024)
von: Schwendinger, Benjamin, et al.
Veröffentlicht: (2024)
A Wasserstein perspective of Vanilla GANs
von: Kunkel, Lea, et al.
Veröffentlicht: (2024)
von: Kunkel, Lea, et al.
Veröffentlicht: (2024)
A PAC-Bayes oracle inequality for sparse neural networks
von: Steffen, Maximilian F., et al.
Veröffentlicht: (2022)
von: Steffen, Maximilian F., et al.
Veröffentlicht: (2022)
Kernel-based estimators for functional causal effects
von: Raykov, Yordan P., et al.
Veröffentlicht: (2025)
von: Raykov, Yordan P., et al.
Veröffentlicht: (2025)
Statistical Inference for Misspecified Contextual Bandits
von: Guo, Yongyi, et al.
Veröffentlicht: (2025)
von: Guo, Yongyi, et al.
Veröffentlicht: (2025)
Generalization of the Gibbs algorithm with high probability at low temperatures
von: Maurer, Andreas
Veröffentlicht: (2025)
von: Maurer, Andreas
Veröffentlicht: (2025)
Adaptive learning of density ratios in RKHS
von: Zellinger, Werner, et al.
Veröffentlicht: (2023)
von: Zellinger, Werner, et al.
Veröffentlicht: (2023)
Adversarial Subspace Generation for Outlier Detection in High-Dimensional Data
von: Cribeiro-Ramallo, Jose, et al.
Veröffentlicht: (2025)
von: Cribeiro-Ramallo, Jose, et al.
Veröffentlicht: (2025)
Towards the Best Solution for Complex System Reliability: Can Statistics Outperform Machine Learning?
von: Gamiz, Maria Luz, et al.
Veröffentlicht: (2024)
von: Gamiz, Maria Luz, et al.
Veröffentlicht: (2024)
On Lai's Upper Confidence Bound in Multi-Armed Bandits
von: Ren, Huachen, et al.
Veröffentlicht: (2024)
von: Ren, Huachen, et al.
Veröffentlicht: (2024)
Batched Single-Index Global Multi-Armed Bandits with Covariates
von: Arya, Sakshi, et al.
Veröffentlicht: (2025)
von: Arya, Sakshi, et al.
Veröffentlicht: (2025)
Active Causal Experimentalist (ACE): Learning Intervention Strategies via Direct Preference Optimization
von: Cooper, Patrick, et al.
Veröffentlicht: (2026)
von: Cooper, Patrick, et al.
Veröffentlicht: (2026)
Cross-Domain Uncertainty Quantification for Selective Prediction: A Comprehensive Bound Ablation with Transfer-Informed Betting
von: Basu, Abhinaba
Veröffentlicht: (2026)
von: Basu, Abhinaba
Veröffentlicht: (2026)
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
On the Convergence of the ELBO to Entropy Sums
von: Lücke, Jörg, et al.
Veröffentlicht: (2022)
von: Lücke, Jörg, et al.
Veröffentlicht: (2022)
Generative Models with ELBOs Converging to Entropy Sums
von: Warnken, Jan, et al.
Veröffentlicht: (2024)
von: Warnken, Jan, et al.
Veröffentlicht: (2024)
GRALIS: A Unified Canonical Framework for Linear Attribution Methods via Riesz Representation
von: Fanale, Raimondo
Veröffentlicht: (2026)
von: Fanale, Raimondo
Veröffentlicht: (2026)
Fundamental limits for weighted empirical approximations of tilted distributions
von: Iyer, Sarvesh Ravichandran, et al.
Veröffentlicht: (2025)
von: Iyer, Sarvesh Ravichandran, et al.
Veröffentlicht: (2025)
Realizing Scaling Laws in Recommender Systems: A Foundation-Expert Paradigm for Hyperscale Model Deployment
von: Li, Dai, et al.
Veröffentlicht: (2025)
von: Li, Dai, et al.
Veröffentlicht: (2025)
Randomness, exchangeability, and conformal prediction
von: Vovk, Vladimir
Veröffentlicht: (2025)
von: Vovk, Vladimir
Veröffentlicht: (2025)
Validity and efficiency of the conformal CUSUM procedure
von: Vovk, Vladimir, et al.
Veröffentlicht: (2024)
von: Vovk, Vladimir, et al.
Veröffentlicht: (2024)
Accelerating System-Level Debug Using Rule Learning and Subgroup Discovery Techniques
von: Khasidashvili, Zurab
Veröffentlicht: (2022)
von: Khasidashvili, Zurab
Veröffentlicht: (2022)
Generative Adversarial Learning from Deterministic Processes
von: Kühl, Joris C., et al.
Veröffentlicht: (2026)
von: Kühl, Joris C., et al.
Veröffentlicht: (2026)
Offline Dynamic Inventory and Pricing Strategy: Addressing Censored and Dependent Demand
von: Gundem, Korel, et al.
Veröffentlicht: (2025)
von: Gundem, Korel, et al.
Veröffentlicht: (2025)
LIBRA: Language Model Informed Bandit Recourse Algorithm for Personalized Treatment Planning
von: Cao, Junyu, et al.
Veröffentlicht: (2026)
von: Cao, Junyu, et al.
Veröffentlicht: (2026)
A packing lemma for VCN${}_k$-dimension and learning high-dimensional data
von: Coregliano, Leonardo N., et al.
Veröffentlicht: (2025)
von: Coregliano, Leonardo N., et al.
Veröffentlicht: (2025)
Towards regularized learning from functional data with covariate shift
von: Holzleitner, Markus, et al.
Veröffentlicht: (2026)
von: Holzleitner, Markus, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Kernel $ε$-Greedy for Multi-Armed Bandits with Covariates
von: Arya, Sakshi, et al.
Veröffentlicht: (2023) -
Asymptotic Optimism of Random-Design Linear and Kernel Regression Models
von: Luo, Hengrui, et al.
Veröffentlicht: (2025) -
Sparse joint shift in multinomial classification
von: Tasche, Dirk
Veröffentlicht: (2023) -
A Frequency-Domain Analysis of the Multi-Armed Bandit Problem: A New Perspective on the Exploration-Exploitation Trade-off
von: Zhang, Di
Veröffentlicht: (2025) -
Optimal In-context Adaptivity and Distributional Robustness of Transformers
von: Ma, Tianyi, et al.
Veröffentlicht: (2025)