Smooth Non-Stationary Bandits
Fuente:
arXiv
Guardado en:
| Autores principales: | Jia, Su, Xie, Qian, Kallus, Nathan, Frazier, Peter I. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
por: Mohamed, Mimoun, et al.
Publicado: (2023)
por: Mohamed, Mimoun, et al.
Publicado: (2023)
Piecewise Polynomial Regression of Tame Functions via Integer Programming
por: Bareilles, Gilles, et al.
Publicado: (2023)
por: Bareilles, Gilles, et al.
Publicado: (2023)
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
por: Nguyen, Minh
Publicado: (2026)
por: Nguyen, Minh
Publicado: (2026)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
por: Bareilles, Gilles, et al.
Publicado: (2026)
por: Bareilles, Gilles, et al.
Publicado: (2026)
Sinkhorn Based Associative Memory Retrieval Using Spherical Hellinger Kantorovich Dynamics
por: Mustafi, Aratrika, et al.
Publicado: (2026)
por: Mustafi, Aratrika, et al.
Publicado: (2026)
Inverse Mixed-Integer Programming: Learning Constraints then Objective Functions
por: Kitaoka, Akira
Publicado: (2025)
por: Kitaoka, Akira
Publicado: (2025)
A Differential and Pointwise Control Approach to Reinforcement Learning
por: Nguyen, Minh, et al.
Publicado: (2024)
por: Nguyen, Minh, et al.
Publicado: (2024)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
por: Han, Qiyang, et al.
Publicado: (2025)
por: Han, Qiyang, et al.
Publicado: (2025)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
por: Chen, Siyu, et al.
Publicado: (2024)
por: Chen, Siyu, et al.
Publicado: (2024)
Learning to Fuse Temporal Proximity Networks: A Case Study in Chimpanzee Social Interactions
por: He, Yixuan, et al.
Publicado: (2025)
por: He, Yixuan, et al.
Publicado: (2025)
Sail into the Headwind: Alignment via Robust Rewards and Dynamic Labels against Reward Hacking
por: Rashidinejad, Paria, et al.
Publicado: (2024)
por: Rashidinejad, Paria, et al.
Publicado: (2024)
FraPPE: Fast and Efficient Preference-based Pure Exploration
por: Das, Udvas, et al.
Publicado: (2025)
por: Das, Udvas, et al.
Publicado: (2025)
Statistical and Algorithmic Foundations of Reinforcement Learning
por: Chi, Yuejie, et al.
Publicado: (2025)
por: Chi, Yuejie, et al.
Publicado: (2025)
Optimism Stabilizes Thompson Sampling for Adaptive Inference
por: Yan, Shunxing, et al.
Publicado: (2026)
por: Yan, Shunxing, et al.
Publicado: (2026)
Federated Optimization of Smooth Loss Functions
por: Jadbabaie, Ali, et al.
Publicado: (2022)
por: Jadbabaie, Ali, et al.
Publicado: (2022)
An Elementary Proof of the Near Optimality of LogSumExp Smoothing
por: Samakhoana, Thabo, et al.
Publicado: (2025)
por: Samakhoana, Thabo, et al.
Publicado: (2025)
Early Stopping in Contextual Bandits and Inferences
por: Cui, Zihan
Publicado: (2025)
por: Cui, Zihan
Publicado: (2025)
Adaptive Smooth Non-Stationary Bandits
por: Suk, Joe
Publicado: (2024)
por: Suk, Joe
Publicado: (2024)
Joint Learning of Linear Dynamical Systems under Smoothness Constraints
por: Tyagi, Hemant
Publicado: (2024)
por: Tyagi, Hemant
Publicado: (2024)
Geometry-induced Regularization in Deep ReLU Neural Networks
por: Bona-Pellissier, Joachim, et al.
Publicado: (2024)
por: Bona-Pellissier, Joachim, et al.
Publicado: (2024)
Probabilistic Guarantees of Stochastic Recursive Gradient in Non-Convex Finite Sum Problems
por: Zhong, Yanjie, et al.
Publicado: (2024)
por: Zhong, Yanjie, et al.
Publicado: (2024)
An Improved Analysis of Langevin Algorithms with Prior Diffusion for Non-Log-Concave Sampling
por: Huang, Xunpeng, et al.
Publicado: (2024)
por: Huang, Xunpeng, et al.
Publicado: (2024)
Efficient Group Lasso Regularized Rank Regression with Data-Driven Parameter Determination
por: Lin, Meixia, et al.
Publicado: (2025)
por: Lin, Meixia, et al.
Publicado: (2025)
Data-Efficient Non-Gaussian Semi-Nonparametric Density Estimation for Nonlinear Dynamical Systems
por: Liao, Aaron R., et al.
Publicado: (2026)
por: Liao, Aaron R., et al.
Publicado: (2026)
Learning the Uncertainty Sets for Control Dynamics via Set Membership: A Non-Asymptotic Analysis
por: Li, Yingying, et al.
Publicado: (2023)
por: Li, Yingying, et al.
Publicado: (2023)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
por: Cheng, Xiuyuan, et al.
Publicado: (2023)
por: Cheng, Xiuyuan, et al.
Publicado: (2023)
The Collusion of Memory and Nonlinearity in Stochastic Approximation With Constant Stepsize
por: Huo, Dongyan, et al.
Publicado: (2024)
por: Huo, Dongyan, et al.
Publicado: (2024)
Non-Smooth Weakly-Convex Finite-sum Coupled Compositional Optimization
por: Hu, Quanqi, et al.
Publicado: (2023)
por: Hu, Quanqi, et al.
Publicado: (2023)
Stopping Rules for Stochastic Gradient Descent via Anytime-Valid Confidence Sequences
por: Aolaritei, Liviu, et al.
Publicado: (2025)
por: Aolaritei, Liviu, et al.
Publicado: (2025)
A review of NMF, PLSA, LBA, EMA, and LCA with a focus on the identifiability issue
por: Qi, Qianqian, et al.
Publicado: (2025)
por: Qi, Qianqian, et al.
Publicado: (2025)
Gradient Equilibrium in Online Learning: Theory and Applications
por: Angelopoulos, Anastasios N., et al.
Publicado: (2025)
por: Angelopoulos, Anastasios N., et al.
Publicado: (2025)
Diagonalisation SGD: Fast & Convergent SGD for Non-Differentiable Models via Reparameterisation and Smoothing
por: Wagner, Dominik, et al.
Publicado: (2024)
por: Wagner, Dominik, et al.
Publicado: (2024)
Stochastic Optimization with Optimal Importance Sampling
por: Aolaritei, Liviu, et al.
Publicado: (2025)
por: Aolaritei, Liviu, et al.
Publicado: (2025)
The Vizier Gaussian Process Bandit Algorithm
por: Song, Xingyou, et al.
Publicado: (2024)
por: Song, Xingyou, et al.
Publicado: (2024)
Analytic Bridge Diffusions for Controlled Path Generation
por: Chertkov, Michael
Publicado: (2026)
por: Chertkov, Michael
Publicado: (2026)
Beyond Maximum Likelihood: Variational Inequality Estimation for Generalized Linear Models
por: Zhu, Linglingzhi, et al.
Publicado: (2025)
por: Zhu, Linglingzhi, et al.
Publicado: (2025)
A Graphical Global Optimization Framework for Parameter Estimation of Statistical Models with Nonconvex Regularization Functions
por: Davarnia, Danial, et al.
Publicado: (2025)
por: Davarnia, Danial, et al.
Publicado: (2025)
A Piecewise Lyapunov Analysis of Sub-quadratic SGD: Applications to Robust and Quantile Regression
por: Zhang, Yixuan, et al.
Publicado: (2025)
por: Zhang, Yixuan, et al.
Publicado: (2025)
Power Constrained Nonstationary Bandits with Habituation and Recovery Dynamics
por: Li, Fengxu, et al.
Publicado: (2025)
por: Li, Fengxu, et al.
Publicado: (2025)
Causal Invariance Learning via Efficient Nonconvex Optimization
por: Wang, Zhenyu, et al.
Publicado: (2024)
por: Wang, Zhenyu, et al.
Publicado: (2024)
Ejemplares similares
-
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
por: Mohamed, Mimoun, et al.
Publicado: (2023) -
Piecewise Polynomial Regression of Tame Functions via Integer Programming
por: Bareilles, Gilles, et al.
Publicado: (2023) -
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
por: Nguyen, Minh
Publicado: (2026) -
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
por: Bareilles, Gilles, et al.
Publicado: (2026) -
Sinkhorn Based Associative Memory Retrieval Using Spherical Hellinger Kantorovich Dynamics
por: Mustafi, Aratrika, et al.
Publicado: (2026)