Multi-Armed Sampling Problem and the End of Exploration
Fuente:
arXiv
Saved in:
| Main Authors: | Pedramfar, Mohammad, Ravanbakhsh, Siamak |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Unified Framework for Analyzing Meta-algorithms in Online Convex Optimization
by: Pedramfar, Mohammad, et al.
Published: (2024)
by: Pedramfar, Mohammad, et al.
Published: (2024)
$γ$-weakly $θ$-up-concavity: A Unified Framework for Non-Convex Optimization Beyond DR-Submodular and OSS Functions
by: Pedramfar, Mohammad, et al.
Published: (2026)
by: Pedramfar, Mohammad, et al.
Published: (2026)
From Linear to Linearizable Optimization: A Novel Framework with Applications to Stationary and Non-stationary DR-submodular Optimization
by: Pedramfar, Mohammad, et al.
Published: (2024)
by: Pedramfar, Mohammad, et al.
Published: (2024)
BAGEL: Projection-Free Algorithm for Adversarially Constrained Online Convex Optimization
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
Decentralized Projection-free Online Upper-Linearizable Optimization with Applications to DR-Submodular Optimization
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
Upper-Linearizability of Online Non-Monotone DR-Submodular Maximization over Down-Closed Convex Sets
by: Lu, Yiyang, et al.
Published: (2026)
by: Lu, Yiyang, et al.
Published: (2026)
Diffusion Tree Sampling: Scalable inference-time alignment of diffusion models
by: Jain, Vineet, et al.
Published: (2025)
by: Jain, Vineet, et al.
Published: (2025)
Multi-Armed Bandits with Minimum Aggregated Revenue Constraints
by: Yahmed, Ahmed Ben, et al.
Published: (2025)
by: Yahmed, Ahmed Ben, et al.
Published: (2025)
Balans: Multi-Armed Bandits-based Adaptive Large Neighborhood Search for Mixed-Integer Programming Problem
by: Cai, Junyang, et al.
Published: (2024)
by: Cai, Junyang, et al.
Published: (2024)
Unified Projection-Free Algorithms for Adversarial DR-Submodular Optimization
by: Pedramfar, Mohammad, et al.
Published: (2024)
by: Pedramfar, Mohammad, et al.
Published: (2024)
SPO-VCS: An End-to-End Smart Predict-then-Optimize Framework with Alternating Differentiation Method for Relocation Problems in Large-Scale Vehicle Crowd Sensing
by: Wang, Xinyu, et al.
Published: (2024)
by: Wang, Xinyu, et al.
Published: (2024)
Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
by: Boveiri, Mohammad, et al.
Published: (2024)
by: Boveiri, Mohammad, et al.
Published: (2024)
Sobolev Training of End-to-End Optimization Proxies
by: Rosemberg, Andrew W., et al.
Published: (2025)
by: Rosemberg, Andrew W., et al.
Published: (2025)
Simultaneous Learning and Optimization via Misspecified Saddle Point Problems
by: Ahmadi, Mohammad Mahdi, et al.
Published: (2025)
by: Ahmadi, Mohammad Mahdi, et al.
Published: (2025)
Closing the Gaps: Optimality of Sample Average Approximation for Data-Driven Newsvendor Problems
by: Lyu, Jiameng, et al.
Published: (2024)
by: Lyu, Jiameng, et al.
Published: (2024)
End-to-End Conformal Calibration for Optimization Under Uncertainty
by: Yeh, Christopher, et al.
Published: (2024)
by: Yeh, Christopher, et al.
Published: (2024)
Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL
by: Yang, Tong, et al.
Published: (2025)
by: Yang, Tong, et al.
Published: (2025)
Constrained Sampling with Primal-Dual Langevin Monte Carlo
by: Chamon, Luiz F. O., et al.
Published: (2024)
by: Chamon, Luiz F. O., et al.
Published: (2024)
Conditional Sampling via Wasserstein Autoencoders and Triangular Transport
by: Al-Jarrah, Mohammad, et al.
Published: (2026)
by: Al-Jarrah, Mohammad, et al.
Published: (2026)
More Optimal Fractional-Order Stochastic Gradient Descent for Non-Convex Optimization Problems
by: Partohaghighi, Mohammad, et al.
Published: (2025)
by: Partohaghighi, Mohammad, et al.
Published: (2025)
Effective Dimension Aware Fractional-Order Stochastic Gradient Descent for Convex Optimization Problems
by: Partohaghighi, Mohammad, et al.
Published: (2025)
by: Partohaghighi, Mohammad, et al.
Published: (2025)
Mirror Descent-Type Algorithms for the Variational Inequality Problem with Functional Constraints
by: Alkousa, Mohammad S., et al.
Published: (2026)
by: Alkousa, Mohammad S., et al.
Published: (2026)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
by: Zhang, Xiaole, et al.
Published: (2025)
by: Zhang, Xiaole, et al.
Published: (2025)
End-to-End Reinforcement Learning of Koopman Models for eNMPC of an Air Separation Unit
by: Mayfrank, Daniel, et al.
Published: (2025)
by: Mayfrank, Daniel, et al.
Published: (2025)
End-to-End Training of High-Dimensional Optimal Control with Implicit Hamiltonians via Jacobian-Free Backpropagation
by: Gelphman, Eric, et al.
Published: (2025)
by: Gelphman, Eric, et al.
Published: (2025)
Collaborative Pareto Set Learning in Multiple Multi-Objective Optimization Problems
by: Shang, Chikai, et al.
Published: (2024)
by: Shang, Chikai, et al.
Published: (2024)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
by: Muehlebach, Michael, et al.
Published: (2025)
by: Muehlebach, Michael, et al.
Published: (2025)
Provably Efficient Exploration in Policy Optimization
by: Cai, Qi, et al.
Published: (2019)
by: Cai, Qi, et al.
Published: (2019)
Scalable Online Exploration via Coverability
by: Amortila, Philip, et al.
Published: (2024)
by: Amortila, Philip, et al.
Published: (2024)
Scalable Solution of the Stochastic Multi-path Traveling Salesman Problem via Neural Networks
by: Chou, Xiaochen, et al.
Published: (2026)
by: Chou, Xiaochen, et al.
Published: (2026)
Bayesian Optimization by Kernel Regression and Density-based Exploration
by: Zhu, Tansheng, et al.
Published: (2025)
by: Zhu, Tansheng, et al.
Published: (2025)
Efficient Model-Free Exploration in Low-Rank MDPs
by: Mhammedi, Zakaria, et al.
Published: (2023)
by: Mhammedi, Zakaria, et al.
Published: (2023)
DOGE-Train: Discrete Optimization on GPU with End-to-end Training
by: Abbas, Ahmed, et al.
Published: (2022)
by: Abbas, Ahmed, et al.
Published: (2022)
Solving the Paint Shop Problem with Flexible Management of Multi-Lane Buffers Using Reinforcement Learning and Action Masking
by: Stappert, Mirko, et al.
Published: (2025)
by: Stappert, Mirko, et al.
Published: (2025)
A Rolling-Space Branch-and-Price Algorithm for the Multi-Compartment Vehicle Routing Problem with Multiple Time Windows
by: Raqabi, El Mehdi Er, et al.
Published: (2026)
by: Raqabi, El Mehdi Er, et al.
Published: (2026)
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
by: Huang, Yilie, et al.
Published: (2025)
by: Huang, Yilie, et al.
Published: (2025)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
by: Zhang, Junkai, et al.
Published: (2023)
by: Zhang, Junkai, et al.
Published: (2023)
Calibrated Adversarial Sampling: Multi-Armed Bandit-Guided Generalization Against Unforeseen Attacks
by: Wang, Rui, et al.
Published: (2025)
by: Wang, Rui, et al.
Published: (2025)
Bayesian Optimization with Preference Exploration using a Monotonic Neural Network Ensemble
by: Wang, Hanyang, et al.
Published: (2025)
by: Wang, Hanyang, et al.
Published: (2025)
The Role of Symmetry in Optimizing Overparameterized Networks
by: Sareen, Kusha, et al.
Published: (2026)
by: Sareen, Kusha, et al.
Published: (2026)
Similar Items
-
A Unified Framework for Analyzing Meta-algorithms in Online Convex Optimization
by: Pedramfar, Mohammad, et al.
Published: (2024) -
$γ$-weakly $θ$-up-concavity: A Unified Framework for Non-Convex Optimization Beyond DR-Submodular and OSS Functions
by: Pedramfar, Mohammad, et al.
Published: (2026) -
From Linear to Linearizable Optimization: A Novel Framework with Applications to Stationary and Non-stationary DR-submodular Optimization
by: Pedramfar, Mohammad, et al.
Published: (2024) -
BAGEL: Projection-Free Algorithm for Adversarially Constrained Online Convex Optimization
by: Lu, Yiyang, et al.
Published: (2025) -
Decentralized Projection-free Online Upper-Linearizable Optimization with Applications to DR-Submodular Optimization
by: Lu, Yiyang, et al.
Published: (2025)