TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Sorokin, D., Kostin, A., Savchenko, L., Gusev, G., Savchenko, A. V. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Latent Guided Sampling for Combinatorial Optimization
by: Surendran, Sobihan, et al.
Published: (2025)
by: Surendran, Sobihan, et al.
Published: (2025)
WeightLoRA: Keep Only Necessary Adapters
by: Veprikov, Andrey, et al.
Published: (2025)
by: Veprikov, Andrey, et al.
Published: (2025)
Reheated Gradient-based Discrete Sampling for Combinatorial Optimization
by: Li, Muheng, et al.
Published: (2025)
by: Li, Muheng, et al.
Published: (2025)
Structured Reinforcement Learning for Combinatorial Decision-Making
by: Hoppe, Heiko, et al.
Published: (2025)
by: Hoppe, Heiko, et al.
Published: (2025)
Combinatorial Optimization Augmented Machine Learning
by: Schiffer, Maximilian, et al.
Published: (2026)
by: Schiffer, Maximilian, et al.
Published: (2026)
Heuristics for Combinatorial Optimization via Value-based Reinforcement Learning: A Unified Framework and Analysis
by: Davidovich, Orit, et al.
Published: (2025)
by: Davidovich, Orit, et al.
Published: (2025)
Sample Efficient Reinforcement Learning with Partial Dynamics Knowledge
by: Alharbi, Meshal, et al.
Published: (2023)
by: Alharbi, Meshal, et al.
Published: (2023)
Sample-Efficient Reinforcement Learning of Koopman eNMPC
by: Mayfrank, Daniel, et al.
Published: (2025)
by: Mayfrank, Daniel, et al.
Published: (2025)
Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning
by: Wan, Jia, et al.
Published: (2024)
by: Wan, Jia, et al.
Published: (2024)
Pointer Networks with Q-Learning for Combinatorial Optimization
by: Barro, Alessandro
Published: (2023)
by: Barro, Alessandro
Published: (2023)
InterQ: A DQN Framework for Optimal Intermittent Control
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Exploring New Frontiers in Vertical Federated Learning: the Role of Saddle Point Reformulation
by: Beznosikov, Aleksandr, et al.
Published: (2026)
by: Beznosikov, Aleksandr, et al.
Published: (2026)
Coreset-Based Task Selection for Sample-Efficient Meta-Reinforcement Learning
by: Zhan, Donglin, et al.
Published: (2025)
by: Zhan, Donglin, et al.
Published: (2025)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
by: Qiu, Shuang, et al.
Published: (2024)
by: Qiu, Shuang, et al.
Published: (2024)
Learning to Handle Parameter Perturbations in Combinatorial Optimization: an Application to Facility Location
by: Lodi, Andrea, et al.
Published: (2019)
by: Lodi, Andrea, et al.
Published: (2019)
Provably Efficient Exploration in Policy Optimization
by: Cai, Qi, et al.
Published: (2019)
by: Cai, Qi, et al.
Published: (2019)
Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems
by: Suttle, Wesley A., et al.
Published: (2024)
by: Suttle, Wesley A., et al.
Published: (2024)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
by: Dus, Mathias
Published: (2026)
by: Dus, Mathias
Published: (2026)
Accuracy of Discretely Sampled Stochastic Policies in Continuous-time Reinforcement Learning
by: Jia, Yanwei, et al.
Published: (2025)
by: Jia, Yanwei, et al.
Published: (2025)
Achieve Performatively Optimal Policy for Performative Reinforcement Learning
by: Chen, Ziyi, et al.
Published: (2025)
by: Chen, Ziyi, et al.
Published: (2025)
Optimizing Posterior Samples for Bayesian Optimization via Rootfinding
by: Adebiyi, Taiwo A., et al.
Published: (2024)
by: Adebiyi, Taiwo A., et al.
Published: (2024)
Enhancing GNNs Performance on Combinatorial Optimization by Recurrent Feature Update
by: Pugacheva, Daria, et al.
Published: (2024)
by: Pugacheva, Daria, et al.
Published: (2024)
Imitation Learning for Combinatorial Optimisation under Uncertainty
by: Gawas, Prakash, et al.
Published: (2026)
by: Gawas, Prakash, et al.
Published: (2026)
A Simple Mixture Policy Parameterization for Improving Sample Efficiency of CVaR Optimization
by: Luo, Yudong, et al.
Published: (2024)
by: Luo, Yudong, et al.
Published: (2024)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
by: Chen, Zijun, et al.
Published: (2025)
by: Chen, Zijun, et al.
Published: (2025)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Tackling Prevalent Conditions in Unsupervised Combinatorial Optimization: Cardinality, Minimum, Covering, and More
by: Bu, Fanchen, et al.
Published: (2024)
by: Bu, Fanchen, et al.
Published: (2024)
Single- vs. Dual-Policy Reinforcement Learning for Dynamic Bike Rebalancing
by: Liang, Jiaqi, et al.
Published: (2024)
by: Liang, Jiaqi, et al.
Published: (2024)
Convergence and Sample Complexity of First-Order Methods for Agnostic Reinforcement Learning
by: Sherman, Uri, et al.
Published: (2025)
by: Sherman, Uri, et al.
Published: (2025)
Offline Reinforcement Learning via Inverse Optimization
by: Dimanidis, Ioannis, et al.
Published: (2025)
by: Dimanidis, Ioannis, et al.
Published: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
by: Muehlebach, Michael, et al.
Published: (2025)
by: Muehlebach, Michael, et al.
Published: (2025)
Deep Reinforcement Learning: A Convex Optimization Approach
by: Gattami, Ather
Published: (2024)
by: Gattami, Ather
Published: (2024)
Methodology for Interpretable Reinforcement Learning for Optimizing Mechanical Ventilation
by: Lee, Joo Seung, et al.
Published: (2024)
by: Lee, Joo Seung, et al.
Published: (2024)
Reinforcement Learning Paycheck Optimization for Multivariate Financial Goals
by: Alaluf, Melda, et al.
Published: (2024)
by: Alaluf, Melda, et al.
Published: (2024)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
by: Carmona, René, et al.
Published: (2019)
by: Carmona, René, et al.
Published: (2019)
Does DQN Learn?
by: Gopalan, Aditya, et al.
Published: (2022)
by: Gopalan, Aditya, et al.
Published: (2022)
Offline Policy Learning with Weight Clipping and Heaviside Composite Optimization
by: Liu, Jingren, et al.
Published: (2026)
by: Liu, Jingren, et al.
Published: (2026)
An Efficient On-Policy Deep Learning Framework for Stochastic Optimal Control
by: Hua, Mengjian, et al.
Published: (2024)
by: Hua, Mengjian, et al.
Published: (2024)
Similar Items
-
Latent Guided Sampling for Combinatorial Optimization
by: Surendran, Sobihan, et al.
Published: (2025) -
WeightLoRA: Keep Only Necessary Adapters
by: Veprikov, Andrey, et al.
Published: (2025) -
Reheated Gradient-based Discrete Sampling for Combinatorial Optimization
by: Li, Muheng, et al.
Published: (2025) -
Structured Reinforcement Learning for Combinatorial Decision-Making
by: Hoppe, Heiko, et al.
Published: (2025) -
Combinatorial Optimization Augmented Machine Learning
by: Schiffer, Maximilian, et al.
Published: (2026)