Learning Branching Policies for MILPs with Proximal Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Mhamed, Abdelouahed Ben, Kamal-Idrissi, Assia, Seghrouchni, Amal El Fallah |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
by: Anand, Akhil S, et al.
Published: (2025)
by: Anand, Akhil S, et al.
Published: (2025)
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Learning Concave Bid Shading Strategies in Online Auctions via Measure-valued Proximal Optimization
by: Nodozi, Iman, et al.
Published: (2025)
by: Nodozi, Iman, et al.
Published: (2025)
Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients
by: Alvo, Matias, et al.
Published: (2026)
by: Alvo, Matias, et al.
Published: (2026)
Delightful Policy Gradient
by: Osband, Ian
Published: (2026)
by: Osband, Ian
Published: (2026)
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
by: Cheng, Ziheng, et al.
Published: (2025)
by: Cheng, Ziheng, et al.
Published: (2025)
Delightful Distributed Policy Gradient
by: Osband, Ian
Published: (2026)
by: Osband, Ian
Published: (2026)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
by: Xiao, Minheng, et al.
Published: (2024)
by: Xiao, Minheng, et al.
Published: (2024)
R2L: Reliable Reinforcement Learning: Guaranteed Return & Reliable Policies in Reinforcement Learning
by: Farhi, Nadir
Published: (2025)
by: Farhi, Nadir
Published: (2025)
Kinematic Tokenization: Optimization-Based Continuous-Time Tokens for Learnable Decision Policies in Noisy Time Series
by: Kearney, Griffin
Published: (2026)
by: Kearney, Griffin
Published: (2026)
Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2026)
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2026)
Classical and Deep Reinforcement Learning Inventory Control Policies for Pharmaceutical Supply Chains with Perishability and Non-Stationarity
by: Stranieri, Francesco, et al.
Published: (2025)
by: Stranieri, Francesco, et al.
Published: (2025)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
by: Ibrahim, Sinan, et al.
Published: (2026)
by: Ibrahim, Sinan, et al.
Published: (2026)
Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
Learning Efficient and Fair Policies for Uncertainty-Aware Collaborative Human-Robot Order Picking
by: Smit, Igor G., et al.
Published: (2024)
by: Smit, Igor G., et al.
Published: (2024)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
CAMBranch: Contrastive Learning with Augmented MILPs for Branching
by: Lin, Jiacheng, et al.
Published: (2024)
by: Lin, Jiacheng, et al.
Published: (2024)
Examining Policy Entropy of Reinforcement Learning Agents for Personalization Tasks
by: Dereventsov, Anton, et al.
Published: (2022)
by: Dereventsov, Anton, et al.
Published: (2022)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Constraint-Anchored Attribution: Feasibility-Certified Counterfactuals and Bonferroni-PAC Sufficient Subsets for Neural CO Policies
by: Lafifi, Sohaib
Published: (2026)
by: Lafifi, Sohaib
Published: (2026)
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
by: Malaniya, Georgiy, et al.
Published: (2025)
by: Malaniya, Georgiy, et al.
Published: (2025)
Generalization Guarantees for Learning Branch-and-Cut Policies in Integer Programming
by: Cheng, Hongyu, et al.
Published: (2025)
by: Cheng, Hongyu, et al.
Published: (2025)
Understanding Optimization in Deep Learning with Central Flows
by: Cohen, Jeremy M., et al.
Published: (2024)
by: Cohen, Jeremy M., et al.
Published: (2024)
Deep Learning for Two-Stage Robust Integer Optimization
by: Dumouchelle, Justin, et al.
Published: (2023)
by: Dumouchelle, Justin, et al.
Published: (2023)
Convex and Bilevel Optimization for Neuro-Symbolic Inference and Learning
by: Dickens, Charles, et al.
Published: (2024)
by: Dickens, Charles, et al.
Published: (2024)
Hyperparameter Optimization for Driving Strategies Based on Reinforcement Learning
by: Adde, Nihal Acharya, et al.
Published: (2024)
by: Adde, Nihal Acharya, et al.
Published: (2024)
HoP: Homeomorphic Polar Learning for Hard Constrained Optimization
by: Deng, Ke, et al.
Published: (2025)
by: Deng, Ke, et al.
Published: (2025)
Multi-Objective Optimization for Sparse Deep Multi-Task Learning
by: Hotegni, S. S., et al.
Published: (2023)
by: Hotegni, S. S., et al.
Published: (2023)
Rolling Ball Optimizer: Learning by ironing out loss landscape wrinkles
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)
Hindsight-Guided Momentum (HGM) Optimizer: An Approach to Adaptive Learning Rate
by: Sarkar, Krisanu
Published: (2025)
by: Sarkar, Krisanu
Published: (2025)
LLM4Branch: Large Language Model for Discovering Efficient Branching Policies of Integer Programs
by: Hou, Zhinan, et al.
Published: (2026)
by: Hou, Zhinan, et al.
Published: (2026)
Learning to Fuse Temporal Proximity Networks: A Case Study in Chimpanzee Social Interactions
by: He, Yixuan, et al.
Published: (2025)
by: He, Yixuan, et al.
Published: (2025)
Isotropic Curvature Model for Understanding Deep Learning Optimization: Is Gradient Orthogonalization Optimal?
by: Su, Weijie
Published: (2025)
by: Su, Weijie
Published: (2025)
Learning-Guided Rolling Horizon Optimization for Long-Horizon Flexible Job-Shop Scheduling
by: Li, Sirui, et al.
Published: (2025)
by: Li, Sirui, et al.
Published: (2025)
A Benchmark for Maximum Cut: Towards Standardization of the Evaluation of Learned Heuristics for Combinatorial Optimization
by: Nath, Ankur, et al.
Published: (2024)
by: Nath, Ankur, et al.
Published: (2024)
Is Scaling Learned Optimizers Worth It? Evaluating The Value of VeLO's 4000 TPU Months
by: Rezk, Fady, et al.
Published: (2023)
by: Rezk, Fady, et al.
Published: (2023)
Unsupervised Machine Learning Hybrid Approach Integrating Linear Programming in Loss Function: A Robust Optimization Technique
by: Kiruluta, Andrew, et al.
Published: (2024)
by: Kiruluta, Andrew, et al.
Published: (2024)
Optimizing the Optimizer for Physics-Informed Neural Networks and Kolmogorov-Arnold Networks
by: Kiyani, Elham, et al.
Published: (2025)
by: Kiyani, Elham, et al.
Published: (2025)
Differentiable Optimization for Deep Learning-Enhanced DC Approximation of AC Optimal Power Flow
by: Rosemberg, Andrew, et al.
Published: (2025)
by: Rosemberg, Andrew, et al.
Published: (2025)
Similar Items
-
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
by: Anand, Akhil S, et al.
Published: (2025) -
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025) -
Learning Concave Bid Shading Strategies in Online Auctions via Measure-valued Proximal Optimization
by: Nodozi, Iman, et al.
Published: (2025) -
Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients
by: Alvo, Matias, et al.
Published: (2026) -
Delightful Policy Gradient
by: Osband, Ian
Published: (2026)