Dynamic resource matching in manufacturing using deep reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Panda, Saunak Kumar, Xiang, Yisha, Liu, Ruiqi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization
by: Xu, Ruoran, et al.
Published: (2026)
by: Xu, Ruoran, et al.
Published: (2026)
PyEPO: A PyTorch-based End-to-End Predict-then-Optimize Library for Linear and Integer Programming
by: Tang, Bo, et al.
Published: (2022)
by: Tang, Bo, et al.
Published: (2022)
Enhancing Model Based Derivative Free Optimization using Direct Search
by: Li, Zijun, et al.
Published: (2026)
by: Li, Zijun, et al.
Published: (2026)
Multi-level meta-reinforcement learning with skill-based curriculum
by: Yang, Sichen, et al.
Published: (2026)
by: Yang, Sichen, et al.
Published: (2026)
Online Statistical Inference of Constant Sample-averaged Q-Learning
by: Panda, Saunak Kumar, et al.
Published: (2026)
by: Panda, Saunak Kumar, et al.
Published: (2026)
Interior-Point Vanishing Problem in Semidefinite Relaxations for Neural Network Verification
by: Ueda, Ryota, et al.
Published: (2025)
by: Ueda, Ryota, et al.
Published: (2025)
Amazon Locker Capacity Management
by: Sethuraman, Samyukta, et al.
Published: (2023)
by: Sethuraman, Samyukta, et al.
Published: (2023)
Escaping Saddle Points via Curvature-Calibrated Perturbations: A Complete Analysis with Explicit Constants and Empirical Validation
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
UniPROT: Uniform Prototype Selection via Partial Optimal Transport with Submodular Guarantees
by: Chanda, Prateek, et al.
Published: (2026)
by: Chanda, Prateek, et al.
Published: (2026)
Implicit Bias and Invariance: How Hopfield Networks Efficiently Learn Graph Orbits
by: Murray, Michael, et al.
Published: (2025)
by: Murray, Michael, et al.
Published: (2025)
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward
by: Fang, Jiawei, et al.
Published: (2025)
by: Fang, Jiawei, et al.
Published: (2025)
NOVAK: Unified adaptive optimizer for deep neural networks
by: Kavun, Sergii
Published: (2026)
by: Kavun, Sergii
Published: (2026)
Efficient Training of Physics-enhanced Neural ODEs via Direct Collocation and Nonlinear Programming
by: Langenkamp, Linus, et al.
Published: (2025)
by: Langenkamp, Linus, et al.
Published: (2025)
Vanishing L2 regularization for the softmax Multi Armed Bandit
by: Anita, Stefana-Lucia, et al.
Published: (2026)
by: Anita, Stefana-Lucia, et al.
Published: (2026)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
by: Khanal, Kshitiz
Published: (2025)
by: Khanal, Kshitiz
Published: (2025)
Robustness, Cost, and Attack-Surface Concentration in Phishing Detection
by: Allagan, Julian, et al.
Published: (2026)
by: Allagan, Julian, et al.
Published: (2026)
SALE-Based Offline Reinforcement Learning with Ensemble Q-Networks
by: Chun, Zheng
Published: (2025)
by: Chun, Zheng
Published: (2025)
Optimizing Donor Outreach for Blood Collection Sessions: A Scalable Decision Support Framework
by: Carneiro, André, et al.
Published: (2026)
by: Carneiro, André, et al.
Published: (2026)
CAO: Curvature-Adaptive Optimization via Periodic Low-Rank Hessian Sketching
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Machine Learning Algorithms for Improving Black Box Optimization Solvers
by: Kimiaei, Morteza, et al.
Published: (2025)
by: Kimiaei, Morteza, et al.
Published: (2025)
TOPSIS-like metaheuristic for LABS problem
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
Iterative Exploration-Driven Sparse SDP Clustering via Thompson Sampling
by: Mun, Jongmin, et al.
Published: (2025)
by: Mun, Jongmin, et al.
Published: (2025)
Autonomous AI Agents for Real-Time Affordable Housing Site Selection: Multi-Objective Reinforcement Learning Under Regulatory Constraints
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
From Non-Identifiability to Goal-Integrated Decision-Making in Parametric Inverse Optimization
by: Ahmadi, Farzin, et al.
Published: (2026)
by: Ahmadi, Farzin, et al.
Published: (2026)
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
LOFA: Online Influence Maximization under Full-Bandit Feedback using Lazy Forward Selection
by: Xu, Jinyu, et al.
Published: (2026)
by: Xu, Jinyu, et al.
Published: (2026)
Active perception and disentangled representations allow continual, episodic zero and few-shot learning
by: Rawlinson, David, et al.
Published: (2026)
by: Rawlinson, David, et al.
Published: (2026)
Pseudoconvex Problems in Operational Decision Systems: Algorithms for Joint Learning and Optimization
by: Li, Zijun, et al.
Published: (2026)
by: Li, Zijun, et al.
Published: (2026)
Tomographic projection optimization for volumetric additive manufacturing with general band constraint Lp-norm minimization
by: Li, Chi Chung, et al.
Published: (2023)
by: Li, Chi Chung, et al.
Published: (2023)
Adaptive Conditional Forest Sampling for Spectral Risk Optimisation under Decision-Dependent Uncertainty
by: Kurbucz, Marcell T.
Published: (2026)
by: Kurbucz, Marcell T.
Published: (2026)
Multi-Agent Decision-Focused Learning via Value-Aware Sequential Communication
by: Amoh, Benjamin, et al.
Published: (2026)
by: Amoh, Benjamin, et al.
Published: (2026)
Stochastic convergence of parallel asynchronous adaptive first-order methods
by: Gratton, Serge, et al.
Published: (2026)
by: Gratton, Serge, et al.
Published: (2026)
Generative AI in Transportation Planning: A Survey
by: Da, Longchao, et al.
Published: (2025)
by: Da, Longchao, et al.
Published: (2025)
Dual-Channel Feature Fusion for Joint Prediction in Dynamic Signed Weighted Networks
by: Zhang, Gaoxin, et al.
Published: (2026)
by: Zhang, Gaoxin, et al.
Published: (2026)
FedWCM: Unleashing the Potential of Momentum-based Federated Learning in Long-Tailed Scenarios
by: Li, Tianle, et al.
Published: (2025)
by: Li, Tianle, et al.
Published: (2025)
Robustness of Reinforcement Learning-Based Traffic Signal Control under Incidents: A Comparative Study
by: Nguyen, Dang Viet Anh, et al.
Published: (2025)
by: Nguyen, Dang Viet Anh, et al.
Published: (2025)
Explainable Artificial Intelligence Credit Risk Assessment using Machine Learning
by: Shreya, et al.
Published: (2025)
by: Shreya, et al.
Published: (2025)
Modeling Vehicle-Type-Specific Pedestrian Crash Avoidance Behavior in Safety-Critical Interactions Using Smooth-Mamba Deep Reinforcement Learning
by: Pu, Qingwen, et al.
Published: (2026)
by: Pu, Qingwen, et al.
Published: (2026)
Multi-beam phase mask optimization for holographic volumetric additive manufacturing
by: Li, Chi Chung, et al.
Published: (2024)
by: Li, Chi Chung, et al.
Published: (2024)
Multi-Objective Optimization with Desirability and Morris-Mitchell Criterion
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
Similar Items
-
Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization
by: Xu, Ruoran, et al.
Published: (2026) -
PyEPO: A PyTorch-based End-to-End Predict-then-Optimize Library for Linear and Integer Programming
by: Tang, Bo, et al.
Published: (2022) -
Enhancing Model Based Derivative Free Optimization using Direct Search
by: Li, Zijun, et al.
Published: (2026) -
Multi-level meta-reinforcement learning with skill-based curriculum
by: Yang, Sichen, et al.
Published: (2026) -
Online Statistical Inference of Constant Sample-averaged Q-Learning
by: Panda, Saunak Kumar, et al.
Published: (2026)