Deep Policy Iteration with Integer Programming for Inventory Management
Fuente:
arXiv
Saved in:
| Main Authors: | Harsha, Pavithra, Jagmohan, Ashish, Kalagnanam, Jayant, Quanz, Brian, Singhvi, Divya |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Industrial Injection Molding Processes with Explainable AI for Quality Classification
by: Rottenwalter, Georg, et al.
Published: (2025)
by: Rottenwalter, Georg, et al.
Published: (2025)
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
by: Priezzhev, I. I., et al.
Published: (2025)
by: Priezzhev, I. I., et al.
Published: (2025)
Advancements in synthetic data extraction for industrial injection molding
by: Rottenwalter, Georg, et al.
Published: (2025)
by: Rottenwalter, Georg, et al.
Published: (2025)
Predicting Future Actions of Reinforcement Learning Agents
by: Chung, Stephen, et al.
Published: (2024)
by: Chung, Stephen, et al.
Published: (2024)
N-Agent Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2024)
by: Wang, Caroline, et al.
Published: (2024)
Inter-Series Transformer: Attending to Products in Time Series Forecasting
by: Cristian, Rares, et al.
Published: (2024)
by: Cristian, Rares, et al.
Published: (2024)
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance
by: Silue, Bram, et al.
Published: (2025)
by: Silue, Bram, et al.
Published: (2025)
Spiking Neural Network Architecture Search: A Survey
by: Svoboda, Kama, et al.
Published: (2025)
by: Svoboda, Kama, et al.
Published: (2025)
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2025)
by: Wang, Caroline, et al.
Published: (2025)
Cryptogenic stroke and migraine: using probabilistic independence and machine learning to uncover latent sources of disease from the electronic health record
by: Betts, Joshua W., et al.
Published: (2025)
by: Betts, Joshua W., et al.
Published: (2025)
Deep Convolutional Autoencoder for Assessment of Drive-Cycle Anomalies in Connected Vehicle Sensor Data
by: Geglio, Anthony, et al.
Published: (2022)
by: Geglio, Anthony, et al.
Published: (2022)
Soft Actor-Critic with Beta Policy via Implicit Reparameterization Gradients
by: Della Libera, Luca
Published: (2024)
by: Della Libera, Luca
Published: (2024)
Iterative Feature Boosting for Explainable Speech Emotion Recognition
by: Nfissi, Alaa, et al.
Published: (2024)
by: Nfissi, Alaa, et al.
Published: (2024)
FDQN: A Flexible Deep Q-Network Framework for Game Automation
by: Gujavarthy, Prabhath Reddy
Published: (2024)
by: Gujavarthy, Prabhath Reddy
Published: (2024)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
by: Zolduoarrati, Elijah, et al.
Published: (2025)
by: Zolduoarrati, Elijah, et al.
Published: (2025)
Multi-Agent Pathfinding with Non-Unit Integer Edge Costs via Enhanced Conflict-Based Search and Graph Discretization
by: Fan, Hongkai, et al.
Published: (2026)
by: Fan, Hongkai, et al.
Published: (2026)
Thinker: Learning to Think Fast and Slow
by: Chung, Stephen, et al.
Published: (2025)
by: Chung, Stephen, et al.
Published: (2025)
CellARC: Measuring Intelligence with Cellular Automata
by: Lžičař, Miroslav
Published: (2025)
by: Lžičař, Miroslav
Published: (2025)
Inductive transfer learning from regression to classification in ECG analysis
by: Jayasundara, Ridma, et al.
Published: (2025)
by: Jayasundara, Ridma, et al.
Published: (2025)
Evading Overlapping Community Detection via Proxy Node Injection
by: Loi, Dario, et al.
Published: (2025)
by: Loi, Dario, et al.
Published: (2025)
Centrally Coordinated Multi-Agent Reinforcement Learning for Power Grid Topology Control
by: de Mol, Barbera, et al.
Published: (2025)
by: de Mol, Barbera, et al.
Published: (2025)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
by: Hong, Yoosung
Published: (2026)
by: Hong, Yoosung
Published: (2026)
Scalable Nested Optimization for Deep Learning
by: Lorraine, Jonathan
Published: (2024)
by: Lorraine, Jonathan
Published: (2024)
Exploring the Technology Landscape through Topic Modeling, Expert Involvement, and Reinforcement Learning
by: Nazari, Ali, et al.
Published: (2025)
by: Nazari, Ali, et al.
Published: (2025)
TML-Bench: Benchmark for Data Science Agents on Tabular ML Tasks
by: Pinchuk, Mykola
Published: (2026)
by: Pinchuk, Mykola
Published: (2026)
LakeMLB: Data Lake Machine Learning Benchmark
by: Pan, Feiyu, et al.
Published: (2026)
by: Pan, Feiyu, et al.
Published: (2026)
LEFT: Learnable Fusion of Tri-view Tokens for Unsupervised Time Series Anomaly Detection
by: Wang, Dezheng, et al.
Published: (2026)
by: Wang, Dezheng, et al.
Published: (2026)
Amortized Molecular Optimization via Group Relative Policy Optimization
by: Javaid, Muhammad bin, et al.
Published: (2026)
by: Javaid, Muhammad bin, et al.
Published: (2026)
Imitation learning for sim-to-real transfer of robotic cutting policies based on residual Gaussian process disturbance force model
by: Hathaway, Jamie, et al.
Published: (2023)
by: Hathaway, Jamie, et al.
Published: (2023)
End-to-end example-based sim-to-real RL policy transfer based on neural stylisation with application to robotic cutting
by: Hathaway, Jamie, et al.
Published: (2026)
by: Hathaway, Jamie, et al.
Published: (2026)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
by: Sun, Yuhui, et al.
Published: (2025)
by: Sun, Yuhui, et al.
Published: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
by: Reddy, Sandeep, et al.
Published: (2025)
by: Reddy, Sandeep, et al.
Published: (2025)
Open-TI: Open Traffic Intelligence with Augmented Language Model
by: Da, Longchao, et al.
Published: (2023)
by: Da, Longchao, et al.
Published: (2023)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
by: Zeytuncu, Yunus E.
Published: (2026)
by: Zeytuncu, Yunus E.
Published: (2026)
GraphWalk: Enabling Reasoning in Large Language Models through Tool-Based Graph Navigation
by: Ghandi, Taraneh, et al.
Published: (2026)
by: Ghandi, Taraneh, et al.
Published: (2026)
Quantum Abduction: A New Paradigm for Reasoning under Uncertainty
by: Pareschi, Remo
Published: (2025)
by: Pareschi, Remo
Published: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
On Divergence Measures for Training GFlowNets
by: da Silva, Tiago, et al.
Published: (2024)
by: da Silva, Tiago, et al.
Published: (2024)
Similar Items
-
Improving Industrial Injection Molding Processes with Explainable AI for Quality Classification
by: Rottenwalter, Georg, et al.
Published: (2025) -
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
by: Priezzhev, I. I., et al.
Published: (2025) -
Advancements in synthetic data extraction for industrial injection molding
by: Rottenwalter, Georg, et al.
Published: (2025) -
Predicting Future Actions of Reinforcement Learning Agents
by: Chung, Stephen, et al.
Published: (2024) -
N-Agent Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2024)