Smaller Batches, Bigger Gains? Investigating the Impact of Batch Sizes on Reinforcement Learning Based Real-World Production Scheduling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Müller, Arthur, Grumbach, Felix, Sabatelli, Matthia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
von: Falzari, Massimiliano, et al.
Veröffentlicht: (2025)
von: Falzari, Massimiliano, et al.
Veröffentlicht: (2025)
On The Presence of Double-Descent in Deep Reinforcement Learning
von: Veselý, Viktor, et al.
Veröffentlicht: (2025)
von: Veselý, Viktor, et al.
Veröffentlicht: (2025)
Reinforcement Learning as an Improvement Heuristic for Real-World Production Scheduling
von: Müller, Arthur, et al.
Veröffentlicht: (2024)
von: Müller, Arthur, et al.
Veröffentlicht: (2024)
Upside-Down Reinforcement Learning for More Interpretable Optimal Control
von: Cardenas-Cartagena, Juan, et al.
Veröffentlicht: (2024)
von: Cardenas-Cartagena, Juan, et al.
Veröffentlicht: (2024)
Video-Driven Graph Network-Based Simulators
von: Szewczyk, Franciszek, et al.
Veröffentlicht: (2024)
von: Szewczyk, Franciszek, et al.
Veröffentlicht: (2024)
VDSC: Enhancing Exploration Timing with Value Discrepancy and State Counts
von: Captari, Marius, et al.
Veröffentlicht: (2024)
von: Captari, Marius, et al.
Veröffentlicht: (2024)
$ε$-Optimally Solving Zero-Sum POSGs
von: Escudie, Erwan, et al.
Veröffentlicht: (2024)
von: Escudie, Erwan, et al.
Veröffentlicht: (2024)
Is Bigger Edit Batch Size Always Better? -- An Empirical Study on Model Editing with Llama-3
von: Yoon, Junsang, et al.
Veröffentlicht: (2024)
von: Yoon, Junsang, et al.
Veröffentlicht: (2024)
Accelerating SGDM via Learning Rate and Batch Size Schedules: A Lyapunov-Based Analysis
von: Kondo, Yuichi, et al.
Veröffentlicht: (2025)
von: Kondo, Yuichi, et al.
Veröffentlicht: (2025)
Sparsity-Driven Plasticity in Multi-Task Reinforcement Learning
von: Todorov, Aleksandar, et al.
Veröffentlicht: (2025)
von: Todorov, Aleksandar, et al.
Veröffentlicht: (2025)
On the Generalisation of Koopman Representations for Chaotic System Control
von: Hjikakou, Kyriakos, et al.
Veröffentlicht: (2025)
von: Hjikakou, Kyriakos, et al.
Veröffentlicht: (2025)
Seesaw: Accelerating Training by Balancing Learning Rate and Batch Size Scheduling
von: Meterez, Alexandru, et al.
Veröffentlicht: (2025)
von: Meterez, Alexandru, et al.
Veröffentlicht: (2025)
Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler
von: Shen, Yikang, et al.
Veröffentlicht: (2024)
von: Shen, Yikang, et al.
Veröffentlicht: (2024)
AdaBatchGrad: Combining Adaptive Batch Size and Adaptive Step Size
von: Ostroukhov, Petr, et al.
Veröffentlicht: (2024)
von: Ostroukhov, Petr, et al.
Veröffentlicht: (2024)
Optimal Growth Schedules for Batch Size and Learning Rate in SGD that Reduce SFO Complexity
von: Umeda, Hikaru, et al.
Veröffentlicht: (2025)
von: Umeda, Hikaru, et al.
Veröffentlicht: (2025)
A Production Scheduling Framework for Reinforcement Learning Under Real-World Constraints
von: Hoss, Jonathan, et al.
Veröffentlicht: (2025)
von: Hoss, Jonathan, et al.
Veröffentlicht: (2025)
One Size Does Not Fit All: Architecture-Aware Adaptive Batch Scheduling with DEBA
von: Belias, François, et al.
Veröffentlicht: (2025)
von: Belias, François, et al.
Veröffentlicht: (2025)
Distilling Reinforcement Learning into Single-Batch Datasets
von: Wilhelm, Connor, et al.
Veröffentlicht: (2025)
von: Wilhelm, Connor, et al.
Veröffentlicht: (2025)
Surge Phenomenon in Optimal Learning Rate and Batch Size Scaling
von: Li, Shuaipeng, et al.
Veröffentlicht: (2024)
von: Li, Shuaipeng, et al.
Veröffentlicht: (2024)
Collaborative Batch Size Optimization for Federated Learning
von: Geimer, Arno, et al.
Veröffentlicht: (2025)
von: Geimer, Arno, et al.
Veröffentlicht: (2025)
Adaptive Batch Size Schedules for Distributed Training of Language Models with Data and Model Parallelism
von: Lau, Tim Tsz-Kit, et al.
Veröffentlicht: (2024)
von: Lau, Tim Tsz-Kit, et al.
Veröffentlicht: (2024)
Full-Graph vs. Mini-Batch Training: Comprehensive Analysis from a Batch Size and Fan-Out Size Perspective
von: Liu, Mengfan, et al.
Veröffentlicht: (2026)
von: Liu, Mengfan, et al.
Veröffentlicht: (2026)
Critical Batch Size Revisited: A Simple Empirical Approach to Large-Batch Language Model Training
von: Merrill, William, et al.
Veröffentlicht: (2025)
von: Merrill, William, et al.
Veröffentlicht: (2025)
Adaptive Batch Size and Learning Rate Scheduler for Stochastic Gradient Descent Based on Minimization of Stochastic First-order Oracle Complexity
von: Umeda, Hikaru, et al.
Veröffentlicht: (2025)
von: Umeda, Hikaru, et al.
Veröffentlicht: (2025)
Convergence Bound and Critical Batch Size of Muon Optimizer
von: Sato, Naoki, et al.
Veröffentlicht: (2025)
von: Sato, Naoki, et al.
Veröffentlicht: (2025)
Measuring Orthogonality as the Blind-Spot of Uncertainty Disentanglement
von: de Jong, Ivo Pascal, et al.
Veröffentlicht: (2024)
von: de Jong, Ivo Pascal, et al.
Veröffentlicht: (2024)
Batch Bayesian Active Learning with Partial Batch Label Sampling
von: Hu, Kangping, et al.
Veröffentlicht: (2025)
von: Hu, Kangping, et al.
Veröffentlicht: (2025)
Switching the Loss Reduces the Cost in Batch (Offline) Reinforcement Learning
von: Ayoub, Alex, et al.
Veröffentlicht: (2024)
von: Ayoub, Alex, et al.
Veröffentlicht: (2024)
Fast Catch-Up, Late Switching: Optimal Batch Size Scheduling via Functional Scaling Laws
von: Wang, Jinbo, et al.
Veröffentlicht: (2026)
von: Wang, Jinbo, et al.
Veröffentlicht: (2026)
Labeled TrustSet Guided: Batch Active Learning with Reinforcement Learning
von: Cui, Guofeng, et al.
Veröffentlicht: (2026)
von: Cui, Guofeng, et al.
Veröffentlicht: (2026)
On the Role of Batch Size in Stochastic Conditional Gradient Methods
von: Islamov, Rustem, et al.
Veröffentlicht: (2026)
von: Islamov, Rustem, et al.
Veröffentlicht: (2026)
Impact of Batch Normalization on Convolutional Network Representations
von: Potgieter, Hermanus L., et al.
Veröffentlicht: (2025)
von: Potgieter, Hermanus L., et al.
Veröffentlicht: (2025)
Batched Energy-Entropy acquisition for Bayesian Optimization
von: Teufel, Felix, et al.
Veröffentlicht: (2024)
von: Teufel, Felix, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Multi-Step Lookahead Information Via Adaptive Batching
von: Merlis, Nadav
Veröffentlicht: (2026)
von: Merlis, Nadav
Veröffentlicht: (2026)
Adaptive Batch-Wise Sample Scheduling for Direct Preference Optimization
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
Towards Batch-to-Streaming Deep Reinforcement Learning for Continuous Control
von: De Monte, Riccardo, et al.
Veröffentlicht: (2026)
von: De Monte, Riccardo, et al.
Veröffentlicht: (2026)
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
Scalable On-Policy Reinforcement Learning via Adaptive Batch Scaling
von: Park, Jongchan
Veröffentlicht: (2026)
von: Park, Jongchan
Veröffentlicht: (2026)
Increasing Batch Size Improves Convergence of Stochastic Gradient Descent with Momentum
von: Kamo, Keisuke, et al.
Veröffentlicht: (2025)
von: Kamo, Keisuke, et al.
Veröffentlicht: (2025)
Increasing Both Batch Size and Learning Rate Accelerates Stochastic Gradient Descent
von: Umeda, Hikaru, et al.
Veröffentlicht: (2024)
von: Umeda, Hikaru, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
von: Falzari, Massimiliano, et al.
Veröffentlicht: (2025) -
On The Presence of Double-Descent in Deep Reinforcement Learning
von: Veselý, Viktor, et al.
Veröffentlicht: (2025) -
Reinforcement Learning as an Improvement Heuristic for Real-World Production Scheduling
von: Müller, Arthur, et al.
Veröffentlicht: (2024) -
Upside-Down Reinforcement Learning for More Interpretable Optimal Control
von: Cardenas-Cartagena, Juan, et al.
Veröffentlicht: (2024) -
Video-Driven Graph Network-Based Simulators
von: Szewczyk, Franciszek, et al.
Veröffentlicht: (2024)