PipeOptim: Ensuring Effective 1F1B Schedule with Optimizer-Dependent Weight Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Guan, Lei, Li, Dongsheng, Chen, Yongle, Liang, Jiye, Wang, Wenjian, Lu, Xicheng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
XGrad: Boosting Gradient-Based Optimizers With Weight Prediction
by: Guan, Lei, et al.
Published: (2023)
by: Guan, Lei, et al.
Published: (2023)
FlashOptim: Optimizers for Memory-Efficient Training
by: Ortiz, Jose Javier Gonzalez, et al.
Published: (2026)
by: Ortiz, Jose Javier Gonzalez, et al.
Published: (2026)
DualOptim: Enhancing Efficacy and Stability in Machine Unlearning with Dual Optimizers
by: Zhong, Xuyang, et al.
Published: (2025)
by: Zhong, Xuyang, et al.
Published: (2025)
OptimAI: Optimization from Natural Language Using LLM-Powered AI Agents
by: Thind, Raghav, et al.
Published: (2025)
by: Thind, Raghav, et al.
Published: (2025)
X-SAM: Boosting Sharpness-Aware Minimization with Dominant-Eigenvector Gradient Correction
by: Duan, Hongru, et al.
Published: (2026)
by: Duan, Hongru, et al.
Published: (2026)
KC-GenRe: A Knowledge-constrained Generative Re-ranking Method Based on Large Language Models for Knowledge Graph Completion
by: Wang, Yilin, et al.
Published: (2024)
by: Wang, Yilin, et al.
Published: (2024)
OptPipe: Memory- and Scheduling-Optimized Pipeline Parallelism for LLM Training
by: Li, Hongpei, et al.
Published: (2025)
by: Li, Hongpei, et al.
Published: (2025)
Predict+Optimize Problem in Renewable Energy Scheduling
by: Bergmeir, Christoph, et al.
Published: (2022)
by: Bergmeir, Christoph, et al.
Published: (2022)
Learnable Game-theoretic Policy Optimization for Data-centric Self-explanation Rationalization
by: Zhao, Yunxiao, et al.
Published: (2025)
by: Zhao, Yunxiao, et al.
Published: (2025)
Ensured: Explanations for Decreasing the Epistemic Uncertainty in Predictions
by: Löfström, Helena, et al.
Published: (2024)
by: Löfström, Helena, et al.
Published: (2024)
PipeSpec: Breaking Stage Dependencies in Hierarchical LLM Decoding
by: McDanel, Bradley, et al.
Published: (2025)
by: McDanel, Bradley, et al.
Published: (2025)
Automated proving in planar geometry based on the complex number identity method and elimination
by: Kovács, Zoltán, et al.
Published: (2025)
by: Kovács, Zoltán, et al.
Published: (2025)
Experience is the Best Teacher: Motivating Effective Exploration in Reinforcement Learning for LLMs
by: Zhang, Wenjian, et al.
Published: (2026)
by: Zhang, Wenjian, et al.
Published: (2026)
Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization
by: Huang, Jiahao, et al.
Published: (2026)
by: Huang, Jiahao, et al.
Published: (2026)
Constrained Language Model Policy Optimization via Risk-aware Stepwise Alignment
by: Zhang, Lijun, et al.
Published: (2025)
by: Zhang, Lijun, et al.
Published: (2025)
STAHGNet: Modeling Hybrid-grained Heterogenous Dependency Efficiently for Traffic Prediction
by: Wang, Jiyao, et al.
Published: (2024)
by: Wang, Jiyao, et al.
Published: (2024)
ME$^3$-BEV: Mamba-Enhanced Deep Reinforcement Learning for End-to-End Autonomous Driving with BEV-Perception
by: Lu, Siyi, et al.
Published: (2025)
by: Lu, Siyi, et al.
Published: (2025)
Evolved Sample Weights for Bias Mitigation: Effectiveness Depends on the Fairness Objective
by: Saini, Anil K., et al.
Published: (2025)
by: Saini, Anil K., et al.
Published: (2025)
Runtime Analysis of Evolutionary Algorithms for Multi-party Multi-objective Optimization
by: Sun, Yuetong, et al.
Published: (2025)
by: Sun, Yuetong, et al.
Published: (2025)
TawPipe: Topology-Aware Weight Pipeline Parallelism for Accelerating Long-Context Large Models Training
by: Wu, Houming, et al.
Published: (2025)
by: Wu, Houming, et al.
Published: (2025)
Leveraging Generative AI for Clinical Evidence Summarization Needs to Ensure Trustworthiness
by: Zhang, Gongbo, et al.
Published: (2023)
by: Zhang, Gongbo, et al.
Published: (2023)
FinePhys: Fine-grained Human Action Generation by Explicitly Incorporating Physical Laws for Effective Skeletal Guidance
by: Shao, Dian, et al.
Published: (2025)
by: Shao, Dian, et al.
Published: (2025)
Benchmark for CEC 2024 Competition on Multiparty Multiobjective Optimization
by: Luo, Wenjian, et al.
Published: (2024)
by: Luo, Wenjian, et al.
Published: (2024)
Talking like Piping and Instrumentation Diagrams (P&IDs)
by: Alimin, Achmad Anggawirya, et al.
Published: (2025)
by: Alimin, Achmad Anggawirya, et al.
Published: (2025)
Improving Large Language Models in Event Relation Logical Prediction
by: Chen, Meiqi, et al.
Published: (2023)
by: Chen, Meiqi, et al.
Published: (2023)
When Speed meets Accuracy: an Efficient and Effective Graph Model for Temporal Link Prediction
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
Beyond Accuracy: Ensuring Correct Predictions With Correct Rationales
by: Li, Tang, et al.
Published: (2024)
by: Li, Tang, et al.
Published: (2024)
Trajectory Entropy: Modeling Game State Stability from Multimodality Trajectory Prediction
by: Zhang, Yesheng, et al.
Published: (2025)
by: Zhang, Yesheng, et al.
Published: (2025)
Domain-Agnostic Scalable AI Safety Ensuring Framework
by: Kim, Beomjun, et al.
Published: (2025)
by: Kim, Beomjun, et al.
Published: (2025)
Optimization with SpotOptim
by: Bartz-Beielstein, Thomas
Published: (2026)
by: Bartz-Beielstein, Thomas
Published: (2026)
wd1: Weighted Policy Optimization for Reasoning in Diffusion Language Models
by: Tang, Xiaohang, et al.
Published: (2025)
by: Tang, Xiaohang, et al.
Published: (2025)
Game-Oriented ASR Error Correction via RAG-Enhanced LLM
by: Jiang, Yan, et al.
Published: (2025)
by: Jiang, Yan, et al.
Published: (2025)
A Novel Immune Algorithm for Multiparty Multiobjective Optimization
by: Chen, Kesheng, et al.
Published: (2026)
by: Chen, Kesheng, et al.
Published: (2026)
PipeOffload: Improving Scalability of Pipeline Parallelism with Memory Optimization
by: Wan, Xinyi, et al.
Published: (2025)
by: Wan, Xinyi, et al.
Published: (2025)
Dependency-Aware CAV Task Scheduling via Diffusion-Based Reinforcement Learning
by: Cheng, Xiang, et al.
Published: (2024)
by: Cheng, Xiang, et al.
Published: (2024)
Showing Many Labels in Multi-label Classification Models: An Empirical Study of Adversarial Examples
by: Liu, Yujiang, et al.
Published: (2024)
by: Liu, Yujiang, et al.
Published: (2024)
Analysis of Schedule-Free Nonconvex Optimization
by: Brown, Connor
Published: (2025)
by: Brown, Connor
Published: (2025)
DSevolve: Enabling Real-Time Adaptive Scheduling on Dynamic Shop Floor with LLM-Evolved Heuristic Portfolios
by: Huang, Jin, et al.
Published: (2026)
by: Huang, Jin, et al.
Published: (2026)
Time-varying Interaction Graph ODE for Dynamic Graph Representation Learning
by: Wang, Xiaoyi, et al.
Published: (2026)
by: Wang, Xiaoyi, et al.
Published: (2026)
ShipTraj-R1: Reinforcing Ship Trajectory Prediction in Large Language Models via Group Relative Policy Optimization
by: Zhan, Yang, et al.
Published: (2026)
by: Zhan, Yang, et al.
Published: (2026)
Similar Items
-
XGrad: Boosting Gradient-Based Optimizers With Weight Prediction
by: Guan, Lei, et al.
Published: (2023) -
FlashOptim: Optimizers for Memory-Efficient Training
by: Ortiz, Jose Javier Gonzalez, et al.
Published: (2026) -
DualOptim: Enhancing Efficacy and Stability in Machine Unlearning with Dual Optimizers
by: Zhong, Xuyang, et al.
Published: (2025) -
OptimAI: Optimization from Natural Language Using LLM-Powered AI Agents
by: Thind, Raghav, et al.
Published: (2025) -
X-SAM: Boosting Sharpness-Aware Minimization with Dominant-Eigenvector Gradient Correction
by: Duan, Hongru, et al.
Published: (2026)