GPU Memory Usage Optimization for Backward Propagation in Deep Network Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hong, Ding-Yong, Tsai, Tzu-Hsien, Wang, Ning, Liu, Pangfeng, Wu, Jan-Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Training Overparametrized Neural Networks in Sublinear Time
von: Deng, Yichuan, et al.
Veröffentlicht: (2022)
von: Deng, Yichuan, et al.
Veröffentlicht: (2022)
Capacity Provisioning Motivated Online Non-Convex Optimization Problem with Memory and Switching Cost
von: Vaze, Rahul, et al.
Veröffentlicht: (2024)
von: Vaze, Rahul, et al.
Veröffentlicht: (2024)
Scaling Up Graph Propagation Computation on Large Graphs: A Local Chebyshev Approximation Approach
von: Yang, Yichun, et al.
Veröffentlicht: (2024)
von: Yang, Yichun, et al.
Veröffentlicht: (2024)
Optimal Clustering with Dependent Costs in Bayesian Networks
von: Wu, Paul Pao-Yen, et al.
Veröffentlicht: (2023)
von: Wu, Paul Pao-Yen, et al.
Veröffentlicht: (2023)
Memory-Efficient Sequential Pattern Mining with Hybrid Tries
von: Hosseininasab, Amin, et al.
Veröffentlicht: (2022)
von: Hosseininasab, Amin, et al.
Veröffentlicht: (2022)
Accelerating Matroid Optimization through Fast Imprecise Oracles
von: Eberle, Franziska, et al.
Veröffentlicht: (2024)
von: Eberle, Franziska, et al.
Veröffentlicht: (2024)
Understanding Memory-Regret Trade-Off for Streaming Stochastic Multi-Armed Bandits
von: He, Yuchen, et al.
Veröffentlicht: (2024)
von: He, Yuchen, et al.
Veröffentlicht: (2024)
Tight Gap-Dependent Memory-Regret Trade-Off for Single-Pass Streaming Stochastic Multi-Armed Bandits
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
von: Ye, Zichun, et al.
Veröffentlicht: (2025)
Approximation Algorithms for Combinatorial Optimization with Predictions
von: Antoniadis, Antonios, et al.
Veröffentlicht: (2024)
von: Antoniadis, Antonios, et al.
Veröffentlicht: (2024)
Semi-Bandit Learning for Monotone Stochastic Optimization
von: Agarwal, Arpit, et al.
Veröffentlicht: (2023)
von: Agarwal, Arpit, et al.
Veröffentlicht: (2023)
Optimization of Inter-group Criteria for Clustering with Minimum Size Constraints
von: Laber, Eduardo S., et al.
Veröffentlicht: (2024)
von: Laber, Eduardo S., et al.
Veröffentlicht: (2024)
Sample-and-Search: An Effective Algorithm for Learning-Augmented k-Median Clustering in High dimensions
von: Cheng, Kangke, et al.
Veröffentlicht: (2026)
von: Cheng, Kangke, et al.
Veröffentlicht: (2026)
Improved Robust Estimation for Erdős-Rényi Graphs: The Sparse Regime and Optimal Breakdown Point
von: Chen, Hongjie, et al.
Veröffentlicht: (2025)
von: Chen, Hongjie, et al.
Veröffentlicht: (2025)
Relax and Merge: A Simple Yet Effective Framework for Solving Fair $k$-Means and $k$-sparse Wasserstein Barycenter Problems
von: Song, Shihong, et al.
Veröffentlicht: (2024)
von: Song, Shihong, et al.
Veröffentlicht: (2024)
Private Edge Density Estimation for Random Graphs: Optimal, Efficient and Robust
von: Chen, Hongjie, et al.
Veröffentlicht: (2024)
von: Chen, Hongjie, et al.
Veröffentlicht: (2024)
Theoretically Grounded Pruning of Large Ground Sets for Constrained, Discrete Optimization
von: Nath, Ankur, et al.
Veröffentlicht: (2024)
von: Nath, Ankur, et al.
Veröffentlicht: (2024)
Sample-Efficient Optimization over Generative Priors via Coarse Learnability
von: Awasthi, Pranjal, et al.
Veröffentlicht: (2025)
von: Awasthi, Pranjal, et al.
Veröffentlicht: (2025)
Stochastic Bandits with ReLU Neural Networks
von: Xu, Kan, et al.
Veröffentlicht: (2024)
von: Xu, Kan, et al.
Veröffentlicht: (2024)
Minimum-Cost Network Flow with Dual Predictions
von: Chen, Zhiyang, et al.
Veröffentlicht: (2026)
von: Chen, Zhiyang, et al.
Veröffentlicht: (2026)
Ads that Stick: Near-Optimal Ad Optimization through Psychological Behavior Models
von: Darmasubramanian, Kailash Gopal, et al.
Veröffentlicht: (2025)
von: Darmasubramanian, Kailash Gopal, et al.
Veröffentlicht: (2025)
Handling Delayed Feedback in Distributed Online Optimization : A Projection-Free Approach
von: Nguyen, Tuan-Anh, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuan-Anh, et al.
Veröffentlicht: (2024)
Corporate Needs You to Find the Difference: Revisiting Submodular and Supermodular Ratio Optimization Problems
von: Harb, Elfarouk, et al.
Veröffentlicht: (2025)
von: Harb, Elfarouk, et al.
Veröffentlicht: (2025)
Differentially Private and Scalable Estimation of the Network Principal Component
von: Khayatian, Alireza, et al.
Veröffentlicht: (2025)
von: Khayatian, Alireza, et al.
Veröffentlicht: (2025)
$O(\sqrt{T})$ Static Regret and Instance Dependent Constraint Violation for Constrained Online Convex Optimization
von: Vaze, Rahul, et al.
Veröffentlicht: (2025)
von: Vaze, Rahul, et al.
Veröffentlicht: (2025)
Learning Neural Networks with Distribution Shift: Efficiently Certifiable Guarantees
von: Chandrasekaran, Gautam, et al.
Veröffentlicht: (2025)
von: Chandrasekaran, Gautam, et al.
Veröffentlicht: (2025)
Finite and Corruption-Robust Regret Bounds in Online Inverse Linear Optimization under M-Convex Action Sets
von: Oki, Taihei, et al.
Veröffentlicht: (2026)
von: Oki, Taihei, et al.
Veröffentlicht: (2026)
Dynamic Data Layout Optimization with Worst-case Guarantees
von: Rong, Kexin, et al.
Veröffentlicht: (2024)
von: Rong, Kexin, et al.
Veröffentlicht: (2024)
Local Fragments, Global Gains: Subgraph Counting using Graph Neural Networks
von: Roy, Shubhajit, et al.
Veröffentlicht: (2023)
von: Roy, Shubhajit, et al.
Veröffentlicht: (2023)
An Efficient Matrix Multiplication Algorithm for Accelerating Inference in Binary and Ternary Neural Networks
von: Dehghankar, Mohsen, et al.
Veröffentlicht: (2024)
von: Dehghankar, Mohsen, et al.
Veröffentlicht: (2024)
Graph Neural Network-Informed Predictive Flows for Faster Ford-Fulkerson and PAC-Learnability
von: Wiesler, Eleanor, et al.
Veröffentlicht: (2026)
von: Wiesler, Eleanor, et al.
Veröffentlicht: (2026)
Calibration Error for Decision Making
von: Hu, Lunjia, et al.
Veröffentlicht: (2024)
von: Hu, Lunjia, et al.
Veröffentlicht: (2024)
Near-optimal Swap Regret Minimization for Convex Losses
von: Hu, Lunjia, et al.
Veröffentlicht: (2026)
von: Hu, Lunjia, et al.
Veröffentlicht: (2026)
The Cost of Parallelizing Boosting
von: Lyu, Xin, et al.
Veröffentlicht: (2024)
von: Lyu, Xin, et al.
Veröffentlicht: (2024)
A Perfectly Truthful Calibration Measure
von: Hartline, Jason, et al.
Veröffentlicht: (2025)
von: Hartline, Jason, et al.
Veröffentlicht: (2025)
Smooth Calibration and Decision Making
von: Hartline, Jason, et al.
Veröffentlicht: (2025)
von: Hartline, Jason, et al.
Veröffentlicht: (2025)
Interpreting the Curse of Dimensionality from Distance Concentration and Manifold Effect
von: Peng, Dehua, et al.
Veröffentlicht: (2023)
von: Peng, Dehua, et al.
Veröffentlicht: (2023)
Changing Base Without Losing Pace: A GPU-Efficient Alternative to MatMul in DNNs
von: Ailon, Nir, et al.
Veröffentlicht: (2025)
von: Ailon, Nir, et al.
Veröffentlicht: (2025)
Towards Universal Convergence of Backward Error in Linear System Solvers
von: Dereziński, Michał, et al.
Veröffentlicht: (2026)
von: Dereziński, Michał, et al.
Veröffentlicht: (2026)
Truthful Calibration Errors for Multi-Class Prediction
von: Lu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Lu, Yuxuan, et al.
Veröffentlicht: (2025)
Online Prediction with Limited Selectivity
von: Liu, Licheng, et al.
Veröffentlicht: (2025)
von: Liu, Licheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Training Overparametrized Neural Networks in Sublinear Time
von: Deng, Yichuan, et al.
Veröffentlicht: (2022) -
Capacity Provisioning Motivated Online Non-Convex Optimization Problem with Memory and Switching Cost
von: Vaze, Rahul, et al.
Veröffentlicht: (2024) -
Scaling Up Graph Propagation Computation on Large Graphs: A Local Chebyshev Approximation Approach
von: Yang, Yichun, et al.
Veröffentlicht: (2024) -
Optimal Clustering with Dependent Costs in Bayesian Networks
von: Wu, Paul Pao-Yen, et al.
Veröffentlicht: (2023) -
Memory-Efficient Sequential Pattern Mining with Hybrid Tries
von: Hosseininasab, Amin, et al.
Veröffentlicht: (2022)