EinDecomp: Decomposition of Declaratively-Specified Machine Learning and Numerical Computations for Parallel Execution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bourgeois, Daniel, Ding, Zhimin, Jankov, Dimitrije, Li, Jiehui, Sleem, Mahmoud, Tang, Yuxin, Yao, Jiawen, Yao, Xinyu, Jermaine, Chris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TURNIP: A "Nondeterministic" GPU Runtime with CPU RAM Offload
von: Ding, Zhimin, et al.
Veröffentlicht: (2024)
von: Ding, Zhimin, et al.
Veröffentlicht: (2024)
DOPPLER: Dual-Policy Learning for Device Assignment in Asynchronous Dataflow Graphs
von: Yao, Xinyu, et al.
Veröffentlicht: (2025)
von: Yao, Xinyu, et al.
Veröffentlicht: (2025)
Declarative Data Pipeline for Large Scale ML Services
von: Yang, Yunzhao, et al.
Veröffentlicht: (2025)
von: Yang, Yunzhao, et al.
Veröffentlicht: (2025)
Efficient Parallel Execution of Blockchain Transactions Leveraging Conflict Specifications
von: Anjana, Parwat Singh, et al.
Veröffentlicht: (2025)
von: Anjana, Parwat Singh, et al.
Veröffentlicht: (2025)
GPU-Based Parallel Computing Methods for Medical Photoacoustic Image Reconstruction
von: Yi, Xinyao, et al.
Veröffentlicht: (2024)
von: Yi, Xinyao, et al.
Veröffentlicht: (2024)
Boosting Blockchain Throughput: Parallel EVM Execution with Asynchronous Storage for Reddio
von: Qi, Xiaodong, et al.
Veröffentlicht: (2025)
von: Qi, Xiaodong, et al.
Veröffentlicht: (2025)
Deferred Objects to Enhance Smart Contract Programming with Optimistic Parallel Execution
von: Mitenkov, George, et al.
Veröffentlicht: (2024)
von: Mitenkov, George, et al.
Veröffentlicht: (2024)
Distributed And Parallel Low-Diameter Decompositions for Arbitrary and Restricted Graphs
von: Dou, Jinfeng, et al.
Veröffentlicht: (2024)
von: Dou, Jinfeng, et al.
Veröffentlicht: (2024)
APEX: An Extensible and Dynamism-Aware Simulator for Automated Parallel Execution in LLM Serving
von: Lin, Yi-Chien, et al.
Veröffentlicht: (2024)
von: Lin, Yi-Chien, et al.
Veröffentlicht: (2024)
NEMO: Faster Parallel Execution for Highly Contended Blockchain Workloads (Full version)
von: Ezard, François, et al.
Veröffentlicht: (2025)
von: Ezard, François, et al.
Veröffentlicht: (2025)
Committee Configuration Optimization for Parallel Byzantine Consensus in a Trusted Execution Environment
von: Xie, Yifei, et al.
Veröffentlicht: (2026)
von: Xie, Yifei, et al.
Veröffentlicht: (2026)
Load Balanced Parallel Node Generation for Meshless Numerical Methods
von: Vehovar, Jon, et al.
Veröffentlicht: (2026)
von: Vehovar, Jon, et al.
Veröffentlicht: (2026)
A Parallel and Distributed Rust Library for Core Decomposition on Large Graphs
von: Rucci, Davide, et al.
Veröffentlicht: (2025)
von: Rucci, Davide, et al.
Veröffentlicht: (2025)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
Fractional Payment Transactions: Executing Payment Transactions in Parallel with Less than f+1 Validations
von: Bazzi, Rida, et al.
Veröffentlicht: (2024)
von: Bazzi, Rida, et al.
Veröffentlicht: (2024)
Malleus: Straggler-Resilient Hybrid Parallel Training of Large-scale Models via Malleable Data and Model Parallelization
von: Li, Haoyang, et al.
Veröffentlicht: (2024)
von: Li, Haoyang, et al.
Veröffentlicht: (2024)
A New Execution Model and Executor for Adaptively Optimizing the Performance of Parallel Algorithms Using HPX Runtime System
von: Mohammadiporshokooh, Karame, et al.
Veröffentlicht: (2025)
von: Mohammadiporshokooh, Karame, et al.
Veröffentlicht: (2025)
ParallelSFL: A Novel Split Federated Learning Framework Tackling Heterogeneity Issues
von: Liao, Yunming, et al.
Veröffentlicht: (2024)
von: Liao, Yunming, et al.
Veröffentlicht: (2024)
Declarative Application Management in the Fog. A bacteria-inspired decentralised approach
von: Brogi, Antonio, et al.
Veröffentlicht: (2025)
von: Brogi, Antonio, et al.
Veröffentlicht: (2025)
Combining Declarative and Linear Programming for Application Management in the Cloud-Edge Continuum
von: Massa, Jacopo, et al.
Veröffentlicht: (2025)
von: Massa, Jacopo, et al.
Veröffentlicht: (2025)
Proteus: Append-Only Ledgers for (Mostly) Trusted Execution Environments
von: Mishra, Shubham, et al.
Veröffentlicht: (2026)
von: Mishra, Shubham, et al.
Veröffentlicht: (2026)
cuFastTuckerPlus: A Stochastic Parallel Sparse FastTucker Decomposition Using GPU Tensor Cores
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
STELLAR: Storage Tuning Engine Leveraging LLM Autonomous Reasoning for High Performance Parallel File Systems
von: Egersdoerfer, Chris, et al.
Veröffentlicht: (2026)
von: Egersdoerfer, Chris, et al.
Veröffentlicht: (2026)
Exploring System-Heterogeneous Federated Learning with Dynamic Model Selection
von: Yao, Dixi
Veröffentlicht: (2024)
von: Yao, Dixi
Veröffentlicht: (2024)
TENT: A Declarative Slice Spraying Engine for Performant and Resilient Data Movement in Disaggregated LLM Serving
von: Ren, Feng, et al.
Veröffentlicht: (2026)
von: Ren, Feng, et al.
Veröffentlicht: (2026)
Balancing Pipeline Parallelism with Vocabulary Parallelism
von: Yeung, Man Tsung, et al.
Veröffentlicht: (2024)
von: Yeung, Man Tsung, et al.
Veröffentlicht: (2024)
The GA4GH Task Execution API: Enabling Easy Multi Cloud Task Execution
von: Kanitz, Alexander, et al.
Veröffentlicht: (2024)
von: Kanitz, Alexander, et al.
Veröffentlicht: (2024)
DreamDDP: Accelerating Data Parallel Distributed LLM Training with Layer-wise Scheduled Partial Synchronization
von: Tang, Zhenheng, et al.
Veröffentlicht: (2025)
von: Tang, Zhenheng, et al.
Veröffentlicht: (2025)
ElasWave: An Elastic-Native System for Scalable Hybrid-Parallel Training
von: Kang, Xueze, et al.
Veröffentlicht: (2025)
von: Kang, Xueze, et al.
Veröffentlicht: (2025)
Pilotfish: Distributed Execution for Scalable Blockchains
von: Kniep, Quentin, et al.
Veröffentlicht: (2024)
von: Kniep, Quentin, et al.
Veröffentlicht: (2024)
The Entropy of Parallel Systems
von: Adefemi, Temitayo
Veröffentlicht: (2025)
von: Adefemi, Temitayo
Veröffentlicht: (2025)
Lectures on Parallel Computing
von: Träff, Jesper Larsson
Veröffentlicht: (2024)
von: Träff, Jesper Larsson
Veröffentlicht: (2024)
Parallel Algorithms for Hierarchical Nucleus Decomposition
von: Shi, Jessica, et al.
Veröffentlicht: (2023)
von: Shi, Jessica, et al.
Veröffentlicht: (2023)
High-Performance N-Queens Solver on GPU: Iterative DFS with Zero Bank Conflicts
von: Yao, Guangchao, et al.
Veröffentlicht: (2025)
von: Yao, Guangchao, et al.
Veröffentlicht: (2025)
Unleashing Multicore Strength for Efficient Execution of Transactions
von: Ravish, Ankit, et al.
Veröffentlicht: (2024)
von: Ravish, Ankit, et al.
Veröffentlicht: (2024)
Orchestrating the Execution of Serverless Functions in Hybrid Clouds
von: Peri, Aristotelis, et al.
Veröffentlicht: (2024)
von: Peri, Aristotelis, et al.
Veröffentlicht: (2024)
Distributed Speculative Execution for Resilient Cloud Applications
von: Li, Tianyu, et al.
Veröffentlicht: (2024)
von: Li, Tianyu, et al.
Veröffentlicht: (2024)
Zenix: Efficient Execution of Bulky Serverless Applications
von: Guo, Zhiyuan, et al.
Veröffentlicht: (2022)
von: Guo, Zhiyuan, et al.
Veröffentlicht: (2022)
NanoCP: Request-Level Dynamic Context Parallelism for Data-Expert Parallel Decoding
von: Chen, Jiefei, et al.
Veröffentlicht: (2026)
von: Chen, Jiefei, et al.
Veröffentlicht: (2026)
ZeroPP: Unleashing Exceptional Parallelism Efficiency through Tensor-Parallelism-Free Methodology
von: Tang, Ding, et al.
Veröffentlicht: (2024)
von: Tang, Ding, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TURNIP: A "Nondeterministic" GPU Runtime with CPU RAM Offload
von: Ding, Zhimin, et al.
Veröffentlicht: (2024) -
DOPPLER: Dual-Policy Learning for Device Assignment in Asynchronous Dataflow Graphs
von: Yao, Xinyu, et al.
Veröffentlicht: (2025) -
Declarative Data Pipeline for Large Scale ML Services
von: Yang, Yunzhao, et al.
Veröffentlicht: (2025) -
Efficient Parallel Execution of Blockchain Transactions Leveraging Conflict Specifications
von: Anjana, Parwat Singh, et al.
Veröffentlicht: (2025) -
GPU-Based Parallel Computing Methods for Medical Photoacoustic Image Reconstruction
von: Yi, Xinyao, et al.
Veröffentlicht: (2024)