FT K-means: A High-Performance K-means on GPU with Fault Tolerance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Shixun, Ding, Yitong, Zhai, Yujia, Liu, Jinyang, Huang, Jiajun, Jian, Zizhe, Dai, Huangliang, Di, Sheng, Wong, Bryan M., Chen, Zizhong, Cappello, Franck |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TurboFFT: A High-Performance Fast Fourier Transform with Fault Tolerance on GPU
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
TurboFFT: Co-Designed High-Performance and Fault-Tolerant Fast Fourier Transform on GPUs
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
FT-Transformer: Resilient and Reliable Transformer with End-to-End Fault Tolerant Attention
von: Dai, Huangliang, et al.
Veröffentlicht: (2025)
von: Dai, Huangliang, et al.
Veröffentlicht: (2025)
TurboFNO: High-Performance Fourier Neural Operator with Fused FFT-GEMM-iFFT on GPU
von: Wu, Shixun, et al.
Veröffentlicht: (2025)
von: Wu, Shixun, et al.
Veröffentlicht: (2025)
DGRO: Diameter-Guided Ring Optimization for Integrated Research Infrastructure Membership
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
gZCCL: Compression-Accelerated Collective Communication Framework for GPU Clusters
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
High-performance Effective Scientific Error-bounded Lossy Compression with Auto-tuned Multi-component Interpolation
von: Liu, Jinyang, et al.
Veröffentlicht: (2023)
von: Liu, Jinyang, et al.
Veröffentlicht: (2023)
Boosting Scientific Error-Bounded Lossy Compression through Optimized Synergistic Lossy-Lossless Orchestration
von: Wu, Shixun, et al.
Veröffentlicht: (2025)
von: Wu, Shixun, et al.
Veröffentlicht: (2025)
cuSZ-$i$: High-Ratio Scientific Lossy Compression on GPUs with Optimized Multi-Level Interpolation
von: Liu, Jinyang, et al.
Veröffentlicht: (2023)
von: Liu, Jinyang, et al.
Veröffentlicht: (2023)
An Optimized Error-controlled MPI Collective Framework Integrated with Lossy Compression
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
ZCCL: Significantly Improving Collective Communication With Error-Bounded Lossy Compression
von: Huang, Jiajun, et al.
Veröffentlicht: (2025)
von: Huang, Jiajun, et al.
Veröffentlicht: (2025)
GPZ: GPU-Accelerated Lossy Compressor for Particle Data
von: Li, Ruoyu, et al.
Veröffentlicht: (2025)
von: Li, Ruoyu, et al.
Veröffentlicht: (2025)
IPComp: Interpolation Based Progressive Lossy Compression for Scientific Applications
von: Yang, Zhuoxun, et al.
Veröffentlicht: (2025)
von: Yang, Zhuoxun, et al.
Veröffentlicht: (2025)
Breaking the Memory Wall: A Study of I/O Patterns and GPU Memory Utilization for Hybrid CPU-GPU Offloaded Optimizers
von: Maurya, Avinash, et al.
Veröffentlicht: (2024)
von: Maurya, Avinash, et al.
Veröffentlicht: (2024)
BlockRaFT: A Distributed Framework for Fault-Tolerant and Scalable Blockchain Nodes
von: Piduguralla, Manaswini, et al.
Veröffentlicht: (2026)
von: Piduguralla, Manaswini, et al.
Veröffentlicht: (2026)
SPARe: Stacked Parallelism with Adaptive Reordering for Fault-Tolerant LLM Pretraining Systems with 100k+ GPUs
von: Lee, Jin, et al.
Veröffentlicht: (2026)
von: Lee, Jin, et al.
Veröffentlicht: (2026)
ReCoVer: Resilient LLM Pre-Training System via Fault-Tolerant Collective and Versatile Workload
von: Liu, Ziyue, et al.
Veröffentlicht: (2026)
von: Liu, Ziyue, et al.
Veröffentlicht: (2026)
Federated K-means Clustering
von: Garst, Swier, et al.
Veröffentlicht: (2023)
von: Garst, Swier, et al.
Veröffentlicht: (2023)
Popcorn: Accelerating Kernel K-means on GPUs through Sparse Linear Algebra
von: Bellavita, Julian, et al.
Veröffentlicht: (2025)
von: Bellavita, Julian, et al.
Veröffentlicht: (2025)
Characterization-Guided GPU Fault Resilience in NVIDIA MPS
von: Liu, Rixin, et al.
Veröffentlicht: (2026)
von: Liu, Rixin, et al.
Veröffentlicht: (2026)
To Compress or Not To Compress: Energy Trade-Offs and Benefits of Lossy Compressed I/O
von: Wilkins, Grant, et al.
Veröffentlicht: (2024)
von: Wilkins, Grant, et al.
Veröffentlicht: (2024)
Beyond Optimal Fault Tolerance
von: Lewis-Pye, Andrew, et al.
Veröffentlicht: (2025)
von: Lewis-Pye, Andrew, et al.
Veröffentlicht: (2025)
Heat: Satellite's meat is GPU's poison
von: Yuan, Zhehu, et al.
Veröffentlicht: (2024)
von: Yuan, Zhehu, et al.
Veröffentlicht: (2024)
Improving GPU Multi-Tenancy Through Dynamic Multi-Instance GPU Reconfiguration
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
Stream-K Optimization and Exploration
von: Rackley, Nick, et al.
Veröffentlicht: (2024)
von: Rackley, Nick, et al.
Veröffentlicht: (2024)
HoSZp: An Efficient Homomorphic Error-bounded Lossy Compressor for Scientific Data
von: Agarwal, Tripti, et al.
Veröffentlicht: (2024)
von: Agarwal, Tripti, et al.
Veröffentlicht: (2024)
Byzantine Fault Tolerant Causal Ordering
von: Misra, Anshuman, et al.
Veröffentlicht: (2021)
von: Misra, Anshuman, et al.
Veröffentlicht: (2021)
Byzantine Fault-Tolerant Min-Max Optimization
von: Liu, Shuo, et al.
Veröffentlicht: (2022)
von: Liu, Shuo, et al.
Veröffentlicht: (2022)
Approximate Byzantine Fault-Tolerance in Distributed Optimization
von: Liu, Shuo, et al.
Veröffentlicht: (2021)
von: Liu, Shuo, et al.
Veröffentlicht: (2021)
Optimal Fault-Tolerant Dispersion on Oriented Grids
von: Banerjee, Rik, et al.
Veröffentlicht: (2024)
von: Banerjee, Rik, et al.
Veröffentlicht: (2024)
Probabilistic Byzantine Fault Tolerance (Extended Version)
von: Avelãs, Diogo, et al.
Veröffentlicht: (2024)
von: Avelãs, Diogo, et al.
Veröffentlicht: (2024)
Byzantine-Tolerant Consensus in GPU-Inspired Shared Memory
von: Georgiou, Chryssis, et al.
Veröffentlicht: (2025)
von: Georgiou, Chryssis, et al.
Veröffentlicht: (2025)
Optimizing Robot Dispersion on Grids: with and without Fault Tolerance
von: Banerjee, Rik, et al.
Veröffentlicht: (2024)
von: Banerjee, Rik, et al.
Veröffentlicht: (2024)
VBFT: Veloce Byzantine Fault Tolerant Consensus for Blockchains
von: Jalalzai, Mohammad M., et al.
Veröffentlicht: (2023)
von: Jalalzai, Mohammad M., et al.
Veröffentlicht: (2023)
A Fault Tolerance Mechanism for Hybrid Scientific Workflows
von: Mulone, Alberto, et al.
Veröffentlicht: (2024)
von: Mulone, Alberto, et al.
Veröffentlicht: (2024)
Asynchronous Fault-Tolerant Distributed Proper Coloring of Graphs
von: Balliu, Alkida, et al.
Veröffentlicht: (2024)
von: Balliu, Alkida, et al.
Veröffentlicht: (2024)
Arma: Byzantine Fault Tolerant Consensus with Horizontal Scalability
von: Manevich, Yacov, et al.
Veröffentlicht: (2024)
von: Manevich, Yacov, et al.
Veröffentlicht: (2024)
The Case for ABI Interoperability in a Fault Tolerant MPI
von: Xu, Yao, et al.
Veröffentlicht: (2025)
von: Xu, Yao, et al.
Veröffentlicht: (2025)
Stabl: Blockchain Fault Tolerance
von: Gramoli, Vincent, et al.
Veröffentlicht: (2024)
von: Gramoli, Vincent, et al.
Veröffentlicht: (2024)
Mitigating Artifacts in Pre-quantization Based Scientific Data Compressors with Quantization-aware Interpolation
von: Jiao, Pu, et al.
Veröffentlicht: (2026)
von: Jiao, Pu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
TurboFFT: A High-Performance Fast Fourier Transform with Fault Tolerance on GPU
von: Wu, Shixun, et al.
Veröffentlicht: (2024) -
TurboFFT: Co-Designed High-Performance and Fault-Tolerant Fast Fourier Transform on GPUs
von: Wu, Shixun, et al.
Veröffentlicht: (2024) -
FT-Transformer: Resilient and Reliable Transformer with End-to-End Fault Tolerant Attention
von: Dai, Huangliang, et al.
Veröffentlicht: (2025) -
TurboFNO: High-Performance Fourier Neural Operator with Fused FFT-GEMM-iFFT on GPU
von: Wu, Shixun, et al.
Veröffentlicht: (2025) -
DGRO: Diameter-Guided Ring Optimization for Integrated Research Infrastructure Membership
von: Wu, Shixun, et al.
Veröffentlicht: (2024)