High-Performance Star-M SVD for Big Data Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Hussain, Md Taufique, Ballard, Grey, Devarakonda, Aditya, Eswar, Srinivas, Pesricha, Naman, Rao, Vishwas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Distributed-memory Algorithms for Sparse Matrix Permutation, Extraction, and Assignment
by: Hassani, Elaheh, et al.
Published: (2025)
by: Hassani, Elaheh, et al.
Published: (2025)
Communication Lower Bounds and Algorithms for Sketching with Random Dense Matrices
by: Daas, Hussam Al, et al.
Published: (2026)
by: Daas, Hussam Al, et al.
Published: (2026)
Cost-Effective Big Data Orchestration Using Dagster: A Multi-Platform Approach
by: Picatto, Hernan, et al.
Published: (2024)
by: Picatto, Hernan, et al.
Published: (2024)
Distributed-Memory Parallel Algorithms for Sparse Matrix and Sparse Tall-and-Skinny Matrix Multiplication
by: Ranawaka, Isuru, et al.
Published: (2024)
by: Ranawaka, Isuru, et al.
Published: (2024)
A Unifying Framework to Enable Artificial Intelligence in High Performance Computing Workflows
by: Domke, Jens, et al.
Published: (2025)
by: Domke, Jens, et al.
Published: (2025)
ATOM: Asynchronous Training of Massive Models for Deep Learning in a Decentralized Environment
by: Wu, Xiaofeng, et al.
Published: (2024)
by: Wu, Xiaofeng, et al.
Published: (2024)
Do Large Language Models Understand Performance Optimization?
by: Cui, Bowen, et al.
Published: (2025)
by: Cui, Bowen, et al.
Published: (2025)
Efficiently Reproducing Distributed Workflows in Notebook-based Systems
by: Azaz, Talha, et al.
Published: (2026)
by: Azaz, Talha, et al.
Published: (2026)
SPES: Towards Optimizing Performance-Resource Trade-Off for Serverless Functions
by: Lee, Cheryl, et al.
Published: (2024)
by: Lee, Cheryl, et al.
Published: (2024)
Toward Portable GPU Performance: Julia Recursive Implementation of TRMM and TRSM
by: Carrica, Vicki, et al.
Published: (2025)
by: Carrica, Vicki, et al.
Published: (2025)
Comprehensive Review of Performance Optimization Strategies for Serverless Applications on AWS Lambda
by: Bechir, Mohamed Lemine El, et al.
Published: (2024)
by: Bechir, Mohamed Lemine El, et al.
Published: (2024)
Implementing Multi-GPU Scientific Computing Miniapps Across Performance Portable Frameworks
by: Villalobos, Johansell, et al.
Published: (2025)
by: Villalobos, Johansell, et al.
Published: (2025)
Communication Lower Bounds and Optimal Algorithms for Symmetric Matrix Computations
by: Daas, Hussam Al, et al.
Published: (2024)
by: Daas, Hussam Al, et al.
Published: (2024)
Minimizing Communication for Parallel Symmetric Tensor Times Same Vector Computation
by: Daas, Hussam Al, et al.
Published: (2025)
by: Daas, Hussam Al, et al.
Published: (2025)
Performant Unified GPU Kernels for Portable Singular Value Computation Across Hardware and Precision
by: Ringoot, Evelyne, et al.
Published: (2025)
by: Ringoot, Evelyne, et al.
Published: (2025)
Efficient N-to-M Checkpointing Algorithm for Finite Element Simulations
by: Ham, David A., et al.
Published: (2024)
by: Ham, David A., et al.
Published: (2024)
Configurable Runtime Orchestration for Dynamic Data Retrieval in Distributed Systems
by: Kandiraju, Abhiram
Published: (2026)
by: Kandiraju, Abhiram
Published: (2026)
Cost-Performance Analysis of Cloud-Based Retail Point-of-Sale Systems: A Comparative Study of Google Cloud Platform and Microsoft Azure
by: Pagidoju, Ravi Teja
Published: (2026)
by: Pagidoju, Ravi Teja
Published: (2026)
Performant Automatic BLAS Offloading on Unified Memory Architecture with OpenMP First-Touch Style Data Movement
by: Li, Junjie
Published: (2024)
by: Li, Junjie
Published: (2024)
ShuffleBench: A Benchmark for Large-Scale Data Shuffling Operations with Distributed Stream Processing Frameworks
by: Henning, Sören, et al.
Published: (2024)
by: Henning, Sören, et al.
Published: (2024)
CLAID: Closing the Loop on AI & Data Collection -- A Cross-Platform Transparent Computing Middleware Framework for Smart Edge-Cloud and Digital Biomarker Applications
by: Langer, Patrick, et al.
Published: (2023)
by: Langer, Patrick, et al.
Published: (2023)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
by: Ramesh, Risshab Srinivas
Published: (2024)
by: Ramesh, Risshab Srinivas
Published: (2024)
MPI Implementation Profiling for Better Application Performance
by: Shipley, Riley, et al.
Published: (2024)
by: Shipley, Riley, et al.
Published: (2024)
Modular Architecture for High-Performance and Low Overhead Data Transfers
by: Swargo, Rasman Mubtasim, et al.
Published: (2025)
by: Swargo, Rasman Mubtasim, et al.
Published: (2025)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
by: Nichols, Daniel, et al.
Published: (2025)
by: Nichols, Daniel, et al.
Published: (2025)
Performance measurements of modern Fortran MPI applications with Score-P
by: Corbin, Gregor
Published: (2025)
by: Corbin, Gregor
Published: (2025)
NApy: Efficient Statistics in Python for Large-Scale Heterogeneous Data with Enhanced Support for Missing Data
by: Woller, Fabian, et al.
Published: (2025)
by: Woller, Fabian, et al.
Published: (2025)
Federated Learning within Global Energy Budget over Heterogeneous Edge Accelerators
by: Banerjee, Roopkatha, et al.
Published: (2025)
by: Banerjee, Roopkatha, et al.
Published: (2025)
DistShap: Scalable GNN Explanations with Distributed Shapley Values
by: Akkas, Selahattin, et al.
Published: (2025)
by: Akkas, Selahattin, et al.
Published: (2025)
MARCO: Multi-Agent Code Optimization with Real-Time Knowledge Integration for High-Performance Computing
by: Rahman, Asif, et al.
Published: (2025)
by: Rahman, Asif, et al.
Published: (2025)
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
by: Tariq, Syed Salauddin Mohammad, et al.
Published: (2024)
by: Tariq, Syed Salauddin Mohammad, et al.
Published: (2024)
KnapsackLB: Enabling Performance-Aware Layer-4 Load Balancing
by: Gandhi, Rohan, et al.
Published: (2024)
by: Gandhi, Rohan, et al.
Published: (2024)
High-level Stream Processing: A Complementary Analysis of Fault Recovery
by: Vogel, Adriano, et al.
Published: (2024)
by: Vogel, Adriano, et al.
Published: (2024)
When Should I Run My Application Benchmark?: Studying Cloud Performance Variability for the Case of Stream Processing Applications
by: Henning, Sören, et al.
Published: (2025)
by: Henning, Sören, et al.
Published: (2025)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
by: Diehl, Patrick, et al.
Published: (2025)
by: Diehl, Patrick, et al.
Published: (2025)
SeBS-Flow: Benchmarking Serverless Cloud Function Workflows
by: Schmid, Larissa, et al.
Published: (2024)
by: Schmid, Larissa, et al.
Published: (2024)
CloudHeatMap: Heatmap-Based Monitoring for Large-Scale Cloud Systems
by: Sohana, Sarah, et al.
Published: (2024)
by: Sohana, Sarah, et al.
Published: (2024)
Integrating Odeint Time Stepping into OpenFPM for Distributed and GPU Accelerated Numerical Solvers
by: Singh, Abhinav, et al.
Published: (2023)
by: Singh, Abhinav, et al.
Published: (2023)
$μ$OpTime: Statically Reducing the Execution Time of Microbenchmark Suites Using Stability Metrics
by: Japke, Nils, et al.
Published: (2025)
by: Japke, Nils, et al.
Published: (2025)
Adaptable TeaStore
by: Bliudze, Simon, et al.
Published: (2024)
by: Bliudze, Simon, et al.
Published: (2024)
Similar Items
-
Distributed-memory Algorithms for Sparse Matrix Permutation, Extraction, and Assignment
by: Hassani, Elaheh, et al.
Published: (2025) -
Communication Lower Bounds and Algorithms for Sketching with Random Dense Matrices
by: Daas, Hussam Al, et al.
Published: (2026) -
Cost-Effective Big Data Orchestration Using Dagster: A Multi-Platform Approach
by: Picatto, Hernan, et al.
Published: (2024) -
Distributed-Memory Parallel Algorithms for Sparse Matrix and Sparse Tall-and-Skinny Matrix Multiplication
by: Ranawaka, Isuru, et al.
Published: (2024) -
A Unifying Framework to Enable Artificial Intelligence in High Performance Computing Workflows
by: Domke, Jens, et al.
Published: (2025)