Matryoshka: Optimization of Dynamic Diverse Quantum Chemistry Systems via Elastic Parallelism Transformation
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Tuowei, Li, Kun, Bai, Donglin, Ju, Fusong, Xia, Leo, Cao, Ting, Ren, Ju, Zhang, Yaoxue, Yang, Mao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Optimal Parallel Scheduling under Concave Speedup Functions
di: Li, Chengzhang, et al.
Pubblicazione: (2025)
di: Li, Chengzhang, et al.
Pubblicazione: (2025)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
di: Ather, Hammad, et al.
Pubblicazione: (2024)
di: Ather, Hammad, et al.
Pubblicazione: (2024)
Cloud Resource Allocation with Convex Optimization
di: Boghani, Shayan, et al.
Pubblicazione: (2025)
di: Boghani, Shayan, et al.
Pubblicazione: (2025)
On Orchestrating Parallel Broadcasts for Distributed Ledgers
di: Sheng, Peiyao, et al.
Pubblicazione: (2024)
di: Sheng, Peiyao, et al.
Pubblicazione: (2024)
Automated Programmatic Performance Analysis of Parallel Programs
di: Cankur, Onur, et al.
Pubblicazione: (2024)
di: Cankur, Onur, et al.
Pubblicazione: (2024)
Shifting the Sweet Spot: High-Performance Matrix-Free Method for High-Order Elasticity
di: Chang, Dali, et al.
Pubblicazione: (2026)
di: Chang, Dali, et al.
Pubblicazione: (2026)
Recorder: Comprehensive Parallel I/O Tracing and Analysis
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
ParaLog: Consistent Host-side Logging for Parallel Checkpoints
di: Chien, Steven W. D., et al.
Pubblicazione: (2024)
di: Chien, Steven W. D., et al.
Pubblicazione: (2024)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
di: Lacey, Dane C., et al.
Pubblicazione: (2024)
di: Lacey, Dane C., et al.
Pubblicazione: (2024)
Fine-Grained Energy Prediction For Parallellized LLM Inference With PIE-P
di: Dutt, Anurag, et al.
Pubblicazione: (2025)
di: Dutt, Anurag, et al.
Pubblicazione: (2025)
Automated Calibration of Parallel and Distributed Computing Simulators: A Case Study
di: McDonald, Jesse, et al.
Pubblicazione: (2024)
di: McDonald, Jesse, et al.
Pubblicazione: (2024)
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
CARAT: Client-Side Adaptive RPC and Cache Co-Tuning for Parallel File Systems
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
HeteGen: Heterogeneous Parallel Inference for Large Language Models on Resource-Constrained Devices
di: Zhao, Xuanlei, et al.
Pubblicazione: (2024)
di: Zhao, Xuanlei, et al.
Pubblicazione: (2024)
AcceleratedKernels.jl: Cross-Architecture Parallel Algorithms from a Unified, Transpiled Codebase
di: Nicusan, Andrei-Leonard, et al.
Pubblicazione: (2025)
di: Nicusan, Andrei-Leonard, et al.
Pubblicazione: (2025)
DIAL: Decentralized I/O AutoTuning via Learned Client-side Local Metrics for Parallel File System
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
PASTA: A Modular Program Analysis Tool Framework for Accelerators
di: Lin, Mao, et al.
Pubblicazione: (2026)
di: Lin, Mao, et al.
Pubblicazione: (2026)
LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing
di: Xia, Yuning, et al.
Pubblicazione: (2026)
di: Xia, Yuning, et al.
Pubblicazione: (2026)
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
di: Sun, Tingyang, et al.
Pubblicazione: (2026)
di: Sun, Tingyang, et al.
Pubblicazione: (2026)
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
di: Lin, Mao, et al.
Pubblicazione: (2026)
di: Lin, Mao, et al.
Pubblicazione: (2026)
DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers
di: Maurya, Avinash, et al.
Pubblicazione: (2026)
di: Maurya, Avinash, et al.
Pubblicazione: (2026)
An Auto-tuning Method for Run-time Data Transformation for Sparse Matrix-Vector Multiplication
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
Inductive Loop Analysis for Practical HPC Application Optimization
di: Schaad, Philipp, et al.
Pubblicazione: (2025)
di: Schaad, Philipp, et al.
Pubblicazione: (2025)
Kino-PAX: Highly Parallel Kinodynamic Sampling-based Planner
di: Perrault, Nicolas, et al.
Pubblicazione: (2024)
di: Perrault, Nicolas, et al.
Pubblicazione: (2024)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
di: Mao, Ying, et al.
Pubblicazione: (2020)
di: Mao, Ying, et al.
Pubblicazione: (2020)
Optimizations on Graph-Level for Domain Specific Computations in Julia and Application to QED
di: Reinhard, Anton, et al.
Pubblicazione: (2025)
di: Reinhard, Anton, et al.
Pubblicazione: (2025)
Data-Driven Analysis to Understand GPU Hardware Resource Usage of Optimizations
di: Islam, Tanzima Z., et al.
Pubblicazione: (2024)
di: Islam, Tanzima Z., et al.
Pubblicazione: (2024)
BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems
di: Wang, Yuxin, et al.
Pubblicazione: (2024)
di: Wang, Yuxin, et al.
Pubblicazione: (2024)
Performance Optimization in Stream Processing Systems: Experiment-Driven Configuration Tuning for Kafka Streams
di: Chen, David, et al.
Pubblicazione: (2026)
di: Chen, David, et al.
Pubblicazione: (2026)
Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2025)
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2025)
Towards a Peer-to-Peer Data Distribution Layer for Efficient and Collaborative Resource Optimization of Distributed Dataflow Applications
di: Scheinert, Dominik, et al.
Pubblicazione: (2023)
di: Scheinert, Dominik, et al.
Pubblicazione: (2023)
Learning-Augmented Performance Model for Tensor Product Factorization in High-Order FEM
di: Ren, Xuanzhengbo, et al.
Pubblicazione: (2026)
di: Ren, Xuanzhengbo, et al.
Pubblicazione: (2026)
Opt4GPTQ: Co-Optimizing Memory and Computation for 4-bit GPTQ Quantized LLM Inference on Heterogeneous Platforms
di: Zhang, Yaozheng, et al.
Pubblicazione: (2025)
di: Zhang, Yaozheng, et al.
Pubblicazione: (2025)
GigaAPI for GPU Parallelization
di: Suvarna, M., et al.
Pubblicazione: (2025)
di: Suvarna, M., et al.
Pubblicazione: (2025)
Evaluation of Quantum and Hybrid Solvers for Combinatorial Optimization
di: Bertuzzi, Amedeo, et al.
Pubblicazione: (2024)
di: Bertuzzi, Amedeo, et al.
Pubblicazione: (2024)
Parallelizing a modern GPU simulator
di: Huerta, Rodrigo, et al.
Pubblicazione: (2025)
di: Huerta, Rodrigo, et al.
Pubblicazione: (2025)
CGSim: A Simulation Framework for Large Scale Distributed Computing Environment
di: Vatsavai, Sairam Sri, et al.
Pubblicazione: (2025)
di: Vatsavai, Sairam Sri, et al.
Pubblicazione: (2025)
Comparing Parallel Functional Array Languages: Programming and Performance
di: van Balen, David, et al.
Pubblicazione: (2025)
di: van Balen, David, et al.
Pubblicazione: (2025)
Can Large Language Models Predict Parallel Code Performance?
di: Bolet, Gregory, et al.
Pubblicazione: (2025)
di: Bolet, Gregory, et al.
Pubblicazione: (2025)
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
di: Barron, Ryan, et al.
Pubblicazione: (2024)
di: Barron, Ryan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Optimal Parallel Scheduling under Concave Speedup Functions
di: Li, Chengzhang, et al.
Pubblicazione: (2025) -
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
di: Ather, Hammad, et al.
Pubblicazione: (2024) -
Cloud Resource Allocation with Convex Optimization
di: Boghani, Shayan, et al.
Pubblicazione: (2025) -
On Orchestrating Parallel Broadcasts for Distributed Ledgers
di: Sheng, Peiyao, et al.
Pubblicazione: (2024) -
Automated Programmatic Performance Analysis of Parallel Programs
di: Cankur, Onur, et al.
Pubblicazione: (2024)