Optimization of a Radiofrequency Ablation FEM Application Using Parallel Sparse Solvers
Fuente:
arXiv
Salvato in:
| Autori principali: | Miletto, Marcelo Cogo, Schepke, Claudio, Schnorr, Lucas Mello |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Parallelization Strategies for Dense LLM Deployment: Navigating Through Application-Specific Tradeoffs and Bottlenecks
di: Topcu, Burak, et al.
Pubblicazione: (2026)
di: Topcu, Burak, et al.
Pubblicazione: (2026)
Temporal Load Imbalance on Ondes3D Seismic Simulator for Different Multicore Architectures
di: Solórzano, Ana Luisa Veroneze, et al.
Pubblicazione: (2024)
di: Solórzano, Ana Luisa Veroneze, et al.
Pubblicazione: (2024)
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
di: Chen, Mu-Chi, et al.
Pubblicazione: (2025)
di: Chen, Mu-Chi, et al.
Pubblicazione: (2025)
ParaQAOA: Efficient Parallel Divide-and-Conquer QAOA for Large-Scale Max-Cut Problems Beyond 10,000 Vertices
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2026)
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2026)
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
di: Lysenstøen, Christian
Pubblicazione: (2026)
di: Lysenstøen, Christian
Pubblicazione: (2026)
Vectorized Adaptive Histograms for Sparse Oblique Forests
di: Lubonja, Ariel, et al.
Pubblicazione: (2026)
di: Lubonja, Ariel, et al.
Pubblicazione: (2026)
Efficient Construction of Large Search Spaces for Auto-Tuning
di: Willemsen, Floris-Jan, et al.
Pubblicazione: (2025)
di: Willemsen, Floris-Jan, et al.
Pubblicazione: (2025)
Learning-Augmented Performance Model for Tensor Product Factorization in High-Order FEM
di: Ren, Xuanzhengbo, et al.
Pubblicazione: (2026)
di: Ren, Xuanzhengbo, et al.
Pubblicazione: (2026)
Comparing the Performance of Heterogeneous Conjugate Gradient and Cholesky Solvers on Various Hardware Using SYCL
di: Thüring, Tim, et al.
Pubblicazione: (2026)
di: Thüring, Tim, et al.
Pubblicazione: (2026)
eScope: A Fine-Grained Power Prediction Mechanism for Mobile Applications
di: Mukherjee, Dipayan, et al.
Pubblicazione: (2024)
di: Mukherjee, Dipayan, et al.
Pubblicazione: (2024)
Cross-Platform Fused MoE Dispatch in Triton: Portable Expert Routing Without CUDA
di: Mitra, Subhadip
Pubblicazione: (2026)
di: Mitra, Subhadip
Pubblicazione: (2026)
Libra: Unleashing GPU Heterogeneity for High-Performance Sparse Matrix Multiplication
di: Shi, Jinliang, et al.
Pubblicazione: (2025)
di: Shi, Jinliang, et al.
Pubblicazione: (2025)
Scalability Evaluation of HPC Multi-GPU Training for ECG-based LLMs
di: Mileski, Dimitar, et al.
Pubblicazione: (2025)
di: Mileski, Dimitar, et al.
Pubblicazione: (2025)
Matryoshka: Optimization of Dynamic Diverse Quantum Chemistry Systems via Elastic Parallelism Transformation
di: Wang, Tuowei, et al.
Pubblicazione: (2024)
di: Wang, Tuowei, et al.
Pubblicazione: (2024)
Intelligent Cloud Orchestration: A Hybrid Predictive and Heuristic Framework for Cost Optimization
di: Nagoriya, Heet, et al.
Pubblicazione: (2026)
di: Nagoriya, Heet, et al.
Pubblicazione: (2026)
Xabclib:A Fully Auto-tuned Sparse Iterative Solver
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
di: Ather, Hammad, et al.
Pubblicazione: (2024)
di: Ather, Hammad, et al.
Pubblicazione: (2024)
Is Sparse Matrix Reordering Effective for Sparse Matrix-Vector Multiplication?
di: Asudeh, Omid, et al.
Pubblicazione: (2025)
di: Asudeh, Omid, et al.
Pubblicazione: (2025)
Inductive Loop Analysis for Practical HPC Application Optimization
di: Schaad, Philipp, et al.
Pubblicazione: (2025)
di: Schaad, Philipp, et al.
Pubblicazione: (2025)
On Orchestrating Parallel Broadcasts for Distributed Ledgers
di: Sheng, Peiyao, et al.
Pubblicazione: (2024)
di: Sheng, Peiyao, et al.
Pubblicazione: (2024)
Automated Programmatic Performance Analysis of Parallel Programs
di: Cankur, Onur, et al.
Pubblicazione: (2024)
di: Cankur, Onur, et al.
Pubblicazione: (2024)
Evaluation of Quantum and Hybrid Solvers for Combinatorial Optimization
di: Bertuzzi, Amedeo, et al.
Pubblicazione: (2024)
di: Bertuzzi, Amedeo, et al.
Pubblicazione: (2024)
Optimizations on Graph-Level for Domain Specific Computations in Julia and Application to QED
di: Reinhard, Anton, et al.
Pubblicazione: (2025)
di: Reinhard, Anton, et al.
Pubblicazione: (2025)
Optimal Parallel Scheduling under Concave Speedup Functions
di: Li, Chengzhang, et al.
Pubblicazione: (2025)
di: Li, Chengzhang, et al.
Pubblicazione: (2025)
Recorder: Comprehensive Parallel I/O Tracing and Analysis
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Towards Portability at Scale: A Cross-Architecture Performance Evaluation of a GPU-enabled Shallow Water Solver
di: Villalobos, Johansell, et al.
Pubblicazione: (2025)
di: Villalobos, Johansell, et al.
Pubblicazione: (2025)
ParaLog: Consistent Host-side Logging for Parallel Checkpoints
di: Chien, Steven W. D., et al.
Pubblicazione: (2024)
di: Chien, Steven W. D., et al.
Pubblicazione: (2024)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
di: Lacey, Dane C., et al.
Pubblicazione: (2024)
di: Lacey, Dane C., et al.
Pubblicazione: (2024)
Fine-Grained Energy Prediction For Parallellized LLM Inference With PIE-P
di: Dutt, Anurag, et al.
Pubblicazione: (2025)
di: Dutt, Anurag, et al.
Pubblicazione: (2025)
Automated Calibration of Parallel and Distributed Computing Simulators: A Case Study
di: McDonald, Jesse, et al.
Pubblicazione: (2024)
di: McDonald, Jesse, et al.
Pubblicazione: (2024)
Comparative Analysis of Large Language Model Inference Serving Systems: A Performance Study of vLLM and HuggingFace TGI
di: Kolluru, Saicharan
Pubblicazione: (2025)
di: Kolluru, Saicharan
Pubblicazione: (2025)
Staging Blocked Evaluation over Structured Sparse Matrices
di: Das, Pratyush, et al.
Pubblicazione: (2024)
di: Das, Pratyush, et al.
Pubblicazione: (2024)
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
Efficient Fault Localization in a Cloud Stack Using End-to-End Application Service Topology
di: Mathews, Dhanya R, et al.
Pubblicazione: (2025)
di: Mathews, Dhanya R, et al.
Pubblicazione: (2025)
FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines
di: He, Jiaao, et al.
Pubblicazione: (2024)
di: He, Jiaao, et al.
Pubblicazione: (2024)
Comprehensive Plugin-Based Monitoring of Nexflow Workflow Executions
di: Kharma, Sami, et al.
Pubblicazione: (2026)
di: Kharma, Sami, et al.
Pubblicazione: (2026)
CARAT: Client-Side Adaptive RPC and Cache Co-Tuning for Parallel File Systems
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
HeteGen: Heterogeneous Parallel Inference for Large Language Models on Resource-Constrained Devices
di: Zhao, Xuanlei, et al.
Pubblicazione: (2024)
di: Zhao, Xuanlei, et al.
Pubblicazione: (2024)
Towards a Peer-to-Peer Data Distribution Layer for Efficient and Collaborative Resource Optimization of Distributed Dataflow Applications
di: Scheinert, Dominik, et al.
Pubblicazione: (2023)
di: Scheinert, Dominik, et al.
Pubblicazione: (2023)
SHIRO: Near-Optimal Communication Strategies for Distributed Sparse Matrix Multiplication
di: Zhuang, Chen, et al.
Pubblicazione: (2025)
di: Zhuang, Chen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Parallelization Strategies for Dense LLM Deployment: Navigating Through Application-Specific Tradeoffs and Bottlenecks
di: Topcu, Burak, et al.
Pubblicazione: (2026) -
Temporal Load Imbalance on Ondes3D Seismic Simulator for Different Multicore Architectures
di: Solórzano, Ana Luisa Veroneze, et al.
Pubblicazione: (2024) -
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
di: Chen, Mu-Chi, et al.
Pubblicazione: (2025) -
ParaQAOA: Efficient Parallel Divide-and-Conquer QAOA for Large-Scale Max-Cut Problems Beyond 10,000 Vertices
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2026) -
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
di: Lysenstøen, Christian
Pubblicazione: (2026)