Massively-Parallel Implementation of Inextensible Elastic Rods Using Inter-block GPU Synchronization
Fuente:
arXiv
Saved in:
| Main Authors: | Korzeniowski, Przemyslaw, Hald, Niels, Bello, Fernando |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAFI -- A Ray/Work Forwarding Infrastructure for Data Parallel Multi-Node/Multi-GPU Computing
by: Wald, Ingo, et al.
Published: (2026)
by: Wald, Ingo, et al.
Published: (2026)
GPU Volume Rendering with Hierarchical Compression Using VDB
by: Zellmann, Stefan, et al.
Published: (2025)
by: Zellmann, Stefan, et al.
Published: (2025)
Distributed Simulation of Large Multi-body Systems
by: Kale, Manas, et al.
Published: (2024)
by: Kale, Manas, et al.
Published: (2024)
Barrier-Augmented Lagrangian for GPU-based Elastodynamic Contact
by: Guo, Dewen, et al.
Published: (2024)
by: Guo, Dewen, et al.
Published: (2024)
An efficient GPU approach for designing 3D cultural heritage information systems
by: López, Luis, et al.
Published: (2025)
by: López, Luis, et al.
Published: (2025)
Exploiting ray tracing technology through OptiX to compute particle interactions with cutoff in a 3D environment on GPU
by: David, Algis, et al.
Published: (2024)
by: David, Algis, et al.
Published: (2024)
MSz: An Efficient Parallel Algorithm for Correcting Morse-Smale Segmentations in Error-Bounded Lossy Compressors
by: Li, Yuxiao, et al.
Published: (2024)
by: Li, Yuxiao, et al.
Published: (2024)
A Scalable System for Visual Analysis of Ocean Data
by: Jain, Toshit, et al.
Published: (2025)
by: Jain, Toshit, et al.
Published: (2025)
Arkade: k-Nearest Neighbor Search With Non-Euclidean Distances using GPU Ray Tracing
by: Mandarapu, Durga, et al.
Published: (2023)
by: Mandarapu, Durga, et al.
Published: (2023)
Fast Sparse Matrix Permutation for Mesh-Based Direct Solvers
by: Zarebavami, Behrooz, et al.
Published: (2026)
by: Zarebavami, Behrooz, et al.
Published: (2026)
Capsule: Efficient Player Isolation for Datacenters
by: Du, Zhouheng, et al.
Published: (2025)
by: Du, Zhouheng, et al.
Published: (2025)
GALE: Leveraging Heterogeneous Systems for Efficient Unstructured Mesh Data Analysis
by: Liu, Guoxi, et al.
Published: (2025)
by: Liu, Guoxi, et al.
Published: (2025)
Parallel Ray Tracing of Black Hole Images Using the Schwarzschild Metric
by: Naddell, Liam, et al.
Published: (2025)
by: Naddell, Liam, et al.
Published: (2025)
Scaling Point-based Differentiable Rendering for Large-scale Reconstruction
by: Zhao, Hexu, et al.
Published: (2025)
by: Zhao, Hexu, et al.
Published: (2025)
DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration
by: Lai, Shih-Yu, et al.
Published: (2026)
by: Lai, Shih-Yu, et al.
Published: (2026)
Accelerating HDC-CNN Hybrid Models Using Custom Instructions on RISC-V GPUs
by: Matsumi, Wakuto, et al.
Published: (2025)
by: Matsumi, Wakuto, et al.
Published: (2025)
Accelerating In-transit Isosurface Generation With Topology Preserving Compression
by: Li, Yanliang, et al.
Published: (2024)
by: Li, Yanliang, et al.
Published: (2024)
Efficient Parallel Implementation of the Pilot Assignment Problem in Massive MIMO Systems
by: Alqudah, Eman, et al.
Published: (2025)
by: Alqudah, Eman, et al.
Published: (2025)
Prompt-Aware Scheduling for Efficient Text-to-Image Inferencing System
by: Agarwal, Shubham, et al.
Published: (2025)
by: Agarwal, Shubham, et al.
Published: (2025)
Paris: A Decentralized Trained Open-Weight Diffusion Model
by: Jiang, Zhiying, et al.
Published: (2025)
by: Jiang, Zhiying, et al.
Published: (2025)
HPC-based Solvers of Minimisation Problems for Signal Processing
by: Cammarasana, Simone, et al.
Published: (2023)
by: Cammarasana, Simone, et al.
Published: (2023)
Parallel Track Transformers: Enabling Fast GPU Inference with Reduced Synchronization
by: Wang, Chong, et al.
Published: (2026)
by: Wang, Chong, et al.
Published: (2026)
Stimpack: An Adaptive Rendering Optimization System for Scalable Cloud Gaming
by: Heo, Jin, et al.
Published: (2024)
by: Heo, Jin, et al.
Published: (2024)
STRIELAD -- A Scalable Toolkit for Real-time Interactive Exploration of Large Atmospheric Datasets
by: Schneegans, Simon, et al.
Published: (2025)
by: Schneegans, Simon, et al.
Published: (2025)
DaVE -- A Curated Database of Visualization Examples
by: Koenen, Jens, et al.
Published: (2024)
by: Koenen, Jens, et al.
Published: (2024)
A Framework for Fine-Grained Synchronization of Dependent GPU Kernels
by: Jangda, Abhinav, et al.
Published: (2023)
by: Jangda, Abhinav, et al.
Published: (2023)
Wattchmen: Watching the Wattchers -- High Fidelity, Flexible GPU Energy Modeling
by: Tran, Brandon, et al.
Published: (2026)
by: Tran, Brandon, et al.
Published: (2026)
Scalability Evaluation of HPC Multi-GPU Training for ECG-based LLMs
by: Mileski, Dimitar, et al.
Published: (2025)
by: Mileski, Dimitar, et al.
Published: (2025)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
by: Liang, Antian, et al.
Published: (2025)
by: Liang, Antian, et al.
Published: (2025)
Dilu: Enabling GPU Resourcing-on-Demand for Serverless DL Serving via Introspective Elasticity
by: Lv, Cunchi, et al.
Published: (2025)
by: Lv, Cunchi, et al.
Published: (2025)
ElasWave: An Elastic-Native System for Scalable Hybrid-Parallel Training
by: Kang, Xueze, et al.
Published: (2025)
by: Kang, Xueze, et al.
Published: (2025)
Stream-K++: Adaptive GPU GEMM Kernel Scheduling and Selection using Bloom Filters
by: Sadasivan, Harisankar, et al.
Published: (2024)
by: Sadasivan, Harisankar, et al.
Published: (2024)
Punch Out Model Synthesis: A Stochastic Algorithm for Constraint Based Tiling Generation
by: Zzyzek, Zzyv
Published: (2025)
by: Zzyzek, Zzyv
Published: (2025)
GPU-Based Parallel Computing Methods for Medical Photoacoustic Image Reconstruction
by: Yi, Xinyao, et al.
Published: (2024)
by: Yi, Xinyao, et al.
Published: (2024)
Concurrent Scheduling of High-Level Parallel Programs on Multi-GPU Systems
by: Knorr, Fabian, et al.
Published: (2025)
by: Knorr, Fabian, et al.
Published: (2025)
A Preliminary Study on Accelerating Simulation Optimization with GPU Implementation
by: He, Jinghai, et al.
Published: (2024)
by: He, Jinghai, et al.
Published: (2024)
A Practical GPU-Accelerated Implementation of Orthogonal Matching Pursuit
by: Lubonja, Ariel, et al.
Published: (2024)
by: Lubonja, Ariel, et al.
Published: (2024)
cuFastTuckerPlus: A Stochastic Parallel Sparse FastTucker Decomposition Using GPU Tensor Cores
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
AnchorTP: Resilient LLM Inference with State-Preserving Elastic Tensor Parallelism
by: Xu, Wendong, et al.
Published: (2025)
by: Xu, Wendong, et al.
Published: (2025)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
by: Mo, Zizhao, et al.
Published: (2025)
by: Mo, Zizhao, et al.
Published: (2025)
Similar Items
-
RAFI -- A Ray/Work Forwarding Infrastructure for Data Parallel Multi-Node/Multi-GPU Computing
by: Wald, Ingo, et al.
Published: (2026) -
GPU Volume Rendering with Hierarchical Compression Using VDB
by: Zellmann, Stefan, et al.
Published: (2025) -
Distributed Simulation of Large Multi-body Systems
by: Kale, Manas, et al.
Published: (2024) -
Barrier-Augmented Lagrangian for GPU-based Elastodynamic Contact
by: Guo, Dewen, et al.
Published: (2024) -
An efficient GPU approach for designing 3D cultural heritage information systems
by: López, Luis, et al.
Published: (2025)