Assembly of FETI dual operator using CUDA
Fuente:
arXiv
Salvato in:
| Autori principali: | Homola, Jakub, Vavřík, Radim, Meca, Ondřej, Brzobohatý, Tomáš, Říha, Lubomír |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
di: Homola, Jakub, et al.
Pubblicazione: (2025)
di: Homola, Jakub, et al.
Pubblicazione: (2025)
Bandicoot: A Templated C++ Library for GPU Linear Algebra
di: Curtin, Ryan R., et al.
Pubblicazione: (2025)
di: Curtin, Ryan R., et al.
Pubblicazione: (2025)
Local Adjoints for Simultaneous Preaccumulations with Shared Inputs
di: Blühdorn, Johannes, et al.
Pubblicazione: (2024)
di: Blühdorn, Johannes, et al.
Pubblicazione: (2024)
Hybrid parallel discrete adjoints in SU2
di: Blühdorn, Johannes, et al.
Pubblicazione: (2024)
di: Blühdorn, Johannes, et al.
Pubblicazione: (2024)
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
di: Kanakagiri, Raghavendra, et al.
Pubblicazione: (2023)
di: Kanakagiri, Raghavendra, et al.
Pubblicazione: (2023)
Software Development Aspects of Integrating Linear Algebra Libraries
di: Koch, Marcel, et al.
Pubblicazione: (2025)
di: Koch, Marcel, et al.
Pubblicazione: (2025)
Tensor Decompositions for Count Data that Leverage Stochastic and Deterministic Optimization
di: Myers, Jeremy M., et al.
Pubblicazione: (2022)
di: Myers, Jeremy M., et al.
Pubblicazione: (2022)
Mapping Sparse Triangular Solves to GPUs via Fine-grained Domain Decomposition
di: Gondhalekar, Atharva, et al.
Pubblicazione: (2025)
di: Gondhalekar, Atharva, et al.
Pubblicazione: (2025)
Verifying a Sparse Matrix Algorithm Using Symbolic Execution
di: Wilton, Alexander C.
Pubblicazione: (2025)
di: Wilton, Alexander C.
Pubblicazione: (2025)
Reasoning about expression evaluation under interference
di: Hayes, Ian J., et al.
Pubblicazione: (2024)
di: Hayes, Ian J., et al.
Pubblicazione: (2024)
Data reification in a concurrent rely-guarantee algebra
di: Meinicke, Larissa A., et al.
Pubblicazione: (2024)
di: Meinicke, Larissa A., et al.
Pubblicazione: (2024)
Cascaded Prediction and Asynchronous Execution of Iterative Algorithms on Heterogeneous Platforms
di: Gao, Jianhua, et al.
Pubblicazione: (2024)
di: Gao, Jianhua, et al.
Pubblicazione: (2024)
Memory-Efficient Training with In-Place FFT Implementation
di: Ding, Xinyu, et al.
Pubblicazione: (2025)
di: Ding, Xinyu, et al.
Pubblicazione: (2025)
Reduction of the graph isomorphism problem to equality checking of $n$-variables polynomials and the algorithms that use the reduction
di: Prolubnikov, Alexander
Pubblicazione: (2015)
di: Prolubnikov, Alexander
Pubblicazione: (2015)
A Symbolic Computing Perspective on Software Systems
di: Norman, Arthur C., et al.
Pubblicazione: (2024)
di: Norman, Arthur C., et al.
Pubblicazione: (2024)
Reasoning about concurrent loops and recursion with rely-guarantee rules
di: Hayes, Ian J., et al.
Pubblicazione: (2025)
di: Hayes, Ian J., et al.
Pubblicazione: (2025)
Sampling patterns for Zernike-like bases in non-standard geometries
di: Díaz-Elbal, Sergio, et al.
Pubblicazione: (2025)
di: Díaz-Elbal, Sergio, et al.
Pubblicazione: (2025)
Randomized Approach to Matrix Completion: Applications in Recommendation Systems and Image Inpainting
di: Krajewska, Antonina, et al.
Pubblicazione: (2024)
di: Krajewska, Antonina, et al.
Pubblicazione: (2024)
Mathematical modeling of the mechanical behavior of three-layer plates with a tetrachiral honeycomb core
di: Mazaev, A. V.
Pubblicazione: (2023)
di: Mazaev, A. V.
Pubblicazione: (2023)
Reasoning about distributive laws in a concurrent refinement algebra
di: Meinicke, Larissa A., et al.
Pubblicazione: (2024)
di: Meinicke, Larissa A., et al.
Pubblicazione: (2024)
Restructuring a concurrent refinement algebra
di: Hayes, Ian J., et al.
Pubblicazione: (2024)
di: Hayes, Ian J., et al.
Pubblicazione: (2024)
Pyroclast: A Modular High-Performance Python Solver for Geodynamics
di: Ferrari, Marcel
Pubblicazione: (2026)
di: Ferrari, Marcel
Pubblicazione: (2026)
The ensmallen library for flexible numerical optimization
di: Curtin, Ryan R., et al.
Pubblicazione: (2021)
di: Curtin, Ryan R., et al.
Pubblicazione: (2021)
Trilinos: Enabling Scientific Computing Across Diverse Hardware Architectures at Scale
di: Mayr, Matthias, et al.
Pubblicazione: (2025)
di: Mayr, Matthias, et al.
Pubblicazione: (2025)
Conversational Concurrency
di: Garnock-Jones, Tony
Pubblicazione: (2024)
di: Garnock-Jones, Tony
Pubblicazione: (2024)
On the relativistic viability of multi-automaton systems: essential concepts, challenges and prospects
di: Băbeanu, Alexandru-Ionuţ
Pubblicazione: (2024)
di: Băbeanu, Alexandru-Ionuţ
Pubblicazione: (2024)
Stencil Computations on AMD and Nvidia Graphics Processors: Performance and Tuning Strategies
di: Pekkilä, Johannes, et al.
Pubblicazione: (2024)
di: Pekkilä, Johannes, et al.
Pubblicazione: (2024)
CUDABeaver: Benchmarking LLM-Based Automated CUDA Debugging
di: Li, Shiyang, et al.
Pubblicazione: (2026)
di: Li, Shiyang, et al.
Pubblicazione: (2026)
Problems from Optimization and Computational Algebra Equivalent to Hilbert's Nullstellensatz
di: Bläser, Markus, et al.
Pubblicazione: (2025)
di: Bläser, Markus, et al.
Pubblicazione: (2025)
BurTorch: Revisiting Training from First Principles by Coupling Autodiff, Math Optimization, and Systems
di: Burlachenko, Konstantin, et al.
Pubblicazione: (2025)
di: Burlachenko, Konstantin, et al.
Pubblicazione: (2025)
Formal Verification of COO to CSR Sparse Matrix Conversion (Invited Paper)
di: Appel, Andrew W.
Pubblicazione: (2025)
di: Appel, Andrew W.
Pubblicazione: (2025)
Optimization under uncertainty: understanding orders and testing programs with specifications
di: Jansson, Patrik, et al.
Pubblicazione: (2025)
di: Jansson, Patrik, et al.
Pubblicazione: (2025)
NVLang: Unified Static Typing for Actor-Based Concurrency on the BEAM
di: Guerreiro, Miguel de Oliveira
Pubblicazione: (2025)
di: Guerreiro, Miguel de Oliveira
Pubblicazione: (2025)
Flexible Quaternion Generalized Minimal Residual Method for Ill-Posed Quaternion Inverse Problems
di: Liu, Xuan, et al.
Pubblicazione: (2024)
di: Liu, Xuan, et al.
Pubblicazione: (2024)
Covariance Matrix Analysis for Optimal Portfolio Selection
di: Keith, Lim Hao Shen
Pubblicazione: (2024)
di: Keith, Lim Hao Shen
Pubblicazione: (2024)
Fast GPU Linear Algebra via Compile Time Expression Fusion
di: Curtin, Ryan R., et al.
Pubblicazione: (2026)
di: Curtin, Ryan R., et al.
Pubblicazione: (2026)
Armadillo: An Efficient Framework for Numerical Linear Algebra
di: Sanderson, Conrad, et al.
Pubblicazione: (2025)
di: Sanderson, Conrad, et al.
Pubblicazione: (2025)
Two Iterative Algorithms for Solving Systems of Simultaneous Linear Algebraic Equations with Real Matrices of Coefficients
di: Kondratiev, A. S., et al.
Pubblicazione: (2005)
di: Kondratiev, A. S., et al.
Pubblicazione: (2005)
Model reduction in Smoluchowski-type equations
di: Timokhin, Ivan V., et al.
Pubblicazione: (2020)
di: Timokhin, Ivan V., et al.
Pubblicazione: (2020)
A Two-Level Direct Solver for the Hierarchical Poincaré-Steklov Method
di: Kump, Joseph, et al.
Pubblicazione: (2025)
di: Kump, Joseph, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
di: Homola, Jakub, et al.
Pubblicazione: (2025) -
Bandicoot: A Templated C++ Library for GPU Linear Algebra
di: Curtin, Ryan R., et al.
Pubblicazione: (2025) -
Local Adjoints for Simultaneous Preaccumulations with Shared Inputs
di: Blühdorn, Johannes, et al.
Pubblicazione: (2024) -
Hybrid parallel discrete adjoints in SU2
di: Blühdorn, Johannes, et al.
Pubblicazione: (2024) -
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
di: Kanakagiri, Raghavendra, et al.
Pubblicazione: (2023)