Black-Scholes Option Pricing on Intel CPUs and GPUs: Implementation on SYCL and Optimization Techniques
Fuente:
arXiv
Salvato in:
| Autori principali: | Panova, Elena, Volokitin, Valentin, Gorshkov, Anton, Meyerov, Iosif |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Vectorization of Gradient Boosting of Decision Trees Prediction in the CatBoost Library for RISC-V Processors
di: Kozinov, Evgeny, et al.
Pubblicazione: (2024)
di: Kozinov, Evgeny, et al.
Pubblicazione: (2024)
High-Performance Implementation of the Optimized Event Generator for Strong-Field QED Plasma Simulations
di: Panova, Elena, et al.
Pubblicazione: (2024)
di: Panova, Elena, et al.
Pubblicazione: (2024)
Performance optimization of BLAS algorithms with band matrices for RISC-V processors
di: Pirova, Anna, et al.
Pubblicazione: (2025)
di: Pirova, Anna, et al.
Pubblicazione: (2025)
CloverLeaf on Intel Multi-Core CPUs: A Case Study in Write-Allocate Evasion
di: Laukemann, Jan, et al.
Pubblicazione: (2023)
di: Laukemann, Jan, et al.
Pubblicazione: (2023)
Xabclib:A Fully Auto-tuned Sparse Iterative Solver
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
Automated MPI-X code generation for scalable finite-difference solvers
di: Bisbas, George, et al.
Pubblicazione: (2023)
di: Bisbas, George, et al.
Pubblicazione: (2023)
On the energy efficiency of sparse matrix computations on multi-GPU clusters
di: Bernaschi, Massimo, et al.
Pubblicazione: (2025)
di: Bernaschi, Massimo, et al.
Pubblicazione: (2025)
NApy: Efficient Statistics in Python for Large-Scale Heterogeneous Data with Enhanced Support for Missing Data
di: Woller, Fabian, et al.
Pubblicazione: (2025)
di: Woller, Fabian, et al.
Pubblicazione: (2025)
A Communication Avoiding and Reducing Algorithm for Symmetric Eigenproblem for Very Small Matrices
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
Beating vDSP: A 138 GFLOPS Radix-8 Stockham FFT on Apple Silicon via Two-Tier Register-Threadgroup Memory Decomposition
di: Bergach, Mohamed Amine
Pubblicazione: (2026)
di: Bergach, Mohamed Amine
Pubblicazione: (2026)
Performance measurements of modern Fortran MPI applications with Score-P
di: Corbin, Gregor
Pubblicazione: (2025)
di: Corbin, Gregor
Pubblicazione: (2025)
Accelerating High-Order Finite Element Simulations at Extreme Scale with FP64 Tensor Cores
di: Tu, Jiqun, et al.
Pubblicazione: (2026)
di: Tu, Jiqun, et al.
Pubblicazione: (2026)
SYCL compute kernels for ExaHyPE
di: Loi, Chung Ming, et al.
Pubblicazione: (2023)
di: Loi, Chung Ming, et al.
Pubblicazione: (2023)
pyGinkgo: A Sparse Linear Algebra Operator Framework for Python
di: Tuteja, Keshvi, et al.
Pubblicazione: (2025)
di: Tuteja, Keshvi, et al.
Pubblicazione: (2025)
Performant Automatic BLAS Offloading on Unified Memory Architecture with OpenMP First-Touch Style Data Movement
di: Li, Junjie
Pubblicazione: (2024)
di: Li, Junjie
Pubblicazione: (2024)
Improved vectorization of OpenCV algorithms for RISC-V CPUs
di: Volokitin, V. D., et al.
Pubblicazione: (2023)
di: Volokitin, V. D., et al.
Pubblicazione: (2023)
A Comparison of the Performance of the Molecular Dynamics Simulation Package GROMACS Implemented in the SYCL and CUDA Programming Models
di: Apanasevich, L., et al.
Pubblicazione: (2024)
di: Apanasevich, L., et al.
Pubblicazione: (2024)
Code Generation for Near-Roofline Finite Element Actions on GPUs from Symbolic Variational Forms
di: Kulkarni, Kaushik, et al.
Pubblicazione: (2025)
di: Kulkarni, Kaushik, et al.
Pubblicazione: (2025)
GoldbachGPU: An Open Source GPU-Accelerated Framework for Verification of Goldbach's Conjecture
di: Llorente-Saguer, Isaac
Pubblicazione: (2026)
di: Llorente-Saguer, Isaac
Pubblicazione: (2026)
Hierarchical Recursive Precision for Accelerating Symmetric Linear Solves on MXUs
di: Carrica, Vicki, et al.
Pubblicazione: (2026)
di: Carrica, Vicki, et al.
Pubblicazione: (2026)
Performance Analysis of HPC applications on the Aurora Supercomputer: Exploring the Impact of HBM-Enabled Intel Xeon Max CPUs
di: Ibeid, Huda, et al.
Pubblicazione: (2025)
di: Ibeid, Huda, et al.
Pubblicazione: (2025)
MPI Implementation Profiling for Better Application Performance
di: Shipley, Riley, et al.
Pubblicazione: (2024)
di: Shipley, Riley, et al.
Pubblicazione: (2024)
Should I Run My Cloud Benchmark on Black Friday?
di: Henning, Sören, et al.
Pubblicazione: (2025)
di: Henning, Sören, et al.
Pubblicazione: (2025)
Comparing the Performance of Heterogeneous Conjugate Gradient and Cholesky Solvers on Various Hardware Using SYCL
di: Thüring, Tim, et al.
Pubblicazione: (2026)
di: Thüring, Tim, et al.
Pubblicazione: (2026)
Analyzing the Performance Portability of SYCL across CPUs, GPUs, and Hybrid Systems with SW Sequence Alignment
di: Costanzo, Manuel, et al.
Pubblicazione: (2024)
di: Costanzo, Manuel, et al.
Pubblicazione: (2024)
A dynamic parallel method for performance optimization on hybrid CPUs
di: Yu, Luo, et al.
Pubblicazione: (2024)
di: Yu, Luo, et al.
Pubblicazione: (2024)
Accelerating Bidiagonalization of Banded Matrices through Memory-Aware Bulge-Chasing on GPUs
di: Ringoot, Evelyne, et al.
Pubblicazione: (2025)
di: Ringoot, Evelyne, et al.
Pubblicazione: (2025)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
di: Nichols, Daniel, et al.
Pubblicazione: (2025)
di: Nichols, Daniel, et al.
Pubblicazione: (2025)
Optimizing OpenFaaS on Kubernetes: Comparative Analysis of Language Runtimes and Cluster Distributions
di: Ataie, Ehsan, et al.
Pubblicazione: (2026)
di: Ataie, Ehsan, et al.
Pubblicazione: (2026)
Comparison of Vectorization Capabilities of Different Compilers for X86 and ARM CPUs
di: Sakib, Nazmus, et al.
Pubblicazione: (2025)
di: Sakib, Nazmus, et al.
Pubblicazione: (2025)
Towards High-Performance and Portable Molecular Docking on CPUs through Vectorization
di: Accordi, Gianmarco, et al.
Pubblicazione: (2025)
di: Accordi, Gianmarco, et al.
Pubblicazione: (2025)
Microarchitectural comparison and in-core modeling of state-of-the-art CPUs: Grace, Sapphire Rapids, and Genoa
di: Laukemann, Jan, et al.
Pubblicazione: (2024)
di: Laukemann, Jan, et al.
Pubblicazione: (2024)
THEAS: Efficient Power Management in Multi-Core CPUs via Cache-Aware Resource Scheduling
di: Muhammad, Said, et al.
Pubblicazione: (2025)
di: Muhammad, Said, et al.
Pubblicazione: (2025)
How to Rent GPUs on a Budget
di: Li, Zhouzi, et al.
Pubblicazione: (2024)
di: Li, Zhouzi, et al.
Pubblicazione: (2024)
Easy Acceleration with Distributed Arrays
di: Kepner, Jeremy, et al.
Pubblicazione: (2025)
di: Kepner, Jeremy, et al.
Pubblicazione: (2025)
Fast and energy-efficient derivatives risk analysis: Streaming option Greeks on Xilinx and Intel FPGAs
di: Klaisoongnoen, Mark, et al.
Pubblicazione: (2022)
di: Klaisoongnoen, Mark, et al.
Pubblicazione: (2022)
AI-NativeBench: An Open-Source White-Box Agentic Benchmark Suite for AI-Native Systems
di: Wang, Zirui, et al.
Pubblicazione: (2026)
di: Wang, Zirui, et al.
Pubblicazione: (2026)
High-level Stream Processing: A Complementary Analysis of Fault Recovery
di: Vogel, Adriano, et al.
Pubblicazione: (2024)
di: Vogel, Adriano, et al.
Pubblicazione: (2024)
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2024)
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2024)
When Should I Run My Application Benchmark?: Studying Cloud Performance Variability for the Case of Stream Processing Applications
di: Henning, Sören, et al.
Pubblicazione: (2025)
di: Henning, Sören, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Vectorization of Gradient Boosting of Decision Trees Prediction in the CatBoost Library for RISC-V Processors
di: Kozinov, Evgeny, et al.
Pubblicazione: (2024) -
High-Performance Implementation of the Optimized Event Generator for Strong-Field QED Plasma Simulations
di: Panova, Elena, et al.
Pubblicazione: (2024) -
Performance optimization of BLAS algorithms with band matrices for RISC-V processors
di: Pirova, Anna, et al.
Pubblicazione: (2025) -
CloverLeaf on Intel Multi-Core CPUs: A Case Study in Write-Allocate Evasion
di: Laukemann, Jan, et al.
Pubblicazione: (2023) -
Xabclib:A Fully Auto-tuned Sparse Iterative Solver
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)