Study on the Particle Sorting Performance for Reactor Monte Carlo Neutron Transport on Apple Unified Memory GPUs
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Liu, Changyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MCPT-Solver: An Monte Carlo Algorithm Solver Using MTJ Devices for Particle Transport Problems
von: Fu, Siqing, et al.
Veröffentlicht: (2026)
von: Fu, Siqing, et al.
Veröffentlicht: (2026)
ADS-IMC: Accelerating Data Sorting with In-Memory Computation
von: Dhakad, Narendra Singh, et al.
Veröffentlicht: (2026)
von: Dhakad, Narendra Singh, et al.
Veröffentlicht: (2026)
Apple vs. Oranges: Evaluating the Apple Silicon M-Series SoCs for HPC Performance and Efficiency
von: Hübner, Paul, et al.
Veröffentlicht: (2025)
von: Hübner, Paul, et al.
Veröffentlicht: (2025)
Bandwidth-Effective DRAM Cache for GPUs with Storage-Class Memory
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024)
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024)
Privacy-Preserving Performance Profiling of In-The-Wild GPUs
von: McDougall, Ian, et al.
Veröffentlicht: (2025)
von: McDougall, Ian, et al.
Veröffentlicht: (2025)
Hidden Risks of Unmonitored GPUs in Intelligent Transportation Systems
von: Puspa, Sefatun-Noor, et al.
Veröffentlicht: (2026)
von: Puspa, Sefatun-Noor, et al.
Veröffentlicht: (2026)
Control Flow Management in Modern GPUs
von: Shoushtary, Mojtaba Abaie, et al.
Veröffentlicht: (2024)
von: Shoushtary, Mojtaba Abaie, et al.
Veröffentlicht: (2024)
Memory-Guided Unified Hardware Accelerator for Mixed-Precision Scientific Computing
von: Wang, Chuanzhen, et al.
Veröffentlicht: (2026)
von: Wang, Chuanzhen, et al.
Veröffentlicht: (2026)
A Systematic Characterization of LLM Inference on GPUs
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
Fleet: Hierarchical Task-based Abstraction for Megakernels on Multi-Die GPUs
von: Chowdhary, Sangeeta, et al.
Veröffentlicht: (2026)
von: Chowdhary, Sangeeta, et al.
Veröffentlicht: (2026)
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
von: Seo, Minseok, et al.
Veröffentlicht: (2024)
von: Seo, Minseok, et al.
Veröffentlicht: (2024)
In-Memory Sorting-Searching with Cayley Tree
von: Paul, Subrata, et al.
Veröffentlicht: (2025)
von: Paul, Subrata, et al.
Veröffentlicht: (2025)
Virgo: Cluster-level Matrix Unit Integration in GPUs for Scalability and Energy Efficiency
von: Kim, Hansung, et al.
Veröffentlicht: (2024)
von: Kim, Hansung, et al.
Veröffentlicht: (2024)
CarbonSet: A Dataset to Analyze Trends and Benchmark the Sustainability of CPUs and GPUs
von: Hu, Jiajun, et al.
Veröffentlicht: (2025)
von: Hu, Jiajun, et al.
Veröffentlicht: (2025)
Dissecting Conditional Branch Predictors of Apple Firestorm and Qualcomm Oryon for Software Optimization and Architectural Analysis
von: Chen, Jiajie, et al.
Veröffentlicht: (2024)
von: Chen, Jiajie, et al.
Veröffentlicht: (2024)
Power- and Area-Efficient Unary Sorting Architecture Using FSM-Based Unary Number Generator
von: Jalilvand, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Jalilvand, Amir Hossein, et al.
Veröffentlicht: (2025)
Hermes: A Unified High-Performance NTT Architecture with Hybrid Dataflow
von: Gu, Hang, et al.
Veröffentlicht: (2026)
von: Gu, Hang, et al.
Veröffentlicht: (2026)
Optimizing and Exploring System Performance in Compact Processing-in-Memory-based Chips
von: Chen, Peilin, et al.
Veröffentlicht: (2025)
von: Chen, Peilin, et al.
Veröffentlicht: (2025)
Overmind NSA: A Unified Neuro-Symbolic Computing Architecture with Approximate Nonlinear Activations and Preemptive Memory Bypass
von: Wang, Weilun, et al.
Veröffentlicht: (2026)
von: Wang, Weilun, et al.
Veröffentlicht: (2026)
Performance Characterizations and Usage Guidelines of Samsung CXL Memory Module Hybrid Prototype
von: Zeng, Jianping, et al.
Veröffentlicht: (2025)
von: Zeng, Jianping, et al.
Veröffentlicht: (2025)
AraOS: Analyzing the Impact of Virtual Memory Management on Vector Unit Performance
von: Perotti, Matteo, et al.
Veröffentlicht: (2025)
von: Perotti, Matteo, et al.
Veröffentlicht: (2025)
Per-Bank Memory Bandwidth Regulation for Predictable and Performant Real-Time System
von: Sullivan, Connor Rudy, et al.
Veröffentlicht: (2026)
von: Sullivan, Connor Rudy, et al.
Veröffentlicht: (2026)
Reimagining Memory Access for LLM Inference: Compression-Aware Memory Controller Design
von: Xie, Rui, et al.
Veröffentlicht: (2025)
von: Xie, Rui, et al.
Veröffentlicht: (2025)
UFO-MAC: A Unified Framework for Optimization of High-Performance Multipliers and Multiply-Accumulators
von: Zuo, Dongsheng, et al.
Veröffentlicht: (2024)
von: Zuo, Dongsheng, et al.
Veröffentlicht: (2024)
A Benchmarking Platform for DDR4 Memory Performance in Data-Center-Class FPGAs
von: Galimberti, Andrea, et al.
Veröffentlicht: (2025)
von: Galimberti, Andrea, et al.
Veröffentlicht: (2025)
Range, Not Precision: Block-Floating-Point Half-Precision FFT and SAR Imaging on Apple Silicon
von: Bergach, Mohamed Amine
Veröffentlicht: (2026)
von: Bergach, Mohamed Amine
Veröffentlicht: (2026)
The Future of Memory: Limits and Opportunities
von: Dayo, Samuel, et al.
Veröffentlicht: (2025)
von: Dayo, Samuel, et al.
Veröffentlicht: (2025)
SynDCIM: A Performance-Aware Digital Computing-in-Memory Compiler with Multi-Spec-Oriented Subcircuit Synthesis
von: Shao, Kunming, et al.
Veröffentlicht: (2024)
von: Shao, Kunming, et al.
Veröffentlicht: (2024)
GPIR: Enabling Practical Private Information Retrieval with GPUs
von: Ji, Hyesung, et al.
Veröffentlicht: (2026)
von: Ji, Hyesung, et al.
Veröffentlicht: (2026)
The Case for Replication-Aware Memory-Error Protection in Disaggregated Memory
von: Volos, Haris
Veröffentlicht: (2023)
von: Volos, Haris
Veröffentlicht: (2023)
In-Memory ADC-Based Nonlinear Activation Quantization for Efficient In-Memory Computing
von: Dong, Shuai, et al.
Veröffentlicht: (2026)
von: Dong, Shuai, et al.
Veröffentlicht: (2026)
How Much Progress Has There Been in NVIDIA Datacenter GPUs?
von: Del Sozzo, Emanuele, et al.
Veröffentlicht: (2026)
von: Del Sozzo, Emanuele, et al.
Veröffentlicht: (2026)
Asynchronous Memory Access Unit: Exploiting Massive Parallelism for Far Memory Access
von: Wang, Luming, et al.
Veröffentlicht: (2024)
von: Wang, Luming, et al.
Veröffentlicht: (2024)
An Event-Driven Spiking Compute-In-Memory Macro based on SOT-MRAM
von: Yu, Deyang, et al.
Veröffentlicht: (2025)
von: Yu, Deyang, et al.
Veröffentlicht: (2025)
IMMSched: Interruptible Multi-DNN Scheduling via Parallel Multi-Particle Optimizing Subgraph Isomorphism
von: Zhao, Boran, et al.
Veröffentlicht: (2026)
von: Zhao, Boran, et al.
Veröffentlicht: (2026)
PipeWeave: Synergizing Analytical and Learning Models for Unified GPU Performance Prediction
von: Zhang, Kaixuan, et al.
Veröffentlicht: (2026)
von: Zhang, Kaixuan, et al.
Veröffentlicht: (2026)
PUMA: Efficient and Low-Cost Memory Allocation and Alignment Support for Processing-Using-Memory Architectures
von: Oliveira, Geraldo F., et al.
Veröffentlicht: (2024)
von: Oliveira, Geraldo F., et al.
Veröffentlicht: (2024)
CINM (Cinnamon): A Compilation Infrastructure for Heterogeneous Compute In-Memory and Compute Near-Memory Paradigms
von: Khan, Asif Ali, et al.
Veröffentlicht: (2022)
von: Khan, Asif Ali, et al.
Veröffentlicht: (2022)
PIM-malloc: A Fast and Scalable Dynamic Memory Allocator for Processing-In-Memory (PIM) Architectures
von: Lee, Dongjae, et al.
Veröffentlicht: (2025)
von: Lee, Dongjae, et al.
Veröffentlicht: (2025)
CAMASim: A Comprehensive Simulation Framework for Content-Addressable Memory based Accelerators
von: Li, Mengyuan, et al.
Veröffentlicht: (2024)
von: Li, Mengyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MCPT-Solver: An Monte Carlo Algorithm Solver Using MTJ Devices for Particle Transport Problems
von: Fu, Siqing, et al.
Veröffentlicht: (2026) -
ADS-IMC: Accelerating Data Sorting with In-Memory Computation
von: Dhakad, Narendra Singh, et al.
Veröffentlicht: (2026) -
Apple vs. Oranges: Evaluating the Apple Silicon M-Series SoCs for HPC Performance and Efficiency
von: Hübner, Paul, et al.
Veröffentlicht: (2025) -
Bandwidth-Effective DRAM Cache for GPUs with Storage-Class Memory
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024) -
Privacy-Preserving Performance Profiling of In-The-Wild GPUs
von: McDougall, Ian, et al.
Veröffentlicht: (2025)