Fast Switching Serial and Parallel Paradigms of SNN Inference on Multi-core Heterogeneous Neuromorphic Platform SpiNNaker2
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Jiaxin, Vogginger, Bernhard, Kelber, Florian, Gonzalez, Hector, Knobloch, Klaus, Mayr, Christian Georg |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An End-to-End DNN Inference Framework for the SpiNNaker2 Neuromorphic MPSoC
von: Jobst, Matthias, et al.
Veröffentlicht: (2025)
von: Jobst, Matthias, et al.
Veröffentlicht: (2025)
Hardware-Aware Fine-Tuning of Spiking Q-Networks on the SpiNNaker2 Neuromorphic Platform
von: Arfa, Sirine, et al.
Veröffentlicht: (2025)
von: Arfa, Sirine, et al.
Veröffentlicht: (2025)
Language Modeling on a SpiNNaker 2 Neuromorphic Chip
von: Nazeer, Khaleelulla Khan, et al.
Veröffentlicht: (2023)
von: Nazeer, Khaleelulla Khan, et al.
Veröffentlicht: (2023)
Event-based backpropagation on the neuromorphic platform SpiNNaker2
von: Béna, Gabriel, et al.
Veröffentlicht: (2024)
von: Béna, Gabriel, et al.
Veröffentlicht: (2024)
SpiNNaker2: A Large-Scale Neuromorphic System for Event-Based and Asynchronous Machine Learning
von: Gonzalez, Hector A., et al.
Veröffentlicht: (2024)
von: Gonzalez, Hector A., et al.
Veröffentlicht: (2024)
Efficient Deployment of Spiking Neural Networks on SpiNNaker2 for DVS Gesture Recognition Using Neuromorphic Intermediate Representation
von: Arfa, Sirine, et al.
Veröffentlicht: (2025)
von: Arfa, Sirine, et al.
Veröffentlicht: (2025)
Neuromorphic hardware for sustainable AI data centers
von: Vogginger, Bernhard, et al.
Veröffentlicht: (2024)
von: Vogginger, Bernhard, et al.
Veröffentlicht: (2024)
Neuromorphic visual attention for Sign-language recognition on SpiNNaker
von: Liskova, Sarka, et al.
Veröffentlicht: (2026)
von: Liskova, Sarka, et al.
Veröffentlicht: (2026)
Ghidorah: Fast LLM Inference on Edge with Speculative Decoding and Hetero-Core Parallelism
von: Wei, Jinhui, et al.
Veröffentlicht: (2025)
von: Wei, Jinhui, et al.
Veröffentlicht: (2025)
DynaTrain: Fast Online Parallelism Switching for Elastic LLM Training
von: Wang, Yuanqing, et al.
Veröffentlicht: (2026)
von: Wang, Yuanqing, et al.
Veröffentlicht: (2026)
LLM-CoOpt: A Co-Design and Optimization Framework for Efficient LLM Inference on Heterogeneous Platforms
von: Kong, Jie, et al.
Veröffentlicht: (2026)
von: Kong, Jie, et al.
Veröffentlicht: (2026)
HeteGen: Heterogeneous Parallel Inference for Large Language Models on Resource-Constrained Devices
von: Zhao, Xuanlei, et al.
Veröffentlicht: (2024)
von: Zhao, Xuanlei, et al.
Veröffentlicht: (2024)
A High Energy-Efficiency Multi-core Neuromorphic Architecture for Deep SNN Training
von: Li, Mingjing, et al.
Veröffentlicht: (2024)
von: Li, Mingjing, et al.
Veröffentlicht: (2024)
FastSet: Parallel Claim Settlement
von: Chen, Xiaohong, et al.
Veröffentlicht: (2025)
von: Chen, Xiaohong, et al.
Veröffentlicht: (2025)
MoEntwine: Unleashing the Potential of Wafer-scale Chips for Large-scale Expert Parallel Inference
von: Tang, Xinru, et al.
Veröffentlicht: (2025)
von: Tang, Xinru, et al.
Veröffentlicht: (2025)
Parallel Track Transformers: Enabling Fast GPU Inference with Reduced Synchronization
von: Wang, Chong, et al.
Veröffentlicht: (2026)
von: Wang, Chong, et al.
Veröffentlicht: (2026)
Arctic Inference with Shift Parallelism: Fast and Efficient Open Source Inference System for Enterprise AI
von: Rajbhandari, Samyam, et al.
Veröffentlicht: (2025)
von: Rajbhandari, Samyam, et al.
Veröffentlicht: (2025)
FLYING SERVING: On-the-Fly Parallelism Switching for Large Language Model Serving
von: Gao, Shouwei, et al.
Veröffentlicht: (2026)
von: Gao, Shouwei, et al.
Veröffentlicht: (2026)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
von: Lacey, Dane C., et al.
Veröffentlicht: (2024)
von: Lacey, Dane C., et al.
Veröffentlicht: (2024)
Parallel Paradigms in Modern HPC: A Comparative Analysis of MPI, OpenMP, and CUDA
von: ALHafez, Nizar, et al.
Veröffentlicht: (2025)
von: ALHafez, Nizar, et al.
Veröffentlicht: (2025)
Accelerating Heterogeneous Tensor Parallelism via Flexible Workload Control
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
Resource-efficient Parallel Split Learning in Heterogeneous Edge Computing
von: Zhang, Mingjin, et al.
Veröffentlicht: (2024)
von: Zhang, Mingjin, et al.
Veröffentlicht: (2024)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
von: Liang, Antian, et al.
Veröffentlicht: (2025)
von: Liang, Antian, et al.
Veröffentlicht: (2025)
Heterogeneous Federated Fine-Tuning with Parallel One-Rank Adaptation
von: Zhang, Zikai, et al.
Veröffentlicht: (2026)
von: Zhang, Zikai, et al.
Veröffentlicht: (2026)
ECC-SNN: Cost-Effective Edge-Cloud Collaboration for Spiking Neural Networks
von: Yu, Di, et al.
Veröffentlicht: (2025)
von: Yu, Di, et al.
Veröffentlicht: (2025)
Mapping Large Memory-constrained Workflows onto Heterogeneous Platforms
von: Kulagina, Svetlana, et al.
Veröffentlicht: (2024)
von: Kulagina, Svetlana, et al.
Veröffentlicht: (2024)
Federated Inference for Heterogeneous LLM Communication and Collaboration
von: Chen, Zihan, et al.
Veröffentlicht: (2026)
von: Chen, Zihan, et al.
Veröffentlicht: (2026)
Amoeba: Runtime Tensor Parallel Transformation for LLM Inference Services
von: Chen, Haoyu, et al.
Veröffentlicht: (2025)
von: Chen, Haoyu, et al.
Veröffentlicht: (2025)
HAP: Hybrid Adaptive Parallelism for Efficient Mixture-of-Experts Inference
von: Lin, Haoran, et al.
Veröffentlicht: (2025)
von: Lin, Haoran, et al.
Veröffentlicht: (2025)
Staleness-Centric Optimizations for Parallel Diffusion MoE Inference
von: Luo, Jiajun, et al.
Veröffentlicht: (2024)
von: Luo, Jiajun, et al.
Veröffentlicht: (2024)
Opt4GPTQ: Co-Optimizing Memory and Computation for 4-bit GPTQ Quantized LLM Inference on Heterogeneous Platforms
von: Zhang, Yaozheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yaozheng, et al.
Veröffentlicht: (2025)
DAK: Direct-Access-Enabled GPU Memory Offloading with Optimal Efficiency for LLM Inference
von: Lin, Shouxu, et al.
Veröffentlicht: (2026)
von: Lin, Shouxu, et al.
Veröffentlicht: (2026)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
von: Mo, Zizhao, et al.
Veröffentlicht: (2025)
von: Mo, Zizhao, et al.
Veröffentlicht: (2025)
Accelerating Microswimmer Simulations via a Heterogeneous Pipelined Parallel-in-Time Framework
von: Huang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Huang, Ruixiang, et al.
Veröffentlicht: (2026)
SiDP: Memory-Efficient Data Parallelism for Offline LLM Inference
von: Zhao, Alan, et al.
Veröffentlicht: (2026)
von: Zhao, Alan, et al.
Veröffentlicht: (2026)
cuFastTuckerPlus: A Stochastic Parallel Sparse FastTucker Decomposition Using GPU Tensor Cores
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
ParallelSFL: A Novel Split Federated Learning Framework Tackling Heterogeneity Issues
von: Liao, Yunming, et al.
Veröffentlicht: (2024)
von: Liao, Yunming, et al.
Veröffentlicht: (2024)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
von: Qiao, Tong, et al.
Veröffentlicht: (2025)
von: Qiao, Tong, et al.
Veröffentlicht: (2025)
AnchorTP: Resilient LLM Inference with State-Preserving Elastic Tensor Parallelism
von: Xu, Wendong, et al.
Veröffentlicht: (2025)
von: Xu, Wendong, et al.
Veröffentlicht: (2025)
Surviving Partial Rank Failures in Wide Expert-Parallel MoE Inference
von: Sun, Xun, et al.
Veröffentlicht: (2026)
von: Sun, Xun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
An End-to-End DNN Inference Framework for the SpiNNaker2 Neuromorphic MPSoC
von: Jobst, Matthias, et al.
Veröffentlicht: (2025) -
Hardware-Aware Fine-Tuning of Spiking Q-Networks on the SpiNNaker2 Neuromorphic Platform
von: Arfa, Sirine, et al.
Veröffentlicht: (2025) -
Language Modeling on a SpiNNaker 2 Neuromorphic Chip
von: Nazeer, Khaleelulla Khan, et al.
Veröffentlicht: (2023) -
Event-based backpropagation on the neuromorphic platform SpiNNaker2
von: Béna, Gabriel, et al.
Veröffentlicht: (2024) -
SpiNNaker2: A Large-Scale Neuromorphic System for Event-Based and Asynchronous Machine Learning
von: Gonzalez, Hector A., et al.
Veröffentlicht: (2024)