FPGA-Accelerated SpeckleNN with SNL for Real-time X-ray Single-Particle Imaging

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Dave, Abhilasha, Wang, Cong, Russell, James, Herbst, Ryan, Thayer, Jana
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866915175607566336
author Dave, Abhilasha
Wang, Cong
Russell, James
Herbst, Ryan
Thayer, Jana
author_facet Dave, Abhilasha
Wang, Cong
Russell, James
Herbst, Ryan
Thayer, Jana
contents We implement a specialized version of our SpeckleNN model for real-time speckle pattern classification in X-ray Single-Particle Imaging (SPI) using the SLAC Neural Network Library (SNL) on an FPGA. This hardware is optimized for inference near detectors in high-throughput X-ray free-electron laser (XFEL) facilities like the Linac Coherent Light Source (LCLS). To fit FPGA constraints, we optimized SpeckleNN, reducing parameters from 5.6M to 64.6K (98.8% reduction) with 90% accuracy. We also compressed the latent space from 128 to 50 dimensions. Deployed on a KCU1500 FPGA, the model used 71% of DSPs, 75% of LUTs, and 48% of FFs, with an average power consumption of 9.4W. The FPGA achieved 45.015us inference latency at 200 MHz. On an NVIDIA A100 GPU, the same inference consumed ~73W and had a 400us latency. Our FPGA version achieved an 8.9x speedup and 7.8x power reduction over the GPU. Key advancements include model specialization and dynamic weight loading through SNL, eliminating time-consuming FPGA re-synthesis for fast, continuous deployment of (re)trained models. These innovations enable real-time adaptive classification and efficient speckle pattern vetoing, making SpeckleNN ideal for XFEL facilities. This implementation accelerates SPI experiments and enhances adaptability to evolving conditions.
format Preprint
id arxiv_https___arxiv_org_abs_2502_19734
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle FPGA-Accelerated SpeckleNN with SNL for Real-time X-ray Single-Particle Imaging
Dave, Abhilasha
Wang, Cong
Russell, James
Herbst, Ryan
Thayer, Jana
Instrumentation and Detectors
Machine Learning
Image and Video Processing
We implement a specialized version of our SpeckleNN model for real-time speckle pattern classification in X-ray Single-Particle Imaging (SPI) using the SLAC Neural Network Library (SNL) on an FPGA. This hardware is optimized for inference near detectors in high-throughput X-ray free-electron laser (XFEL) facilities like the Linac Coherent Light Source (LCLS). To fit FPGA constraints, we optimized SpeckleNN, reducing parameters from 5.6M to 64.6K (98.8% reduction) with 90% accuracy. We also compressed the latent space from 128 to 50 dimensions. Deployed on a KCU1500 FPGA, the model used 71% of DSPs, 75% of LUTs, and 48% of FFs, with an average power consumption of 9.4W. The FPGA achieved 45.015us inference latency at 200 MHz. On an NVIDIA A100 GPU, the same inference consumed ~73W and had a 400us latency. Our FPGA version achieved an 8.9x speedup and 7.8x power reduction over the GPU. Key advancements include model specialization and dynamic weight loading through SNL, eliminating time-consuming FPGA re-synthesis for fast, continuous deployment of (re)trained models. These innovations enable real-time adaptive classification and efficient speckle pattern vetoing, making SpeckleNN ideal for XFEL facilities. This implementation accelerates SPI experiments and enhances adaptability to evolving conditions.
title FPGA-Accelerated SpeckleNN with SNL for Real-time X-ray Single-Particle Imaging
topic Instrumentation and Detectors
Machine Learning
Image and Video Processing
url https://arxiv.org/abs/2502.19734