Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Neu, Marc, Baptist, Frank, Lobmaier, Thomas, Papagno, Fabio, Ferber, Torben, Becker, Jürgen
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918494624284672
author Neu, Marc
Baptist, Frank
Lobmaier, Thomas
Papagno, Fabio
Ferber, Torben
Becker, Jürgen
author_facet Neu, Marc
Baptist, Frank
Lobmaier, Thomas
Papagno, Fabio
Ferber, Torben
Becker, Jürgen
contents Graph neural networks are increasingly adopted in trigger systems for collider experiments, where strict latency and throughput constraints render deployment on embedded platforms challenging. As detectors move towards higher granularity, the number of inputs per inference increase and FPGA-only solutions face resource bottlenecks. This work presents an end-to-end demonstrator for the real-time deployment of a dynamic Graph Neural Network for the Belle II electromagnetic calorimeter hardware trigger on the AMD Versal VCK190, leveraging both FPGA fabric and AI Engine tiles. We develop a Python-based semi-automated design flow covering operator fusion, partitioning, mapping, spatial parallelization, and kernel-level optimization. Our design achieves a throughput of 2.94 million events per second at an end-to-end latency of 7.15 microseconds. Compared to the FPGA-only baseline, this represents a 53% throughput improvement while reducing DSP utilization from 99% to 19% at 29% AI Engine tile utilization. To validate the deployment, an interactive visualization pipeline enables real-time monitoring of inference results on the physical demonstrator.
format Preprint
id arxiv_https___arxiv_org_abs_2605_10612
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science
Neu, Marc
Baptist, Frank
Lobmaier, Thomas
Papagno, Fabio
Ferber, Torben
Becker, Jürgen
Hardware Architecture
Machine Learning
Graph neural networks are increasingly adopted in trigger systems for collider experiments, where strict latency and throughput constraints render deployment on embedded platforms challenging. As detectors move towards higher granularity, the number of inputs per inference increase and FPGA-only solutions face resource bottlenecks. This work presents an end-to-end demonstrator for the real-time deployment of a dynamic Graph Neural Network for the Belle II electromagnetic calorimeter hardware trigger on the AMD Versal VCK190, leveraging both FPGA fabric and AI Engine tiles. We develop a Python-based semi-automated design flow covering operator fusion, partitioning, mapping, spatial parallelization, and kernel-level optimization. Our design achieves a throughput of 2.94 million events per second at an end-to-end latency of 7.15 microseconds. Compared to the FPGA-only baseline, this represents a 53% throughput improvement while reducing DSP utilization from 99% to 19% at 29% AI Engine tile utilization. To validate the deployment, an interactive visualization pipeline enables real-time monitoring of inference results on the physical demonstrator.
title Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science
topic Hardware Architecture
Machine Learning
url https://arxiv.org/abs/2605.10612