Using hls4ml, the authors estimate an Alveo U250 FPGA can process the LHCb track-embedding MLP at 1.1 million events per second, exceeding the measured 0.82 million events per second of an RTX 3090 GPU at lower power.
Throughput-Optimized OpenCL-based FPGA Accelerator for Large-Scale Convolutional Neural Network s,
1 Pith paper cite this work, alongside 11 external citations. Polarity classification is still indexing.
1
Pith paper citing it
11
external citations · OpenAlex
citation-role summary
background 1
citation-polarity summary
fields
hep-ex 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Comparative Analysis of FPGA and GPU Performance for Machine Learning-Based Track Reconstruction at LHCb
Using hls4ml, the authors estimate an Alveo U250 FPGA can process the LHCb track-embedding MLP at 1.1 million events per second, exceeding the measured 0.82 million events per second of an RTX 3090 GPU at lower power.