We present a systematic study of tensor network (TN) models—matrix product states and tree TNs—for real-time jet tagging in high-energy physics, with a focus on low-latency deployment on field programmable gate arrays (FPGA). Motivated by the strict requirements of the high-luminosity large hadron collider Level-1 trigger system, we explore TNs as compact and interpretable alternatives to deep neural networks. Using low-level jet constituent features, our models achieve competitive performance compared to state-of-the-art deep learning classifiers. We investigate post-training quantization to enable hardware-efficient implementations without degrading classification performance or latency. Selected models are synthesized for FPGA deployment and implemented to obtain resource usage, latency, and memory occupancy evaluation. The resulting measurements and timing estimates indicate sub-microsecond inference latency, supporting the feasibility of TN-based models for online deployment in real-time trigger systems. Overall, this study highlights the potential of TN-based models for fast and resource-efficient inference in low-latency environments.

Towards tensor network models for low-latency jet tagging on FPGAs

Coppi A.;Borella L.;Pazzini J.;Triossi A.;Montangero S.
2026

Abstract

We present a systematic study of tensor network (TN) models—matrix product states and tree TNs—for real-time jet tagging in high-energy physics, with a focus on low-latency deployment on field programmable gate arrays (FPGA). Motivated by the strict requirements of the high-luminosity large hadron collider Level-1 trigger system, we explore TNs as compact and interpretable alternatives to deep neural networks. Using low-level jet constituent features, our models achieve competitive performance compared to state-of-the-art deep learning classifiers. We investigate post-training quantization to enable hardware-efficient implementations without degrading classification performance or latency. Selected models are synthesized for FPGA deployment and implemented to obtain resource usage, latency, and memory occupancy evaluation. The resulting measurements and timing estimates indicate sub-microsecond inference latency, supporting the feasibility of TN-based models for online deployment in real-time trigger systems. Overall, this study highlights the potential of TN-based models for fast and resource-efficient inference in low-latency environments.
File in questo prodotto:
File Dimensione Formato  
Coppi_2026_Mach._Learn.%3A_Sci._Technol._7_045065.pdf

accesso aperto

Tipologia: Published (Publisher's Version of Record)
Licenza: Creative commons
Dimensione 1.22 MB
Formato Adobe PDF
1.22 MB Adobe PDF Visualizza/Apri
Pubblicazioni consigliate

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11577/3613426
Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 0
  • ???jsp.display-item.citation.isi??? ND
  • OpenAlex 0
social impact