arXiv ScienceSearch

arXiv subjects

Maria Spiropulu

Publications and source records attributed to Maria Spiropulu.

At least 19 recordsLinked to original sources

Entanglement swapping across a five-node relay in a multiplexed quantum-classical network

Quantum networks are resources for scaling quantum computers and distributed sensing technologies while offering post-quantum security benefits. Teleporting non-classical resources like entanglement, via so called entanglement swapping, is essential for networks in particular overcoming rate-loss limits via quantum repeaters. Deploying these systems on real infrastructure will likely require multiplexing photonic qubits into fibers carrying 'classical' light encoding standard Internet communications and control plane signals for multi-node quantum protocols. Here, we report the first demonstration of entanglement swapping and conventional communications operating over the same fibers. Entanglement is swapped across a five-node quantum relay topology connected by four long-distance fibers, each populated with classical data signals. Time-bin entangled photons in the C-band are multiplexed alongside C-band classical signals using dense-wavelength division multiplexing, introducing noise photons generated by high-power classical light. We experimentally and theoretically characterize the trade-off between quantum fidelity and Raman noise photons. Entanglement swapping is demonstrated over a maximum fiber length of 40 km (four 10-km fibers) while simultaneously transmitting 10-Gbps classical data through all fibers. These results represent a significant advancement in the demonstrated complexity of coexisting quantum and classical networks and provide a roadmap for achieving the widespread deployment of advanced quantum technologies.

quant-ph

FQTree: Fine-grained Quantization and Hardware Generation of Boosted Decision Trees

Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing designs often rely on uniform or manually tuned fixed-point formats, which can introduce unnecessary hardware cost or accuracy loss. This work presents the FQTree algorithm{https://github.com/ecs-bristol/FQTree} for fine-grained quantization-aware training of BDTs, together with the QXGB framework for automatic hardware generation. FQTree introduces a hardware-oriented leaf-value quantization scheme that uses a global quantization step together with a tree-wise shift, enabling compact non-negative integer leaf representations, controlled clipping/pruning, and bias folding to reduce datapath cost. This work further applies this quantization during boosting so that later trees adapt to the errors of the already-quantized ensemble, and then lowers the trained model into low-latency hardware implementations through a compiler-based flow. Results on JSC, MNIST, and NID show that our method reduces LUT usage by 26-57\% compared with the state-of-the-art FPGA-based BDT designs while matching or improving accuracy.

cs.AR

Photonic quantum information with time-bins: Principles and applications

Long-range quantum communication, distributed quantum computing, and sensing applications require robust and reliable ways to encode transmitted quantum information. In this context, time-bin encoding has emerged as a promising candidate due to its resilience to mechanical and thermal perturbations, depolarization from refractive index changes, and birefringence in fiber optic media. Time-bin quantum bits (qubits) can be produced in various ways, and each implementation calls for different considerations regarding design parameters, component compatibility (optical, electrical, electro-optical), and measurement procedures. Here, we provide a comprehensive overview of experimental methods for preparing and characterizing time-bin qubits (TBQs) for quantum communication protocols, with an assessment of their advantages and limitations. We discuss challenges in transmitting TBQs over optical fibers and free-space channels, and methods to overcome them. We also analyze the selection of key time-bin parameters and component requirements across experiments. This leads us to explore the preparation and characterization of time-bin entanglement and examine requirements for interference of time-bins from separate sources. Further, we cover preparation and characterization techniques for high-dimensional time-bin states, namely qudits, and the generation of time-bin entangled qudit pairs. We review time-energy entanglement and key experimental realizations. Finally, we present notable applications of time-bin encoded quantum states, from quantum communication protocols to photonic quantum computation. This work serves as an accessible introduction and a comprehensive review of recent developments.

quant-ph

Fast and bright scintillators for ultrafast materials dynamics using 4th generation synchrotron

We present recent advances in fast and bright scintillators for ultrafast X-ray phase contrast imaging of dynamic materials experiments at the upgraded Advanced Photon Source (APS-U), a fourth generation synchrotron. APS-U enables hard X-ray imaging at frame rates of at least 13 MHz (corresponding to 77 ns or shorter interframe intervals), creating a new need for scintillators with faster response and higher light output than lutetium yttrium oxyorthosilicate (LYSO). For indirect imaging and diffraction with ultrafast cameras, commercial lanthanum bromide (LaBr3) and cerium bromide (CeBr3) are promising candidates. These materials exhibit decay times approximately a factor of two shorter than LYSO (around 40 ns) and lutetium oxyorthosilicate (LSO), while maintaining comparable light yield per incident X-ray photon. However, their implementation at APS-U requires addressing several challenges, including material limitations due to hygroscopicity, efficient optical coupling to imaging systems, and high quantum efficiency for conversion of scintillation light, predominantly at wavelengths below 400 nm, into detectable electronic signals. We report results from material characterization, detector integration and packaging, and beamline experiments of materials with impact. In addition, emerging scintillator classes, including perovskites and high-entropy materials, are discussed as potential alternatives for next-generation ultrafast X-ray diagnostics.

physics.ins-det

HGQ-LUT: Fast LUT-Aware Training and Efficient Architectures for DNN Inference

Lookup-table (LUT) based neural networks can deliver ultra-low latency and excellent hardware efficiency on FPGAs by mapping arithmetic operations directly onto the logic primitives. However, state-of-the-art LUT-aware training (LAT) approaches remain difficult to use in practice: they are often orders of magnitude slower to train than conventional networks, require non-trivial manual tuning for hardware efficiency, and lack an end-to-end workflow. This work presents HGQ-LUT, integrated in https://github.com/calad0i/HGQ2, a new LAT approach that achieves state-of-the-art hardware efficiency while accelerating training by over 100 times on modern GPUs. HGQ-LUT introduces LUT-Dense and LUT-Conv layers that are implemented with regular, accelerator-efficient tensor operations during training, which are then compiled into logic LUTs for hardware. By combining these layers with fine-grained, element-wise heterogeneous quantization (including zero-bit pruning) and a LUT-aware resource surrogate, HGQ-LUT enables the automatic exploration of accuracy-resource trade-offs without manual bit-width tuning. We further integrate HGQ-LUT into open-source toolchains, enabling unified design, compilation, and bit-exact verification of hybrid architectures that mix LUT-based with conventional arithmetic blocks. These features make LAT-based DNNs practical for real-world deployment, such as at the CERN Large Hadron Collider's experiments.

cs.AR

da4ml: Distributed Arithmetic for Real-time Neural Networks on FPGAs

Neural networks with a latency requirement on the order of microseconds, like the ones used at the CERN Large Hadron Collider, are typically deployed on FPGAs fully unrolled and pipelined. A bottleneck for the deployment of such neural networks is area utilization, which is directly related to the required constant matrix-vector multiplication (CMVM) operations. In this work, we propose an efficient algorithm for implementing CMVM operations with distributed arithmetic on FPGAs that simultaneously optimizes for area consumption and latency. The algorithm achieves resource reduction similar to state-of-the-art algorithms while being significantly faster to compute. The proposed algorithm is open-sourced and integrated into the \texttt{hls4ml} library, a free and open-source library for running real-time neural network inference on FPGAs. We show that the proposed algorithm can reduce on-chip resources by up to a third for realistic, highly quantized neural networks while simultaneously reducing latency, enabling the implementation of previously infeasible networks.

cs.AR

High-dimensional quantum communication with scalable photonic entanglement in time and frequency

High-dimensional photonic entanglement holds significant promise for advancing quantum communication, computation, and metrology. For example, large-alphabet quantum communication protocols are known to benefit from enhanced noise resilience and information capacity via multi-bit time-bin encoding. Yet, characterizing high-dimensional entangled states is challenging, as full state tomography becomes prohibitively costly and often requires unrealizable measurements. Here, we demonstrate a scan-free method to characterize high-dimensional entanglement in the time-frequency domain. Our reconstruction achieves a record $5.70\pm0.07$ ebits and a fidelity of $65.4\pm0.4\%$ with the maximally entangled state of local dimension $1021$, certifying the presence of $668$-dimensional entanglement. We further prove the attainability of a secure key rate of $15.6$ kB/s in a composable finite-size, entanglement-based protocol, and show that in continuous operation, the setup can quickly approach asymptotic key rates. Using commercial telecom components and state-of-the-art low-jitter single-photon detectors, our scalable architecture offers a practical path towards high-rate, noise-resilient quantum communication testbeds.

quant-ph

Towards High-Efficiency Particle Detection Using Superconducting Microwire Arrays

We present a detailed study of an 8-channel $1\times1$ mm$^{2}$ WSi superconducting microwire single photon detector (SMSPD) array exposed to 120 GeV hadron beam and 120 GeV muon beam at the CERN Super Proton Synchrotron H6 beamline. Following up on our first detailed characterization of the efficiency and response of an SMSPD fabricated on a 3 nm WSi film, we report measurements of enhanced particle detection efficiency using a sensor fabricated from a thicker 4.7nm-thick WSi film. We also report the first SMSPD detection efficiency measurement made for muons. Measurements are enabled by a silicon tracking telescope providing 10 $μ$m in-situ spatial resolution. The results show a fill factor-normalized detection efficiency of 75% and a time resolution of about 130 ps across pixels. These findings represent a significant advancement toward developing high-efficiency SMSPD charged particle tracking systems with simultaneous precision timing, with potential applications in future accelerator-based experiments such as the FCC-ee and Muon Collider.

physics.ins-det

Fast Jet Tagging with MLP-Mixers on FPGAs

We explore the innovative use of MLP-Mixer models for real-time jet tagging and establish their feasibility on resource-constrained hardware like FPGAs. MLP-Mixers excel in processing sequences of jet constituents, achieving state-of-the-art performance on datasets mimicking Large Hadron Collider conditions. By using advanced optimization techniques such as High-Granularity Quantization and Distributed Arithmetic, we achieve unprecedented efficiency. These models match or surpass the accuracy of previous architectures, reduce hardware resource usage by up to 97%, double the throughput, and half the latency. Additionally, non-permutation-invariant architectures enable smart feature prioritization and efficient FPGA deployment, setting a new benchmark for machine learning in real-time data processing at particle colliders.

physics.ins-det

HGQ: High Granularity Quantization for Real-time Neural Networks on FPGAs

Neural networks with sub-microsecond inference latency are required by many critical applications. Targeting such applications deployed on FPGAs, we present High Granularity Quantization (HGQ), a quantization-aware training framework that optimizes parameter bit-widths through gradient descent. Unlike conventional methods, HGQ determines the optimal bit-width for each parameter independently, making it suitable for hardware platforms supporting heterogeneous arbitrary precision arithmetic. In our experiments, HGQ shows superior performance compared to existing network compression methods, achieving orders of magnitude reduction in resource consumption and latency while maintaining the accuracy on several benchmark tasks. These improvements enable the deployment of complex models previously infeasible due to resource or latency constraints. HGQ is open-source and is used for developing next-generation trigger systems at the CERN ATLAS and CMS experiments for particle physics, enabling the use of advanced machine learning models for real-time data selection with sub-microsecond latency.

cs.LG

An Evaluation of Representation Learning Methods in Particle Physics Foundation Models

We present a systematic evaluation of representation learning objectives for particle physics within a unified framework. Our study employs a shared transformer-based particle-cloud encoder with standardized preprocessing, matched sampling, and a consistent evaluation protocol on a jet classification dataset. We compare contrastive (supervised and self-supervised), masked particle modeling, and generative reconstruction objectives under a common training regimen. In addition, we introduce targeted supervised architectural modifications that achieve state-of-the-art performance on benchmark evaluations. This controlled comparison isolates the contributions of the learning objective, highlights their respective strengths and limitations, and provides reproducible baselines. We position this work as a reference point for the future development of foundation models in particle physics, enabling more transparent and robust progress across the community.

cs.LG

JEDI-linear: Fast and Efficient Graph Neural Networks for Jet Tagging on FPGAs

Graph Neural Networks (GNNs), particularly Interaction Networks (INs), have shown exceptional performance for jet tagging at the CERN High-Luminosity Large Hadron Collider (HL-LHC). However, their computational complexity and irregular memory access patterns pose significant challenges for deployment on FPGAs in hardware trigger systems, where strict latency and resource constraints apply. In this work, we propose JEDI-linear, a novel GNN architecture with linear computational complexity that eliminates explicit pairwise interactions by leveraging shared transformations and global aggregation. To further enhance hardware efficiency, we introduce fine-grained quantization-aware training with per-parameter bitwidth optimization and employ multiplier-free multiply-accumulate operations via distributed arithmetic. Evaluation results show that our FPGA-based JEDI-linear achieves 3.7 to 11.5 times lower latency, up to 150 times lower initiation interval, and up to 6.2 times lower LUT usage compared to state-of-the-art GNN designs while also delivering higher model accuracy and eliminating the need for DSP blocks entirely. This is the first interaction-based GNN to achieve less than 60~ns latency and currently meets the requirements for use in the HL-LHC CMS Level-1 trigger system. This work advances the next-generation trigger systems by enabling accurate, scalable, and resource-efficient GNN inference in real-time environments. Our open-sourced templates will further support reproducibility and broader adoption across scientific applications.

hep-ex

RINO: Renormalization Group Invariance with No Labels

A common challenge with supervised machine learning (ML) in high energy physics (HEP) is the reliance on simulations for labeled data, which can often mismodel the underlying collision or detector response. To help mitigate this problem of domain shift, we propose RINO (Renormalization Group Invariance with No Labels), a self-supervised learning approach that can instead pretrain models directly on collision data, learning embeddings invariant to renormalization group flow scales. In this work, we pretrain a transformer-based model on jets originating from quantum chromodynamic (QCD) interactions from the JetClass dataset, emulating real QCD-dominated experimental data, and then finetune on the JetNet dataset -- emulating simulations -- for the task of identifying jets originating from top quark decays. RINO demonstrates improved generalization from the JetNet training data to JetClass data compared to supervised training on JetNet from scratch, demonstrating the potential for RINO pretraining on real collision data followed by fine-tuning on small, high-quality MC datasets, to improve the robustness of ML models in HEP.

hep-ex

Experimental high-dimensional entanglement certification and quantum steering with time-energy measurements

High-dimensional entanglement provides unique ways of transcending the limitations of current approaches in quantum information processing, quantum communications based on qubits. The generation of time-frequency qudit states offer significantly increased quantum capacities while keeping the number of photons constant, but pose significant challenges regarding the possible measurements for certification of entanglement. Here, we develop a new scheme and experimentally demonstrate the certification of 24-dimensional entanglement and a 9-dimensional quantum steering. We then subject our photon-pairs to dispersion conditions equivalent to the transmission through 600-km of fiber and still certify 21-dimensional entanglement. Furthermore, we use a steering inequality to prove 7-dimensional entanglement in a semi-device independent manner, proving that large chromatic dispersion is not an obstacle in distributing and certifying high-dimensional entanglement and quantum steering. Our approach, leveraging intrinsic large-alphabet nature of telecom-band photons, enables scalable, commercially viable, and field-deployable entangled and steerable quantum sources, providing a pathway towards fully scalable quantum information processer and high-dimensional quantum communication networks.

quant-ph

Sub-microsecond Transformers for Jet Tagging on FPGAs

We present the first sub-microsecond transformer implementation on an FPGA achieving competitive performance for state-of-the-art high-energy physics benchmarks. Transformers have shown exceptional performance on multiple tasks in modern machine learning applications, including jet tagging at the CERN Large Hadron Collider (LHC). However, their computational complexity prohibits use in real-time applications, such as the hardware trigger system of the collider experiments up until now. In this work, we demonstrate the first application of transformers for jet tagging on FPGAs, achieving $\mathcal{O}(100)$ nanosecond latency with superior performance compared to alternative baseline models. We leverage high-granularity quantization and distributed arithmetic optimization to fit the entire transformer model on a single FPGA, achieving the required throughput and latency. Furthermore, we add multi-head attention and linear attention support to hls4ml, making our work accessible to the broader fast machine learning community. This work advances the next-generation trigger systems for the High Luminosity LHC, enabling the use of transformers for real-time applications in high-energy physics and beyond.

physics.ins-det

Optimal filtering and generation of entangled photons for quantum applications in the presence of noise

Filtering is commonly used in quantum optics to reject noise photons, and also to enable interference between independent photons. However, filtering the joint spectrum of photon pairs can reduce the inherent coincidence probability or loss-independent heralding efficiency. Here, we investigate filtering for multiphoton applications based on entanglement and interference (e.g., quantum teleportation). We multiplex C-band entangled photons and C-band classical communications into the same long-distance fibers, which enables scalable low-loss quantum networking but requires filtering of spontaneous Raman scattering noise from classical light. Using tunable-bandwidth filters, low-jitter detectors, and polarization filters, we co-propagate time-bin-entangled photons at wavelengths compatible with erbium-ion quantum memories (1536.5 nm) and 10-Gbps C-band classical data over 25 km/25 km of standard fiber. Narrow filtering enables mW-level C-band power, which exceeds comparable studies by roughly an order of magnitude and could feasibly support Tbps classical rates. We evaluate how performance depends on pump and filter bandwidths, multipair emission, filter shapes, loss, phase matching, and how quantum information is measured. We find a trade-off between improving noise impact and single-mode purity and discuss mitigation methods toward optimal multiphoton applications. Importantly, these results apply to noise in free space and in quantum devices (sources, frequency converters, switches, detectors, etc.) and provide insight on filter-induced degradation of single-photon purity and rates even in noise-free environments.

quant-ph

Entanglement swapping systems toward a quantum internet

We demonstrate conditional entanglement swapping, i.e. teleportation of entanglement, between time-bin qubits at the telecommunication wavelength of 1536.4 nm with high fidelity of 87\%. Our system is deployable, utilizing modular, off-the-shelf, fiber-coupled, and electrically controlled components such as electro-optic modulators. It leverages the precise timing resolution of superconducting nanowire detectors, which are controlled and read out via a custom developed graphical user interface. The swapping process is described, interpreted, and guided using characteristic function-based analytical modeling that accounts for realistic imperfections. Our system supports quantum networking protocols, including source-independent quantum key distribution, with an estimated secret key rate of approximately 0.5 bits per sifted bit.

quant-ph

Analytical Modeling of Real-World Photonic Quantum Teleportation

We develop analytical models for realistic photonic quantum teleportation experiments with time-bin qubits, utilizing phase space methods from quantum optics. These models yield analytical expressions for Hong-Ou-Mandel interference visibilities and qubit fidelities, accounting for imperfections such as loss, photon distinguishability, and unwanted multi-photon events in single and entangled photon sources. Our expressions agree with the Hong-Ou-Mandel interference visibilities and teleportation fidelities reported by Valivarthi et al. (PRX Quantum 1, 020317, 2020). We use these models to predict and analyze the outcomes of future teleportation experiments under varying degrees of imperfections.

quant-ph