arXiv ScienceSearch

arXiv subjects

Rui Qu

Publications and source records attributed to Rui Qu.

10 recordsLinked to original sources

Experimental High-Dimensional Quantum Overlapping Tomography

Large-scale quantum systems have advanced rapidly via the exploration of more particles and higher dimensions, offering great potential for developing quantum technologies. However, their characterization becomes prohibitive with increasing local dimensionality and particle number. Here we propose high-dimensional quantum overlapping tomography based on a graph-theoretic formulation, which allows one to efficiently reconstruct few-body marginals of multipartite high-dimensional quantum systems. We experimentally realize it on a photonic four-party entangled state in a $4 \times 4 \times 2 \times 2$ system. Using measurements in mutually unbiased bases, we reconstruct all six two-body marginals with only 25 projective measurement settings, compared with 94 and 225 settings for independent tomography of all two-body reduced states and full state tomography, respectively. The reconstructed marginals reveal a layered entanglement structure vital for high-dimensional quantum networks. We further show that these marginals enable more noise-resilient certification of multipartite high-dimensional entanglement than the fidelity-based criterion. Our work thus offers a scalable route for learning multidimensional quantum systems.

quant-ph

Reflection in the Dark: Exposing and Escaping the Black Box in Reflective Prompt Optimization

Automatic prompt optimization (APO) has emerged as a powerful paradigm for improving LLM performance without manual prompt engineering. Reflective APO methods such as GEPA iteratively refine prompts by diagnosing failure cases, but the optimization process remains black-box and label-free, leading to uninterpretable trajectories and systematic failure. We identify and empirically demonstrate four limitations: on GSM8K with a defective seed, GEPA degrades accuracy from 23.81% to 13.50%. We propose VISTA, a multi-agent APO framework that decouples hypothesis generation from prompt rewriting, enabling semantically labeled hypotheses, parallel minibatch verification, and interpretable optimization trace. A two-layer explore-exploit mechanism combining random restart and epsilon-greedy sampling further escapes local optima. VISTA recovers accuracy to 87.57% on the same defective seed and consistently outperforms baselines across all conditions on GSM8K and AIME2025.

cs.AI

DICE: Discrete Interpretable Comparative Evaluation with Probabilistic Scoring for Retrieval-Augmented Generation

As Retrieval-Augmented Generation (RAG) systems evolve toward more sophisticated architectures, ensuring their trustworthiness through explainable and robust evaluation becomes critical. Existing scalar metrics suffer from limited interpretability, inadequate uncertainty quantification, and computational inefficiency in multi-system comparisons, hindering responsible deployment of RAG technologies. We introduce DICE (Discrete Interpretable Comparative Evaluation), a two-stage, evidence-coupled framework that advances explainability and robustness in RAG evaluation. DICE combines deep analytical reasoning with probabilistic $\{A, B, Tie\}$ scoring to produce transparent, confidence-aware judgments that support accountable system improvement through interpretable reasoning traces, enabling systematic error diagnosis and actionable insights. To address efficiency challenges at scale, DICE employs a Swiss-system tournament that reduces computational complexity from $O(N^2)$ to $O(N \log N)$, achieving a 42.9% reduction in our eight-system evaluation while preserving ranking fidelity. Validation on a curated Chinese financial QA dataset demonstrates that DICE achieves 85.7% agreement with human experts, substantially outperforming existing LLM-based metrics such as RAGAS. Our results establish DICE as a responsible, explainable, and efficient paradigm for trustworthy RAG system assessment.

cs.AI

Experimental demonstration of scalable quantum cryptographic conferencing

Quantum network enables a variety of quantum information processing tasks, where multi-user quantum communication is one of the important objectives. Quantum cryptographic conferencing serves as an essential solution to establish secure keys to realize secure multi-user communications. However, existing QCC implementations have been fundamentally limited by the low probability of multi-user coincidence detection to measure or construct the Greenberger-Horne-Zeilinger (GHZ) entangled state. In this work, we report the experimental realization of QCC eliminating the need for coincidence detection, where the GHZ state is constructed by correlating detection events occurring within the coherence time, thereby greatly enhancing the success probability of GHZ-state measurement. Meanwhile, to establish and maintain high-visibility GHZ measurement among three independent users, we developed a three-party phase compensation scheme combined with precise temporal and polarization alignment within a time-bin-phase encoding framework. Furthermore, we designed an efficient pairing strategy to simplify subsequent data processing and enhance processing efficiency. Based on these techniques, we successfully performed QCC experiments over total channel losses of 66.3 dB, corresponding to 331.5 km of commercial fiber (0.2 dB/km), achieving secure key rates of 5.4 bit/s, whereas previous QCC experiments have been limited to 100 km. The results surpass the multi-user repeaterless bound in quantum networks, establishing a new regime of scalable, multi-user quantum communication and paving the way for metropolitan quantum networks.

quant-ph

Agent.xpu: Efficient Scheduling of Agentic LLM Workloads on Heterogeneous SoC

Personal LLM agents increasingly combine foreground reactive interactions with background proactive monitoring, forming long-lived, stateful LLM flows that interleave prefill and token-by-token decode. While modern heterogeneous SoCs integrate CPUs, iGPUs, and NPUs to support on-device intelligence, existing LLM engines assume static, single-shot inference and lack mechanisms for flow-level concurrency, prioritization, and efficient accelerator coordination. As a result, commodity SoCs remain poorly matched to the dynamic, mixed-criticality execution patterns of personal agents. This paper presents Agent$.$xpu, the first LLM engine that orchestrates concurrent reactive and proactive LLM flows on commodity SoCs. Extensive profiling uncovers unique SoC characteristics of operator-accelerator affinity, asymmetric DDR contention, and stage-divergent batching behaviors distinct from cloud-serving assumptions. Agent$.$xpu introduces three key techniques: a heterogeneous execution graph (HEG) capturing NPU/iGPU affinity and elastic operator binding; flow-aware NPU-iGPU coordination with stage elasticity, decoupling prefill and decode to reduce bandwidth contention and enforce priorities; and fine-grained preemption with slack-aware piggybacking to guarantee reactive responsiveness without starving proactive work. Across realistic personal-agent workloads, Agent$.$xpu delivers 1.2-4.9$\times$ proactive throughput and reduces reactive latency by at least 91%, compared with both industrial iGPU-only serving engine and NPU-iGPU static inference with optimal tensor-partitioning schemes. Agent$.$xpu also minimizes energy consumption and graphics interference via controlled iGPU usage.

cs.DC

FluentLip: A Phonemes-Based Two-stage Approach for Audio-Driven Lip Synthesis with Optical Flow Consistency

Generating consecutive images of lip movements that align with a given speech in audio-driven lip synthesis is a challenging task. While previous studies have made strides in synchronization and visual quality, lip intelligibility and video fluency remain persistent challenges. This work proposes FluentLip, a two-stage approach for audio-driven lip synthesis, incorporating three featured strategies. To improve lip synchronization and intelligibility, we integrate a phoneme extractor and encoder to generate a fusion of audio and phoneme information for multimodal learning. Additionally, we employ optical flow consistency loss to ensure natural transitions between image frames. Furthermore, we incorporate a diffusion chain during the training of Generative Adversarial Networks (GANs) to improve both stability and efficiency. We evaluate our proposed FluentLip through extensive experiments, comparing it with five state-of-the-art (SOTA) approaches across five metrics, including a proposed metric called Phoneme Error Rate (PER) that evaluates lip pose intelligibility and video fluency. The experimental results demonstrate that our FluentLip approach is highly competitive, achieving significant improvements in smoothness and naturalness. In particular, it outperforms these SOTA approaches by approximately $\textbf{16.3%}$ in Fr\'echet Inception Distance (FID) and $\textbf{35.2%}$ in PER.

cs.CV

Optimal Overlapping Tomography

Characterising large-scale quantum systems is central to fundamental physics and essential for applications of quantum technologies. While a full characterisation requires exponentially increasing resources, focusing on application-relevant information can often lead to significantly simplified analysis. Overlapping tomography is such a scheme, allowing one to obtain all the information contained in specific subsystems of multiparticle quantum systems in an efficient manner, but the ultimate limits of this approach remain elusive. We present protocols for overlapping tomography that are optimal with respect to the number of measurement settings. First, by providing algorithmic approaches based on graph theory we find the minimal number of Pauli settings, relating overlapping tomography to the problem of covering arrays in combinatorics. This significantly reduces the number of measurement settings, showing for instance that two-body overlapping tomography of nearest neighbours in qubit systems with planar topologies can always be performed with nine Pauli settings. Second, we prove that using general projective measurements, all $k$-body marginals can be reconstructed with only $3^k$ settings, independently of the system size. Finally, we demonstrate the practical applicability of our methods in a six-photon experiment. Our results will find applications in learning noise and interaction patterns in quantum computers as well as characterising fermionic systems in quantum chemistry.

quant-ph

Witnessing Quantum Incompatibility Structures in High-Dimensional Multimeasurement Systems

Quantum incompatibility, referred as the phenomenon that some quantum measurements cannot be performed simultaneously, is necessary for various quantum information processing tasks, such as nonlocality and steering. When these applications come to high-dimensional multimeasurement scenarios, it is crucial and challenging to witness the incompatibility of measurements with complex structures. To address this problem, we propose a modified quantum state discrimination protocol that decomposes complex compatibility structures into pairwise ones and employs noise robustness to bound incompatibility structures. We then derive arithmetic bounds for arbitrary measurements and analytical bounds for mutually unbiased bases, and capture some quantum incompatibility structures where measurements are partly compatible and partly incompatible. Finally, we experimentally demonstrate our results and connect them with quantum steering, quantum simulability and quantum communications.

quant-ph

Quantum state transfer between two photons with polarization and orbital angular momentum via quantum teleportation technology

Quantum teleportation is a useful quantum information technology to transmit quantum states between different degrees of freedom. We here report a quantum state transfer experiment in the linear optical system, transferring a single photon state in the polarization degree of freedom (DoF) to another photon in the orbital angular momentum (OAM) quantum state via a biphoton OAM entangled channel. Our experimental method is based on quantum teleportation technology. The differences between ours and the original teleportation scheme is that the transfer state is known in ours, and our method is for different particles with different DoFs while the original one is for different particles with same DoF. Besides, our present experiment is implemented with a high Bell-efficiency since each of the four hybrid-entangled Bell states can be discriminated. We use six states of poles of the Bloch sphere to test our experiment, and the fidelity of the quantum state transfer is $91.8\pm1.3\%$.

quant-ph

Retrieving High-Dimensional Quantum Steering From a Noisy Environment with N Measurement Settings

One of the most often implied benefits of high-dimensional (HD) quantum systems is to lead to stronger forms of correlations, featuring increased robustness to noise. Here, we experimentally demonstrate the $n$-setting linear HD quantum steering criterion. We verify the large violation of the steering inequalities without full-state tomography. The lower bound of the violation is $2.24\pm0.01$ in 11 dimensions, exceeding the bound ($V<2$) of 2-setting criteria. Hence, a higher strength of steering has been revealed. Moreover, we demonstrate the method for enhancing the noise robustness without increasing dimension, alternatively, by increasing measurement settings. Using the entanglement in 11 dimensions, we experimentally retrieve steering nonlocality with $63.4\pm1.4\%$ isotropic noise fraction, surpassing the $50\%$ limitation of 2-setting criteria. Our work offers the potential for practical one-sided device-independent quantum information processing that tolerates the noisy environment, lossy detection, and transcends the present transmission distance limitation.

quant-ph