arXiv ScienceSearch

arXiv subjects

Peng Huang

Publications and source records attributed to Peng Huang.

At least 19 recordsLinked to original sources

Fully integrated continuous-variable quantum key distribution with composable security over 100 km

Quantum key distribution (QKD) guarantees information-theoretic security by the laws of physics, but deployment at scale requires compact, manufacturable photonic terminals. Continuous-variable QKD (CV-QKD) is well suited for this transition through telecom-compatible, room-temperature coherent detection. However, unifying full on-chip core terminal integration, room-temperature operation, high loss tolerance, and composable end-to-end security in long-distance QKD remains a key bottleneck. Here we report a fully integrated CV-QKD platform in which two hybrid III--V/Si$_3$N$_4$ integrated lasers, a silicon transmitter, and a silicon coherent receiver implement the core terminal functions, operating with a local local oscillator (LLO) over fibre links of 25--150 km. A Bayesian machine-learning algorithm maintains robust phase lock throughout the long records required for composable security, consistently outperforming the conventional unscented Kalman filter, while rate-matched multidimensional reconciliation approaches the Shannon limit. The system certifies a composable finite-size secret-key rate of 29.3 kbps at 100 km from a 140-billion-symbol block, with 12.9 kbps at 125 km under finite-size analysis and 9.17 kbps at 150 km under asymptotic analysis. By establishing the longest finite-size and asymptotic reaches and the highest secret-key rate per symbol reported for integrated CV-QKD, this work advances the development of practical chip-based quantum networks.

quant-ph

Motional Degrees of Freedom in Network Hamiltonian Models

Network Hamiltonian Models (NHMs) provide an efficient framework for modeling the aggregation of interacting particles (e.g., the condensation of proteins into gel-like, oligomeric, or fibrillar states), representing the system as a network whose edges represent bound interactions. Terms within the network Hamiltonian represent multi-body interactions governing aggregation behavior, and are specified via topological degrees of freedom. Because motional degrees of freedom are not explicitly represented within the NHM, their influence must be indirectly accounted for by introduction of corresponding terms to the Hamiltonian. Here, we describe specifications for terms representing motional degrees of freedom of two, three, and four-body interactions. We also consider the impact of these terms on the aggregation states of a minimal system governed only by a pairwise edge potential, showing that three and four-body interactions favor condensation of the system into small droplet-like structures at low temperature.

q-bio.MN

Spectral Attack on Continuous-Variable Quantum Key Distribution Systems

Continuous-variable quantum key distribution (CVQKD) has attracted extensive attention due to its compatibility and low costs. However, bandwidth mismatch exists to varying degrees between the transmitter and receiver. This may prevent frequency components carrying modulation information from being fully perceived by the legitimate party. In this paper, we identify a practical security loophole caused by bandwidth mismatch and propose a corresponding spectral attack scheme. Different from previous approaches that exploit security loopholes to conceal the excess noise introduced by intercept-resend attacks, this scheme can directly obtain raw-key information without introducing additional disturbances. A proof-of-principle attack on a CVQKD system with filtering operation is constructed to verify the feasibility. Experimental results indicate that Eve can obtain enough information to render the system insecure if this practical security loophole is ignored. Based on the identified security loophole, corresponding defense strategies are proposed. This work helps bridge the gap between theoretical models and practical implementations, providing a reference for defense design in practical quantum communication systems.

quant-ph

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos

We study the challenging problem of novel view video synthesis from single images or monocular videos. Existing methods, which operate under the assumption that pre-trained video models lack native novel view synthesis capability and enforce view alignment via camera conditioning, task-specific fine-tuning, or stepwise hard denoising guidance, often suffer from artifacts and compromised global scene consistency. In this paper, we introduce NeoMap, a novel training-free framework designed to locate high-fidelity, view-consistent novel view solutions from general pre-trained video models. The key to our approach is the core insight that promising novel view solutions are inherently encoded within the natural video data manifold learned by pre-trained models, and the core challenge is simply to locate this optimal solution. We solve this via our core mechanism: convergent manifold alternating projection iterations that optimize the initial noise. Extensive experiments demonstrate that NeoMap significantly outperforms all existing methods across 3 standard novel view synthesis benchmarks, including the challenging Tanks-and-Temples, LLFF and DAVIS datasets, achieving state-of-the-art generation fidelity and top-tier view consistency.

cs.CV

Field Demonstration of a Multi-User Continuous-Variable Quantum Access Network for Quantum-to-the-Home

Realizing scalable Quantum-to-the-Home (QTTH) faces a bottleneck: link asymmetry in broadcast continuous-variable quantum access networks (CV-QANs) hinders the selection of a globally optimal modulation variance. We demonstrate a downstream broadcast CV-QAN connecting a Quantum Line Terminal (QLT) to multiple Quantum Network Units (QNUs) over commercial fiber. Operating within a trusted local network domain, we establish a multi-user utility model to select the optimal shared variance, balancing network efficiency and user fairness. Supported by robust digital signal processing, our 1:16 field trial achieves Mbit/s-level asymptotic secure key rates, bridging theoretical protocols with Fiber-to-the-Home reality and guiding future scalable access architectures.

quant-ph

Ultra-Large-Capacity Passive Quantum Access Network Powered By Single Thermal Source

Quantum Key Distribution (QKD) provides secure keys for classical communications through one-time-pad (OTP) encryption with physical-law security. Advanced PON-based Classical Access Networks (CANs) support up to 256 users with a total rate of 10 Gbps (10-Gbps @ 256-users). The equivalent rate demand of OTP encryption requires QKD Access Networks (QANs) to reach comparable performance, yet state-of-the-art PON-based QANs remain far from this standard. To address this gap, we propose a passive Thermal-State QAN (TS-QAN) distributing polychromatic quantum randomness from a single thermal source and supporting 304 users with an aggregate secret key rate (SKR) of 13 Gbps (13-Gbps @ 304-users). This performance is enabled by three features. First, broadband thermal states with Bose-Einstein statistics can be represented, through the Glauber-Sudarshan representation, as high-bandwidth Gaussian coherent-state ensembles across frequency modes, eliminating many active modulators and quantum random number generators (QRNGs). Second, Electro-Optic (EO) comb beacons provide time-varying polychromatic phase tracking, so each frequency-mode thermal signal can be coherently measured with a Local Local Oscillator (LLO) aided by its beacon, without large-scale phase-locking networks. Third, state broadcasting allows each user to obtain independent final keys via reverse reconciliation after accounting for residual broadcast-induced correlations, expanding network capacity with small SKR losses. Experimentally, we verify a 13-Gbps @ 304-users TS-QAN using Continuous-Variable QKD (CV-QKD) under covariance-matrix-based network security analysis including multimode Holevo leakage and broadcast correlations. This work meets the SKR and capacity demands from CAN to QAN: 13-Gbps @ 304-users satisfies the 10-Gbps @ 256-users benchmark and provides a scalable solution for modern telecommunication systems.

quant-ph

UniMesh: Unifying 3D Mesh Understanding and Generation

Recent advances in 3D vision have led to specialized models for either 3D understanding (e.g., shape classification, segmentation, reconstruction) or 3D generation (e.g., synthesis, completion, and editing). However, these tasks are often tackled in isolation, resulting in fragmented architectures and representations that hinder knowledge transfer and holistic scene modeling. To address these challenges, we propose UniMesh, a unified framework that jointly learns 3D generation and understanding within a single architecture. First, we introduce a novel Mesh Head that acts as a cross model interface, bridging diffusion based image generation with implicit shape decoders. Second, we develop Chain of Mesh (CoM), a geometric instantiation of iterative reasoning that enables user driven semantic mesh editing through a closed loop latent, prompting, and re generation cycle. Third, we incorporate a self reflection mechanism based on an Actor Evaluator Self reflection triad to diagnose and correct failures in high level tasks like 3D captioning. Experimental results demonstrate that UniMesh not only achieves competitive performance on standard benchmarks but also unlocks novel capabilities in iterative editing and mutual enhancement between generation and understanding. Code: https://github.com/AIGeeksGroup/UniMesh. Website: https://aigeeksgroup.github.io/UniMesh.

cs.CV

PhysInOne: Visual Physics Learning and Reasoning in One Suite

We present PhysInOne, a large-scale synthetic dataset addressing the critical scarcity of physically-grounded training data for AI systems. Unlike existing datasets limited to merely hundreds or thousands of examples, PhysInOne provides 2 million videos across 153,810 dynamic 3D scenes, covering 71 basic physical phenomena in mechanics, optics, fluid dynamics, and magnetism. Distinct from previous works, our scenes feature multiobject interactions against complex backgrounds, with comprehensive ground-truth annotations including 3D geometry, semantics, dynamic motion, physical properties, and text descriptions. We demonstrate PhysInOne's efficacy across four emerging applications: physics-aware video generation, long-/short-term future frame prediction, physical property estimation, and motion transfer. Experiments show that fine-tuning foundation models on PhysInOne significantly enhances physical plausibility, while also exposing critical gaps in modeling complex physical dynamics and estimating intrinsic properties. As the largest dataset of its kind, orders of magnitude beyond prior works, PhysInOne establishes a new benchmark for advancing physics-grounded world models in generation, simulation, and embodied AI.

cs.CV

Weight Group-wise Post-Training Quantization for Medical Foundation Model

Foundation models have achieved remarkable results in medical image analysis. However, its large network architecture and high computational complexity significantly impact inference speed, limiting its application on terminal medical devices. Quantization, a technique that compresses models into low-bit versions, is a solution to this challenge. In this paper, we propose a post-training quantization algorithm, Permutation-COMQ. It eliminates the need for backpropagation by using simple dot products and rounding operations, thereby removing hyperparameter tuning and simplifying the process. Additionally, we introduce a weight-aware strategy that reorders the weight within each layer to address the accuracy degradation induced by channel-wise scaling during quantization, while preserving channel structure. Experiments demonstrate that our method achieves the best results in 2-bit, 4-bit, and 8-bit quantization.

cs.CV

Evidence-Based Actor-Verifier Reasoning for Echocardiographic Agents

Echocardiography plays an important role in the screening and diagnosis of cardiovascular diseases. However, automated intelligent analysis of echocardiographic data remains challenging due to complex cardiac dynamics and strong view heterogeneity. In recent years, visual language models (VLM) have opened a new avenue for building ultrasound understanding systems for clinical decision support. Nevertheless, most existing methods formulate this task as a direct mapping from video and question to answer, making them vulnerable to template shortcuts and spurious explanations. To address these issues, we propose EchoTrust, an evidence-driven Actor-Verifier framework for trustworthy reasoning in echocardiography VLM-based agents. EchoTrust produces a structured intermediate representation that is subsequently analyzed by distinct roles, enabling more reliable and interpretable decision-making for high-stakes clinical applications.

cs.CV

ColoDiff: Integrating Dynamic Consistency With Content Awareness for Colonoscopy Video Generation

Colonoscopy video generation delivers dynamic, information-rich data critical for diagnosing intestinal diseases, particularly in data-scarce scenarios. High-quality video generation demands temporal consistency and precise control over clinical attributes, but faces challenges from irregular intestinal structures, diverse disease representations, and various imaging modalities. To this end, we propose ColoDiff, a diffusion-based framework that generates dynamic-consistent and content-aware colonoscopy videos, aiming to alleviate data shortage and assist clinical analysis. At the inter-frame level, our TimeStream module decouples temporal dependency from video sequences through a cross-frame tokenization mechanism, enabling intricate dynamic modeling despite irregular intestinal structures. At the intra-frame level, our Content-Aware module incorporates noise-injected embeddings and learnable prototypes to realize precise control over clinical attributes, breaking through the coarse guidance of diffusion models. Additionally, ColoDiff employs a non-Markovian sampling strategy that cuts steps by over 90% for real-time generation. ColoDiff is evaluated across three public datasets and one hospital database, based on both generation metrics and downstream tasks including disease diagnosis, modality discrimination, bowel preparation scoring, and lesion segmentation. Extensive experiments show ColoDiff generates videos with smooth transitions and rich dynamics. ColoDiff presents an effort in controllable colonoscopy video generation, revealing the potential of synthetic videos in complementing authentic representation and mitigating data scarcity in clinical settings.

cs.CV

Fault-Tolerant Information Processing with Quantum Weak Measurement

Noise is an important factor that influences the reliability of information acquisition, transmission, processing, and storage. In order to suppress the inevitable noise effects, a fault-tolerant information processing approach via quantum weak measurement is proposed, where pairwise orthogonal postselected measurement bases with various tiny angles and optimal compositions of measured results are chosen as a decoding rule. The signal to be protected can be retrieved with a minimal distortion after having been transmitted through a noisy channel. Demonstrated by typical examples of encoding signal on two-level superposition state or Einstein-Podolsky-Rossen state transmitted through random telegraph noise and decoherence noises channel, the mean squared error distortion may be close to $0$ and the fault-tolerant capability could reach $1$ with finite quantum resources. To verify the availability of the proposed approach, classic coherent light and quantum coherent state are used for encoding information in the experiment. Potentially, the proposed approach may provide a solution for suppressing noise effects in long-distance quantum communication, high-sensitivity quantum sensing, and accurate quantum computation.

quant-ph

RoboCOIN: An Open-Sourced Bimanual Robotic Data Collection for Integrated Manipulation

Despite the critical role of bimanual manipulation in endowing robots with human-like dexterity, large-scale and diverse datasets remain scarce due to the significant hardware heterogeneity across bimanual robotic platforms. To bridge this gap, we introduce RoboCOIN, a large-scale multi-embodiment bimanual manipulation dataset comprising over 180,000 demonstrations collected from 15 distinct robotic platforms. Spanning 16 diverse environments-including residential, commercial, and industrial settings-the dataset features 421 bimanual tasks systematically categorized by 39 bimanual collaboration actions and 432 objects. A key innovation of our work is the hierarchical capability pyramid, which provides granular annotations ranging from trajectory-level concepts to segment-level subtasks and frame-level kinematics. Furthermore, we present CoRobot, an efficient data processing pipeline powered by the Robot Trajectory Markup Language (RTML), designed to facilitate quality assessment, automated annotation, and unified multi-embodiment and data management. Extensive experiments demonstrate the effectiveness of RoboCOIN in enhancing the performance of various bimanual manipulation models across a wide spectrum of robotic embodiments. The entire dataset and codebase are fully open-sourced, providing a valuable resource for advancing research in bimanual and multi-embodiment manipulation.

cs.RO

Robust and cost-effective quantum network using Kramers-Kronig receiver

The quantum internet holds the potential to facilitate applications that are fundamentally inaccessible to the classical internet. Among its most prominent applications is quantum key distribution (QKD) networks, which connect two distant nodes to establish a secure key based on the principles of quantum mechanics. However, the subsequent extensive reliance on interferences in existing QKD protocols leads to the weak robustness of the system and the corresponding network. In this work, we propose a robust and cost-effective quantum network using the Kramers-Kronig receiver. We first propose a continuous-variable QKD protocol based on direct detection without interference, which achieves the recovery of quadrature components through the Kramers-Kronig relation. Subsequently, we have extended this protocol to continuous-variable quantum access networks, further highlighting the robustness and cost advantages of interference-free detection. The experimental results show that each user can achieve a secret key rate at 50 kbit/s within the access network range by using only one photodetector without interference structures. This scheme opens up new possibilities in establishing a robust and cost-effective quantum network, serving as a foundational element in the progress toward establishing a large-scale quantum internet.

quant-ph

Arbitrarily-high-dimensional reconciliation via cross-rotation for continuous-variable quantum key distribution

Multidimensional rotation serves as a powerful tool for enhancing information reconciliation and extending the transmission distance in continuous-variable quantum key distribution (CV-QKD). However, the lack of closed-form orthogonal transformations for high-dimensional rotations has limited the maximum reconciliation efficiency to channels with 8 dimensions over the past decade. This paper presents a cross-rotation scheme to overcome this limitation and enable reconciliation in arbitrarily high dimensions, constrained to even multiples of 8. The key treatment involves reshaping the string vector into matrix form and applying orthogonal transformations to its columns and rows in a cross manner, thereby increasing the reconciliation dimension by one order per cross-rotation while significantly reducing the communication overhead over the classical channel. A rigorous performance analysis is also presented from the perspective of achievable sum-rate. Simulation results demonstrate that 64-dimensional cross-rotation nearly approaches the upper bound, making it a recommended choice for practical implementations.

quant-ph

Joint parameter estimation and multidimensional reconciliation for continuous-variable quantum key distribution

Accurate quantum channel parameter estimation is essential for effective information reconciliation in continuous-variable quantum key distribution (CV-QKD). However, conventional maximum likelihood (ML) estimators rely on a large amount of disclosed data, leading to a significant loss in symbol efficiency. Moreover, the separation between the estimation and reconciliation phases can introduce error propagation. In this paper, we propose a novel joint message-passing scheme that unifies channel parameter estimation and information reconciliation within a Bayesian framework. By leveraging the expectation-maximization (EM) algorithm, the proposed method simultaneously estimates unknown parameters during decoding, eliminating the need for separate ML estimation. Furthermore, we introduce a hybrid multidimensional rotation scheme that removes the requirement for norm feedback, significantly reducing classical channel overhead. To the best of our knowledge, this is the first work to unify multidimensional reconciliation and channel parameter estimation in CV-QKD, providing a practical solution for high-efficiency reconciliation with minimal information disclosure.

quant-ph

Long-distance free-space quantum key distribution with continuous variables

Continuous-variable quantum key distribution (CVQKD) enables remote users to share high-rate and unconditionally secure secret keys while maintaining compatibility with classical optical communication networks and effective resistance against background noise. However, CVQKD experiments have only been demonstrated indoors or over short outdoor distances. Here, by developing channel-fluctuation-independent high-precision manipulation of continuous-variable quantum states, high-accuracy quantum signal acquisition and processing, and high-efficiency free-space acquisition, tracking, and pointing technology, we overcome the excess noise due to atmospheric effects especially in daylight without extra wavelength conversion and spectral filtering, and demonstrate for the first time long-distance free-space quantum key distribution over 7-km inland and 9.6-km maritime atmospheric channels with Gaussian-modulated coherent states. This achieved distribution distance of secure quantum secret keys is well beyond the atmosphere's effective thickness, offering a promising alternative for realizing satellite-based quantum cryptography communication in daylight. Moreover, given that the CVQKD system is naturally compatible with existing ground fiber telecommunication networks, it marks an essential step for realizing integrated air-ground quantum access networks with cross-domain applications.

quant-ph

VAP-Diffusion: Enriching Descriptions with MLLMs for Enhanced Medical Image Generation

As the appearance of medical images is influenced by multiple underlying factors, generative models require rich attribute information beyond labels to produce realistic and diverse images. For instance, generating an image of skin lesion with specific patterns demands descriptions that go beyond diagnosis, such as shape, size, texture, and color. However, such detailed descriptions are not always accessible. To address this, we explore a framework, termed Visual Attribute Prompts (VAP)-Diffusion, to leverage external knowledge from pre-trained Multi-modal Large Language Models (MLLMs) to improve the quality and diversity of medical image generation. First, to derive descriptions from MLLMs without hallucination, we design a series of prompts following Chain-of-Thoughts for common medical imaging tasks, including dermatologic, colorectal, and chest X-ray images. Generated descriptions are utilized during training and stored across different categories. During testing, descriptions are randomly retrieved from the corresponding category for inference. Moreover, to make the generator robust to unseen combination of descriptions at the test time, we propose a Prototype Condition Mechanism that restricts test embeddings to be similar to those from training. Experiments on three common types of medical imaging across four datasets verify the effectiveness of VAP-Diffusion.

cs.CV