arXiv ScienceSearch

arXiv subjects

Zhenhua Li

Publications and source records attributed to Zhenhua Li.

At least 19 recordsLinked to original sources

SignDino: Self-Supervised Sign Language Representation Learning via Temporal-Axis Self-Distillation

Self-supervised sign language representation learning must model two properties not central to natural-image SSL: signs are produced by a small set of anatomically distinct articulators, and their meaning depends on the temporal organisation of those articulators. We introduce SignDino, a self-supervised sign-video encoder that moves the DINOv3 student--teacher recipe from the spatial domain of image crops to the temporal domain of tracked sign streams. Each video is decomposed into left-hand, right-hand, and face streams by a detector-first YOLOv8n+ByteTrack pipeline. A frozen DINOv3 ViT-B/16 embeds each per-frame anatomical crop, while lightweight temporal Transformers, not the image backbone, form the student and EMA teacher. They are trained by temporal DINO self-distillation, frame-level masked-token prediction in the style of iBOT, KoLeo feature spreading, and Gram anchoring of the frame-to-frame similarity structure. This design keeps strong image-level visual primitives fixed and learns only how articulator states evolve across time. We evaluate on sign-to-English translation, isolated sign recognition, and fingerspelling detection benchmarks. Across these tasks, SignDino provides a strong public self-supervised representation and shows competitive or state-of-the-art performance under matched downstream evaluation.

cs.CV

SignNet-1M: Large-Scale Multilingual Sign Language Video Dataset with Downstream Benchmarks

Sign language models are typically trained on datasets captured under constrained conditions, with limited viewpoint, background, and signer-identity diversity, leading to poor robustness under real-world distribution shifts. We introduce SignNet-1M, a large-scale augmented dataset spanning ASL, CSL, and German Sign Language (DGS). SignNet-1M synthesizes realistic variations along three axes: (i) novel-view rendering (rotation and zoom) via 3D Gaussian Splatting (3DGS), (ii) scene/identity editing via diffusion models for background replacement and signer substitution while preserving sign motion and linguistic content, and (iii) post-rendering augmentations that emulate capture and compression artifacts (e.g., pose/temporal perturbations and video-level corruptions) to better match in-the-wild recordings. Beyond data release, we provide a unified benchmark suite across downstream tasks (e.g., translation and recognition) and ablations that isolate each augmentation component. Experiments across backbones show that training with SignNet-1M consistently improves generalization under cross-view, cross-background, cross-identity, and post-rendering shifts, while maintaining strong in-distribution performance. The dataset, full augmentation pipeline, and benchmark are available at https://signnet.chatsign.ai/.

cs.CV

Parallel distributed quantum gates for dual-species quantum emitters

We propose a parallel protocol for implementing distributed nonlocal quantum gates between spatially separated stationary qubits encoded in dual-species quantum emitters (i.e., color-center and superconducting qubits). By utilizing entangled photon pairs with distinct frequencies as a quantum data bus, our approach connects spatially separated devices without requiring quantum frequency conversion or preshared entanglement, while maintaining an always-ready and resource-efficient property for distributed quantum computing and networks. Furthermore, we demonstrate the feasibility of implementing parallel distributed nonlocal quantum gates on multiple pairs of spatially separated qubits using a single high-dimensional entangled photon pair, which directly benefits from the enhanced quantum capacity provided by optical qudit encoding. Our protocol establishes a scalable and practically implementable framework for distributed quantum networks, potentially enabling the development of future large-scale quantum computing architectures.

quant-ph

FUSCO: High-Performance Distributed Data Shuffling via Transformation-Communication Fusion

Large-scale Mixture-of-Experts (MoE) models rely on \emph{expert parallelism} for efficient training and inference, which splits experts across devices and necessitates distributed data shuffling to route each token to its assigned experts. However, existing communication libraries handle this shuffling poorly; its overhead can account for over half of end-to-end runtime. We present FUSCO, an MoE-friendly communication library that achieves efficient and lightweight data shuffling through fused data transformation and communication, based on the key observation that MoE's expert-major data layout conflicts with the device-major layout expected by communication operations. FUSCO captures the fine-grained data layout, which is then interpreted by a pipelined communication engine that performs the required shuffling efficiently along the communication path. Lightweight planning and load-balancing mechanisms complement the engine by eliminating redundant communication and dispersing traffic. Evaluations on representative benchmarks illustrate that FUSCO achieves up to 3.84$\times$ and 2.01$\times$ speedups over NCCL and DeepEP (the state-of-the-art MoE communication library), respectively. In end-to-end MoE tasks, compared to NCCL and DeepEP, FUSCO reduces the training latency by 1.17-1.39$\times$ and 1.10-1.19$\times$, and lowers the first-token generation latency in inference by 1.09-1.25$\times$ and 1.06-1.16$\times$.

cs.DC

Experimental demonstration of scalable quantum cryptographic conferencing

Quantum network enables a variety of quantum information processing tasks, where multi-user quantum communication is one of the important objectives. Quantum cryptographic conferencing serves as an essential solution to establish secure keys to realize secure multi-user communications. However, existing QCC implementations have been fundamentally limited by the low probability of multi-user coincidence detection to measure or construct the Greenberger-Horne-Zeilinger (GHZ) entangled state. In this work, we report the experimental realization of QCC eliminating the need for coincidence detection, where the GHZ state is constructed by correlating detection events occurring within the coherence time, thereby greatly enhancing the success probability of GHZ-state measurement. Meanwhile, to establish and maintain high-visibility GHZ measurement among three independent users, we developed a three-party phase compensation scheme combined with precise temporal and polarization alignment within a time-bin-phase encoding framework. Furthermore, we designed an efficient pairing strategy to simplify subsequent data processing and enhance processing efficiency. Based on these techniques, we successfully performed QCC experiments over total channel losses of 66.3 dB, corresponding to 331.5 km of commercial fiber (0.2 dB/km), achieving secure key rates of 5.4 bit/s, whereas previous QCC experiments have been limited to 100 km. The results surpass the multi-user repeaterless bound in quantum networks, establishing a new regime of scalable, multi-user quantum communication and paving the way for metropolitan quantum networks.

quant-ph

Quantum queer superalgebra and its integral form

In this paper, we introduce quantum root vectors for the quantum queer superalgebra ${\boldsymbol U}_{\!{v}}({\mathfrak q_n})$ via a braid-group action, compute their complete commutation relations, and construct a PBW-type basis for the Lusztig integral form ${{\boldsymbol U}_{v,\mathcal{ Z}}}$. This yields an explicit presentation of ${{\boldsymbol U}_{v,\mathcal{ Z}}}$ and provides a way to understand the structure of the quantum queer superalgebra at roots of unity.

math.QA

In situ calibration of camera-refraction interface based on analytical refractive imaging equation

Camera calibration is an essential process in photogrammetry, serving as a crucial link between the 2D image coordinate system and the 3D world coordinate system. However, when observations are conducted through refractive interfaces, the refraction effects at these interfaces render traditional calibration methods ineffective, significantly compromising measurement accuracy. To address this challenge, we propose a novel camera calibration method based on the analytical refractive imaging (ARI) equation. The ARI method facilitates accurate estimation of camera parameters from distorted images and enables in-situ joint calibration of both the camera and the refractive interface. The experimental results indicate that the proposed method reduces the error to only 10% of that produced by conventional ray-tracing (RT) method. Moreover, while maintaining comparable computational accuracy and efficiency, it effectively mitigates the local convergence issues that may arise in the polynomial fitting (PF) approach. Finally, reconstruction experiments further confirm the accuracy of the proposed method. Experimental results demonstrate that the proposed method outperforms existing refractive calibration techniques in terms of accuracy while maintaining high precision in 3D reconstruction tasks.

physics.optics

STAlloc: Enhancing Memory Efficiency in Large-Scale Model Training with Spatio-Temporal Planning

The rapid scaling of large language models (LLMs) has significantly increased GPU memory pressure, which is further aggravated by training optimization techniques such as virtual pipeline and recomputation that disrupt tensor lifespans and introduce considerable memory fragmentation. Such fragmentation stems from the use of online GPU memory allocators in popular deep learning frameworks like PyTorch, which disregard tensor lifespans. As a result, this inefficiency can waste as much as 43% of memory and trigger out-of-memory errors, undermining the effectiveness of optimization methods. To address this, we introduce STAlloc, a GPU memory allocator for deep learning frameworks that reduces fragmentation by exploiting the spatial and temporal regularity in memory allocation behaviors of training workloads. STAlloc introduces a novel paradigm that combines offline planning with online allocation. The offline planning leverages spatio-temporal regularities to generate a near-optimal allocation plan, while the online allocation handles complex and dynamic models such as Mixture-of-Experts (MoE). Built as a pluggable PyTorch memory allocator, STAlloc reduces fragmentation ratio on average by 85.1% (up to 100%) across both dense and MoE models, with negligible overhead. This enables more efficient, high-throughput training configurations and improves throughput performance by up to 32.5%.

cs.LG

The Regular Representation of the twisted queer $q$-Schur Superalgebra

We study the representation theory of the quantum queer superalgebra ${U_{\lcase{v}}(\mathfrak{\lcase{q}}_{n})}$ and obtain some properties of the highest weight modules. Furthermore, based on the realization of ${U_{\lcase{v}}(\mathfrak{\lcase{q}}_{n})}$, we study the representation theory of the twisted queer $q$-Schur superalgebra ${{\widetilde{\mathcal{Q}}}_{\lcase{v}}(\lcase{n},\lcase{r})}$, and obtain the decomposition of its regular module as a direct sum of irreducible submodules, which also means ${{\widetilde{\mathcal{Q}}}_{\lcase{v}}(\lcase{n},\lcase{r})}$ is semisimple.

math.QA

Asymmetric protocols for mode pairing quantum key distribution with finite-key analysis

The mode pairing quantum key distribution (MP-QKD) protocol has attracted considerable attention for its capability to ensure high secure key rates over long distances without requiring global phase locking. However, ensuring symmetric channels for the MP-QKD protocol is challenging in practical quantum communication networks. Previous studies on the asymmetric MP-QKD protocol have relied on ideal decoy state assumptions and infinite-key analysis, which are unattainable for real-world deployment. In this paper, we conduct a security analysis of asymmetric MP-QKD protocol with the finite-key analysis, where we discard the previously impractical assumptions made in the decoy-state method. Combined with statistical fluctuation analysis, we globally optimized the 12 independent parameters in the asymmetric MP-QKD protocol by employing our modified particle swarm optimization. The simulation results demonstrate that our work can achieve significantly enhanced secure key rates and transmission distances compared to the original strategy with adding extra attenuation. We further investigate the relationship between the intensities and probabilities of signal, decoy, and vacuum states with transmission distance, facilitating its more efficient deployment in future quantum networks.

quant-ph

Constructing the quantum queer supergroup using Hecke-Clifford superalgebras

In [DGLW], we use certain special elements and their commutation relations in the Hecke-Clifford algebras $H^c_{r,R}$ to derive some fundamental multiplication formulas associated with the natural bases in queer $q$-Schur superalgebras $Q_q(n,r;R)$ introduced in [DW2]. Here a natural basis element is defined by a special element $T_{A^{\star}}$ in $H^c_{r,R}$ associated with a pair of certain $n\times n$ matrices $A^{\star}=(A^{\bar0}|A^{\bar1})$ over $\mathbb{N}$ with entries sum to $r$. The definition of $T_{A^\star}$ consists of an element $c_{A^{\star}}$ in the Clifford superalgebra and an element $T_A$ in the Hecke algebra, where $A=A^{\bar0}+A^{\bar1}$. Note that all $T_A$ can be used to define the natural basis for the corresponding $q$-Schur algebra $S_q(n,r)$. This paper is a continuation of [DGLW]. We start with standardized queer $v$-Schur superalgebras $ Q^s_v(n,r)$, for $R=\mathbb{Z}[v,v^{-1}]$ and $q=v^2$, and their natural bases. With the $v$-Schur algebra ${ S}_v(n,r)$ at the background, the first key ingredient is a standardisation of the natural basis for $Q^s_v(n,r)$ and their associated standard multiplication formulas. By introducing some long elements of finite sums, we then extend the formulas to these long elements which allow us to explicitly define $\mathbb{Q}(v)$-superalgebra homomorphisms $\xi_{n,r}$ from the quantum queer supergroup $\boldsymbol{U}_v(\mathfrak{q}_n)$ to queer $q$-Schur superalgebras $\boldsymbol{Q}^s_v(n,r)$, for all $r\geq1$. Finally, taking limits of long elements yields certain infinitely long elements as formal infinite series which eventually lead to a new construction for $\boldsymbol{U}_v(\mathfrak{q}_n)$.

math.QA

Measurement-device-independent quantum-secret-sharing networks with linear Bell-state analysis

Quantum secret sharing (QSS) plays a pivotal role in multiparty quantum communication, ensuring the secure distribution of private information among multiple parties. However, the security of QSS schemes can be compromised by attacks exploiting imperfections in measurement devices. Here, we propose a reconfigurable approach to implement QSS based on measurement-device-independent (MDI) principles, utilizing linear two-photon Bell state analysis.By employing single-qubit conjugate operations for encoding private classical information, our approach offers reconfigurability, allowing for the inclusion of additional parties without sacrificing efficiency. Furthermore, we demonstrate the robust security of our MDI-QSS scheme against inter-eavesdropping by dishonest participants and establish lower bounds for secure communication among three legitimate parties. This work presents a flexible configuration for implementing multiparty secure quantum communication with imperfect measurement devices and represents a significant advancement in the development of secure quantum communication technologies.

quant-ph

RADS-Checker: Measuring Compliance with Right of Access by the Data Subject in Android Markets

The latest data protection regulations worldwide, such as the General Data Protection Regulation (GDPR), have established the Right of Access by the Data Subject (RADS), granting users the right to access and obtain a copy of their personal data from the data controllers. This clause can effectively compel data controllers to handle user personal data more cautiously, which is of significant importance for protecting user privacy. However, there is currently no research systematically examining whether RADS has been effectively implemented in mobile apps, which are the most common personal data controllers. In this study, we propose a compliance measurement framework for RADS in apps. In our framework, we first analyze an app's privacy policy text using NLP techniques such as GPT-4 to verify whether it clearly declares offering RADS to users and provides specific details on how the right can be exercised. Next, we assess the authenticity and usability of the identified implementation methods by submitting data access requests to the app. Finally, for the obtained data copies, we further verify their completeness by comparing them with the user personal data actually collected by the app during runtime, as captured by Frida Hook. We analyzed a total of 1,631 apps in the American app market G and the Chinese app market H. The results show that less than 54.50% and 37.05% of apps in G and H, respectively, explicitly state in their privacy policies that they can provide users with copies of their personal data. Additionally, in both app markets, less than 20% of apps could truly provide users with their data copies. Finally, among the obtained data copies, only about 2.94% from G pass the completeness verification.

cs.CR

Measurements of the hyperfine structure of $nP_J$ Rydberg states by microwave spectroscopy in Cs atoms

We present measurements of hyperfine structure (HFS) of the $nP_J$ Rydberg states for large principal quantum number $n$ range ($n=41-55$) employing the microwave spectroscopy in an ultra-cold cesium Rydberg ensemble. A microwave field with 30-$\mu$s duration couples the $ nS \to nP $ transition, yielding a narrow linewidth spectroscopy that approaches the Fourier limit, which allows us to resolve the hyperfine structure of $ nP_J $ states. By analyzing the hyperfine splittings of $nP_J$ states, we determine the magnetic-dipole HFS coupling constant $\bar{A}_{HFS,P_{1/2}}=3.760(26) ~$GHz for $P_{1/2}$ state, $\bar{A}_{HFS,P_{3/2}}= 0.718(27)~$GHz, and $ \bar{B}_{HFS,P_{3/2}}= -0.084(102)~$GHz for $P_{3/2}$ state, respectively. Systematic uncertainties caused by stray electromagnetic field, microwave field power and Rydberg interaction are analyzed. This measurement is significant for the investigation of Rydberg electrometry and quantum simulation with dipole interaction involving $nP_J$ state.

physics.atom-ph

Continuous broadband Rydberg receiver using AC Stark shifts and Floquet States

We demonstrate the continuous broadband microwave receivers based on AC Stark shifts and Floquet States of Rydberg levels in a cesium atomic vapor cell. The resonant transition frequency of two adjacent Rydberg states 78$S_{1/2}$ and 78$P_{1/2}$ is tuned based on AC Stark effect of 70~MHz Radio frequency (RF) field that is applied outside the vapor cell. Meanwhile, the Rydberg states also exhibit Floquet even-order sidebands that are used to extend the bandwidths further. We achieve microwave electric field measurements over 1.172~GHz of continuous frequency range. The sensitivity of the Rydberg receiver with heterodyne technique in the absence of RF field is 280.2~nVcm$^{-1}$Hz$^{-1/2}$, while it is dramatically decreased with tuning the resonant transition frequency in the presence of RF field. Surprisingly, the sensitivity can be greatly improved if the microwave field couples the Floquet sideband transition. The achieving of continuous frequency and high sensitivity microwave detection will promote the application of Rydberg receiver in the radar technique and wireless communication.

physics.atom-ph

Braid Group Action and Quantum Queer Superalgebra

In this paper, we present explicit actions of braid group on the universal enveloping superalgebra ${\boldsymbol U}(\mathfrak{{q}}_n)$ and the quantum queer superalgebra ${\boldsymbol U}_{\!{v}}(\mathfrak{{q}}_{n})$. Then we provide a new definition of root vectors and some explicit expression for them. With these procedures, we obtain the PBW-type basis containing the product of root vectors.

math.QA

High Spectral-Efficiency, Ultra-low MIMO SDM Transmission over a Field-Deployed Multi-Core OAM Fiber

Few-mode multi-core fiber (FM-MCF) based Space-Division Multiplexing (SDM) systems possess the potential to maximize the number of multiplexed spatial channels per fiber by harnessing both the space (fiber cores) and mode (optical mode per core) dimensions. However, to date, no SDM transmissions over field-deployed FM-MCFs in realistic outdoor settings have been reported, which contrasts with SDM schemes demonstrated using single-mode multi-core fibers (SM-MCFs) installed in practical fiber cable ducts. In this paper, we present the successful demonstration of bidirectional SDM transmission over a 5-km field-deployed seven ring-core fiber (7-RCF) with a cladding diameter of 178 ${\mu}$m, achieving a Spectral Efficiency (SE) of 2$\times$201.6 bit/s/Hz. This work establishes a new record for the highest SE attained in SDM demonstrations utilizing field-deployed fiber cables, achieving an approximate 10x increase compared to the SE of reported field-deployed optical fiber cable transmission systems. Notably, these results are realized through the utilization of small-scale modular 4$\times$4 multiple-input multiple-output (MIMO) processing with a time-domain equalization (TDE) tap number not exceeding 15, maintaining a complexity per unit capacity comparable to that of MIMO equalization in SDM demonstrations employing weakly coupled SM-MCF cables. These results underscore the significant potential for achieving heightened SE and expanding capacity per individual fiber using SDM techniques in practical applications.

cs.NI

Scalable and Versatile Linear Computation with Minimalistic Photonic Matrix Processor

The advancement of artificial intelligence demands flexible multimodal data processing with high throughput and energy efficiency. Photonic integrated circuits (PIC) has demonstrated promising potentials in terms of low latency and low power consumption per operation for linear operations such as matrix-vector multiplication. However, the existing schemes face challenges in their scalability due to the use of photonic circuits that expand with the scale of the operants, despite efforts of exploiting the multiple optical parameter dimensions such as time, wavelength and spatial parallelism. They also lacked flexibility and efficiency in switching between different types of operations or tasks and adapting to multimodal data. In this article, we introduce an optical matrix processor (MP) with a minimalistic recursive structure for both multiplications and accumulations. The MP consists of an eletro-optic ring-modulator implemented as a thin-film lithium niobate PIC that allows flexible configurability and time-division multiplexed scheduling. The MP supports not only versatile linear operations including vector/matrix-vector multiplication and single/multi-kernel convolution but also ultrafast task switching and adaptability to data of different sizes, by simply adjusting the data baud rate relative to the ring delay without structural modifications. We demonstrate its capabilities in a optic-electronic convolutional neural network with a computing throughput up to 73.4 billion operations per second. The MP further supports high scalability through appropriate allocation of wavelength and space resources,extending computing parallelism to handle higher data volumes with higher energy efficiency. This novel scheme paves the way for a new class of photonic processorscapable of managing escalating data workloads with unprecedented flexibility, efficiency and scalability.

physics.optics