arXiv ScienceSearch

arXiv subjects

Xiaochen Liu

Publications and source records attributed to Xiaochen Liu.

At least 19 recordsLinked to original sources

Q-BridgeNet: A Quantization Network for Cross-Lingual Sign Language Translation

Most sign language translation (SLT) methods focus on isolated native sign-spoken pairs (e.g., American Sign Language - English). Extending language-specific SLT models to multilingual translation would improve accessibility by enabling communication across diverse sign and spoken language communities. However, existing multilingual SLT approaches still struggle to learn a unified model that minimizes cross-lingual conflicts while capturing shared cross-lingual semantics and preserving language-specific variations across different sign languages. Therefore, we propose Q-BridgeNet, a unified framework for multilingual SLT that jointly mitigates cross-lingual conflicts across both the sign language and spoken language sides. On the sign language side, Q-BridgeNet learns discrete Q-units via adaptive segmentation and residual vector quantization: a shared base codebook provides language-agnostic semantic primitives, while language-specific residual codebooks refine heterogeneous signing semantics. On the spoken language side, a multilingual LLM is fine-tuned to operate in the Q-unit space, leveraging cross-lingual priors to enable a unified SLT model. Experiments on PHOENIX14T, How2Sign, and CSL-Daily show that Q-BridgeNet effectively mitigates cross-lingual conflicts, achieving state-of-the-art performance on native sign-spoken pairs while also demonstrating strong generalization to non-native pairs. Our source code is publicly available at: https://github.com/FengLiQ/Q-BridgeNet

cs.CL

Modeling and Modulation Optimization for OWC Limited by Electronic and Photonic Bandwidth

In contrast to radio frequency (RF), where the modulation bandwidth is restricted by regulations to avoid interference, the available bandwidth in optical wireless communication (OWC) is primarily constrained by system components. To investigate their frequency characteristics, we review the bandwidth limitations of components in the PHY layer of OWC links. Such limitations typically contribute to a decay in the frequency profile of the gain-to-noise ratio (GNR), which can be modeled by a pole-zero transfer function that is generally low-pass. To boost performance, we optimize the signal power spectral density (PSD) of DC-biased optical orthogonal frequency-division multiplexing (DCO-OFDM) which allows for modulation beyond the 3-dB end-to-end bandwidth. We express the Lagrangian-optimized throughput versus the maximum modulation frequency, for an M-zero N-pole low-pass GNR optical link. For optimization implementation, we compare a novel Newton-based algorithm with a newly accelerated version of the Hughes-Hartogs (HH) algorithm, to find the (near-) optimal signal spectrum for theoretical, measured and simulated GNR responses. As demonstrated numerically, employing the proposed multi-stage response model for optimization improves performance in dealing with the successive bandwidth limitations.

eess.SP

Universal Quartic Scaling Law for Kerr-Type Interactions: Projection-Law Factorization Across Nonlinear Quantum Platforms

We present a rigorous derivation and numerical validation of a universal projection-law factorization for quartic nonlinear coupling rates across physically distinct platforms. The central result is that observable Kerr-type interactions -- self-Kerr, cross-Kerr, and cross-phase modulation -- factorize into a dimensionless projection coefficient and an intrinsic quartic energy scale. This structure follows from canonical quantization of a quartic interaction projected onto a finite normal-mode basis. We state this result as a formal proposition with clearly enumerated assumptions, provide a proof sketch based on second quantization of the anharmonic potential, and enumerate the domain of validity. We then implement the scaling law as a lightweight computational toolkit (UEFT-Designer) with platform-specific kernels for superconducting circuits, photonic microcavities, and epsilon-near-zero (ENZ) structures. A complete, step-by-step worked example for a superconducting quarton device -- with full uncertainty propagation -- predicts a cross-Kerr rate $\chi/2\pi = 361\pm 13\,\mathrm{MHz}$, agreeing with the independently measured value of $366\pm 0.5\,\mathrm{MHz}$ \cite{Ye2025} to within 1.4\%. A formal error-propagation analysis establishes that the dominant uncertainty is the Josephson-energy extraction uncertainty ($\sim\!2\%$), not the geometric projection factor. Cross-platform validation against five independent experiments spanning eight orders of magnitude in coupling strength confirms the universality of the factorization to within reported experimental uncertainties. The ghost-sector spectral correction introduced in earlier versions is reframed as an optional phenomenological self-energy model with no mandatory role in the validated scaling law.

quant-ph

Generative Modeling in Protein Design: Neural Representations, Conditional Generation, and Evaluation Standards

Generative modeling has become a central paradigm in protein research, extending machine learning beyond structure prediction toward sequence design, backbone generation, inverse folding, and biomolecular interaction modeling. However, the literature remains fragmented across representations, model classes, and task formulations, making it difficult to compare methods or identify appropriate evaluation standards. This survey provides a systematic synthesis of generative AI in protein research, organized around (i) foundational representations spanning sequence, geometric, and multimodal encodings; (ii) generative architectures including $\mathrm{SE}(3)$-equivariant diffusion, flow matching, and hybrid predictor-generator systems; and (iii) task settings from structure prediction and de novo design to protein-ligand and protein-protein interactions. Beyond cataloging methods, we compare assumptions, conditioning mechanisms, and controllability, and we synthesize evaluation best practices that emphasize leakage-aware splits, physical validity checks, and function-oriented benchmarks. We conclude with critical open challenges: modeling conformational dynamics and intrinsically disordered regions, scaling to large assemblies while maintaining efficiency, and developing robust safety frameworks for dual-use biosecurity risks. By unifying architectural advances with practical evaluation standards and responsible development considerations, this survey aims to accelerate the transition from predictive modeling to reliable, function-driven protein engineering.

cs.LG

Molecular Dynamics Simulations Reveal PolyQ-Length-Dependent Conformational Changes in Huntingtin Exon-1: Implications for Environmental Co-Solvent Modulation of Aggregation-Prone States

Huntington's disease (HD) is caused by CAG-repeat expansion in HTT, which lengthens the polyglutamine (polyQ) tract in huntingtin (HTT) and promotes misfolding and aggregation. While polyQ-length-dependent aggregation is well established, the atomistic conformational dynamics preceding aggregation remain less defined. Here we perform all-atom molecular dynamics simulations of HTT exon-1 constructs containing the N17 domain, polyQ tracts of clinically relevant lengths (Q21, wildtype; Q40, adult onset threshold; Q70, juvenile onset), and the polyproline (polyP) region. Multi-copy simulations (four chains) were run for 100 ns in explicit SPC/E water using the OPLS-AA force field. We quantified radius of gyration (Rg), solvent-accessible surface area (SASA), root-mean-square deviation (RMSD), and intra-protein hydrogen bonds as proxies for conformational expansion and aggregation propensity. PolyQ expansion drove progressive increases in Rg and SASA, consistent with more extended, solvent-exposed ensembles. We further tested organic co-solvents (methanol, hexane, trichloroethylene; 0.5 to 1.0 M), which modulated these landscapes in a solvent-dependent manner. Trichloroethylene induced marked expansion in Q21 and Q40, whereas methanol produced mild compaction in Q21. To our knowledge, this is the first MD study to systematically examine co-solvent effects on HTT exon-1 conformational dynamics. Although limited sampling precludes definitive mechanistic conclusions, the observed trends suggest that hydrophobic co-solvents can bias HTT exon-1 toward more expanded ensembles, motivating computational studies of gene-environment modulation in HD.

cs.CE

DSFlow: Dual Supervision and Step-Aware Architecture for One-Step Flow Matching Speech Synthesis

Flow-matching models have enabled high-quality text-to-speech synthesis, but their iterative sampling process during inference incurs substantial computational cost. Although distillation is widely used to reduce the number of inference steps, existing methods often suffer from process variance due to endpoint error accumulation. Moreover, directly reusing continuous-time architectures for discrete, fixed-step generation introduces structural parameter inefficiencies. To address these challenges, we introduce DSFlow, a modular distillation framework for few-step and one-step synthesis. DSFlow reformulates generation as a discrete prediction task and explicitly adapts the student model to the target inference regime. It improves training stability through a dual supervision strategy that combines endpoint matching with deterministic mean-velocity alignment, enforcing consistent generation trajectories across inference steps. In addition, DSFlow improves parameter efficiency by replacing continuous-time timestep conditioning with lightweight step-aware tokens, aligning model capacity with the significantly reduced timestep space of the discrete task. Extensive experiments across diverse flow-based text-to-speech architectures demonstrate that DSFlow consistently outperforms standard distillation approaches, achieving strong few-step and one-step synthesis quality while reducing model parameters and inference cost.

cs.SD

CryptoBench: A Dynamic Benchmark for Expert-Level Evaluation of LLM Agents in Cryptocurrency

This paper introduces CryptoBench, the first expert-curated, dynamic benchmark designed to rigorously evaluate the real-world capabilities of Large Language Model (LLM) agents in the uniquely demanding and fast-paced cryptocurrency domain. Unlike general-purpose agent benchmarks for search and prediction, professional crypto analysis presents specific challenges: \emph{extreme time-sensitivity}, \emph{a highly adversarial information environment}, and the critical need to synthesize data from \emph{diverse, specialized sources}, such as on-chain intelligence platforms and real-time Decentralized Finance (DeFi) dashboards. CryptoBench thus serves as a much more challenging and valuable scenario for LLM agent assessment. To address these challenges, we constructed a live, dynamic benchmark featuring 50 questions per month, expertly designed by crypto-native professionals to mirror actual analyst workflows. These tasks are rigorously categorized within a four-quadrant system: Simple Retrieval, Complex Retrieval, Simple Prediction, and Complex Prediction. This granular categorization enables a precise assessment of an LLM agent's foundational data-gathering capabilities alongside its advanced analytical and forecasting skills. Our evaluation of ten LLMs, both directly and within an agentic framework, reveals a performance hierarchy and uncovers a failure mode. We observe a \textit{retrieval-prediction imbalance}, where many leading models, despite being proficient at data retrieval, demonstrate a pronounced weakness in tasks requiring predictive analysis. This highlights a problematic tendency for agents to appear factually grounded while lacking the deeper analytical capabilities to synthesize information.

cs.CL

Unified Effective Field Theory for Nonlinear and Quantum Optics

Predicting phenomena that mix few-photon quantum optics with strong field nonlinear optics is hindered by the use of separate theoretical formalisms for each regime. We close this gap with a unified effective field theory valid for frequencies lower than the material-dependent cutoff set by the band gap, plasma frequency, or similar scale. The action couples the electromagnetic gauge field to vector polarisation modes. An isotropic potential generates the optical susceptibilities, while a higher-dimension axion-like term captures magnetoelectric effects; quantisation on the Schwinger-Keldysh contour with doubled BRST ghosts preserves gauge symmetry in dissipative media. One-loop renormalisation-group equations reproduce the measured dispersion of the third-order susceptibility from terahertz to near-visible frequencies after matching a single datum per material. Real-time dynamics solved with a matrix-product-operator engine yield good agreement with published results for GaAs polariton cavities, epsilon-near-zero indium-tin-oxide films and superconducting quarton circuits. The current formulation is limited to these 1-D geometries and sub-cut-off frequencies; higher-dimensional or above-cut-off phenomena will require additional degrees of freedom or numerical methods.

physics.optics

A Unified Spectrum for Turbulence in Microfluidic Flow

We present a predictive master spectrum describing turbulence-like flows in microfluidic systems. Extending Pao's viscous-range closure, the model introduces (i) an adaptive inertial-range slope dependent on measurable dimensionless numbers and (ii) a physics-specific cutoff that captures entropy-producing sinks such as electrokinetic forcing, compliant walls, active stresses, and interfacial tension. This formulation unifies turbulence regimes -- electrokinetic, active, interfacial, and compressible -- within one compact expression. Comparison with reported data reproduces both spectral slopes and dissipation cutoffs while requiring only global observables (velocity, viscosity, Taylor microscale, and forcing strength). The framework provides a design-level predictive tool for turbulent microflows prior to computationally heavy DNS or CFD.

physics.flu-dyn

FOSS solution for Molecular Dynamics Simulation Automation and Collaboration with MDSGAT

The process of setting up and successfully running Molecular Dynamics Simulations (MDS) is outlined to be incredibly labour and computationally expensive with a very high barrier to entry for newcomers wishing to utilise the benefits and insights of MDS. Here, presented, is a unique Free and Open-Source Software (FOSS) solution that aims to not only reduce the barrier of entry for new Molecular Dynamics (MD) users, but also significantly reduce the setup time and hardware utilisation overhead for even highly experienced MD researchers. This is accomplished through the creation of the Molecular Dynamics Simulation Generator and Analysis Tool (MDSGAT) which currently serves as a viable alternative to other restrictive or privatised MDS Graphical solutions with a unique design that allows for seamless collaboration and distribution of exact MD simulation setups and initialisation parameters through a single setup file. This solution is designed from the start with a modular mindset allowing for additional software expansion to incorporate numerous extra MDS packages and analysis methods over time

cs.CE

ETimeline: An Extensive Timeline Generation Dataset based on Large Language Model

Timeline generation is of great significance for a comprehensive understanding of the development of events over time. Its goal is to organize news chronologically, which helps to identify patterns and trends that may be obscured when viewing news in isolation, making it easier to track the development of stories and understand the interrelationships between key events. Timelines are now common in various commercial products, but academic research in this area is notably scarce. Additionally, the current datasets are in need of refinement for enhanced utility and expanded coverage. In this paper, we propose ETimeline, which encompasses over $13,000$ news articles, spanning $600$ bilingual timelines across $28$ news domains. Specifically, we gather a candidate pool of more than $120,000$ news articles and employ the large language model (LLM) Pipeline to improve performance, ultimately yielding the ETimeline. The data analysis underscores the appeal of ETimeline. Additionally, we also provide the news pool data for further research and analysis. This work contributes to the advancement of timeline generation research and supports a wide range of tasks, including topic generation and event relationships. We believe that this dataset will serve as a catalyst for innovative research and bridge the gap between academia and industry in understanding the practical application of technology services. The dataset is available at https://zenodo.org/records/11392212

cs.IR

The disrupting and growing open cluster spiral arm patterns of the Milky Way

Star clusters provide unique advantages for investigating Galactic spiral arms, particularly due to their precise ages, positions, and kinematic properties, which are further enhanced by ongoing updates from the astrometric data. In this study, we employ the latest extensive catalogue of open clusters from Gaia DR3 to examine the positional deviations of clusters belonging to different age groups. Additionally, we employ dynamical simulations to probe the evolutionary behavior of spiral arm positions. Our analysis reveals an absence of a theoretical age pattern in the spiral arms traced by open clusters, and the pattern speeds of the spiral arms are consistent with the rotation curve. Both of these results do not align with the predictions of quasi-stationary density wave theory, suggesting a more dynamic or transient arm scenario for the Milky Way. From this perspective, combined with vertex deviation estimates, it appears that the Local arm is in a state of growth. In contrast, the Sagittarius-Carina arm and the Perseus arm exhibit opposing trends. Consequently, we speculate that the Galactic stellar disk does not exhibit a grand-design spiral pattern with a fixed pattern speed, but rather manifests as a multi-armed structure with arms that continuously emerge and dissipate.

astro-ph.GA

DeRisk: An Effective Deep Learning Framework for Credit Risk Prediction over Real-World Financial Data

Despite the tremendous advances achieved over the past years by deep learning techniques, the latest risk prediction models for industrial applications still rely on highly handtuned stage-wised statistical learning tools, such as gradient boosting and random forest methods. Different from images or languages, real-world financial data are high-dimensional, sparse, noisy and extremely imbalanced, which makes deep neural network models particularly challenging to train and fragile in practice. In this work, we propose DeRisk, an effective deep learning risk prediction framework for credit risk prediction on real-world financial data. DeRisk is the first deep risk prediction model that outperforms statistical learning approaches deployed in our company's production system. We also perform extensive ablation studies on our method to present the most critical factors for the empirical success of DeRisk.

cs.LG

Survey for Distant Stellar Aggregates in Galactic Disk: Detecting Two Thousand Star Clusters and Candidates, along with the Dwarf Galaxy IC10

Despite having data for over 10^9 stars from Gaia, only less than 10^4 star clusters and candidates have been discovered. Particularly, distant star clusters are rarely identified, due to the challenges posed by heavy extinction and great distance. However, Gaia data has continued to improve, enabling even fainter cluster members to be distinguished from field stars. In this work, we will introduce a star cluster search method based on the DBSCAN algorithm; we have made improvements to make it better suited for identifying clusters on dimmer and more distant stars. After removing member stars of known Gaia-based clusters, we have identified 2086 objects with |b|<10 deg, of which 1488 are highly reliable open star clusters, along with 569 candidates, 28 globular cluster candidates and 1 irregular galaxy IC 10 at low Galactic latitudes. We found that the proper motion of IC10 is similar yet slightly different from the water maser observations, which is an important result for the comparison with Gaia and VLBA. Besides, when compared with the star clusters appearing in Gaia DR2/EDR3, we have found nearly three times as many new objects above a distance of 5 kpc, including hundreds of them above Av > 5 mag. And it has enabled us to detect a higher number of old clusters, over a billion years old, that are difficult to detect due to observational limitations. Our findings significantly expand the remote cluster sample and enhance our understanding of the limits of Gaia DR3 data in stellar aggregates research. The full figure set for 2085 clusters can be seen in \url{https://nadc.china-vo.org/res/r101258/}

astro-ph.GA

Unveiling hidden stellar aggregates in the Milky Way: 1656 new star clusters found in Gaia EDR3

We report 1,656 new star clusters found in the Galactic disk (|b|<20 degrees) beyond 1.2 kpc, using Gaia EDR3 data. Based on an unsupervised machine learning algorithm, DBSCAN, and followed our previous studies, we utilized a unique method to do the data preparation and obtained the clustering coefficients, which proved to be an effective way to search blindly for star clusters. We tabulated the physical parameters and member stars of the new clusters, and presented some interesting examples, including a globular cluster candidate. The cluster parameters and member stars are available at CDS via anonymous ftp to https://cdsarc.cds.unistra.fr/ftp/vizier.submit//he22c. We examined the new discoveries and discussed their statistical properties. The proper motion dispersions and radii of the new clusters were the same as the previously reported ones. The new star clusters beyond 1.2 kpc were older than those in the solar neighborhood, and the new objects found in the third Galactic quadrant presented the lowest line-of-sight extinctions. Combined with our previous results, the total population of new clusters detected through our method was 2,541, corresponding to 55% of all newly published clusters in the Gaia era. The number of cataloged Gaia star clusters was also increased to nearly six thousand. In the near future, it is necessary to make a unified confirmation and member star determination for all reported clusters.

astro-ph.GA

A Blind All-sky Search for Star Clusters in Gaia EDR3: 886 Clusters within 1.2 kpc of the Sun

Although previous searches for star clusters have been very successful, many clusters are likely still omitted, especially at high Galactic latitude regions. In this work, based on the astrometry of Gaia EDR3, we searched nearby (parallax > 0.8 mas) all-sky regions, obtaining 886 star clusters, of which 270 candidates have not been cataloged before. At the same time, we have presented the physical parameters of the clusters by fitting theoretical isochrones to their optical magnitudes. More halo members and expanding structures in many star clusters were also found. Most of the new objects are young clusters that are less than 100 million years old. Our work greatly increased the sample size and physical parameters of star clusters in the solar neighborhood, in particular, 46 clusters are newly found with |b| > 20 deg, which represents an increase of nearly three fold of cluster numbers at high Galactic latitude regions. The cluster parameters and member stars are available at CDS via https://cdsarc.u-strasbg.fr/ftp/vizier.submit//hezh22b/, and the cluster figure sets are available via https://doi.org/10.12149/101133.

astro-ph.GA

PSP: Pre-trained Soft Prompts for Few-Shot Abstractive Summarization

Few-shot abstractive summarization has become a challenging task in natural language generation. To support it, we designed a novel soft prompts architecture coupled with a prompt pre-training plus fine-tuning paradigm that is effective and tunes only extremely light parameters. The soft prompts include continuous input embeddings across an encoder and a decoder to fit the structure of the generation models. Importantly, a novel inner-prompt placed in the text is introduced to capture document-level information. The aim is to devote attention to understanding the document that better prompts the model to generate document-related content. The first step in the summarization procedure is to conduct prompt pre-training with self-supervised pseudo-data. This teaches the model basic summarizing capabilities. The model is then fine-tuned with few-shot examples. Experimental results on the CNN/DailyMail and XSum datasets show that our method, with only 0.1% of the parameters, outperforms full-model tuning where all model parameters are tuned. It also surpasses Prompt Tuning by a large margin and delivers competitive results against Prefix-Tuning with 3% of the parameters.

cs.CL

AutoCast: Scalable Infrastructure-less Cooperative Perception for Distributed Collaborative Driving

Autonomous vehicles use 3D sensors for perception. Cooperative perception enables vehicles to share sensor readings with each other to improve safety. Prior work in cooperative perception scales poorly even with infrastructure support. AutoCast enables scalable infrastructure-less cooperative perception using direct vehicle-to-vehicle communication. It carefully determines which objects to share based on positional relationships between traffic participants, and the time evolution of their trajectories. It coordinates vehicles and optimally schedules transmissions in a distributed fashion. Extensive evaluation results under different scenarios show that, unlike competing approaches, AutoCast can avoid crashes and near-misses which occur frequently without cooperative perception, its performance scales gracefully in dense traffic scenarios providing 2-4x visibility into safety critical objects compared to existing cooperative perception schemes, its transmission schedules can be completed on the real radio testbed, and its scheduling algorithm is near-optimal with negligible computation overhead.

cs.NI