arXiv ScienceSearch

arXiv subjects

Zhiyang Wang

Publications and source records attributed to Zhiyang Wang.

At least 19 recordsLinked to original sources

Size Transferability of Graph Transformers with Convolutional Positional Encodings

Transformers have achieved remarkable success across domains, motivating the rise of Graph Transformers (GTs) as attention-based architectures for graph-structured data. A key design choice in GTs is the use of Graph Neural Network (GNN)-based positional encodings to incorporate structural information. In this work, we study GTs through the lens of manifold limit models for graph sequences and establish a theoretical connection between GTs with GNN positional encodings and Manifold Neural Networks (MNNs). Building on transferability results for GNNs under manifold convergence, we show that GTs inherit transferability guarantees from their positional encodings. In particular, GTs trained on small graphs provably generalize to larger graphs under mild assumptions. We complement our theory with extensive experiments on standard graph benchmarks, demonstrating that GTs exhibit scalable behavior on par with GNNs. To further show the efficiency in a real-world scenario, we implement GTs for shortest path distance estimation over terrains to better illustrate the efficiency of the transferable GTs. Our results provide new insights into the understanding of GTs and suggest practical directions for efficient training of GTs in large-scale settings.

cs.LG

Hidden Crossover and Relaxor-Like Response from Emerging Polar Skyrmion Correlations in Ferroelectric Superlattices

Polar skyrmions in ferroelectric superlattices are nanoscale topological polarization textures typically regarded as weakly coupled objects confined to individual layers, with a role secondary to that of the underlying symmetry-breaking order parameter. Here using large-scale phase-field simulations of ferroelectric superlattices, we uncover a hidden thermal crossover deep inside the ferroelectric phase, where polar skyrmions evolve from an uncorrelated, layer-resolved state into an interlayer-correlated ensemble. This crossover occurs without additional symmetry breaking or a new order parameter, but produces a pronounced broad peak in the dielectric susceptibility. The anomaly originates from the competition between correlation-enhanced response, associated with the growth of interlayer skyrmion correlations, and polarization-induced stiffness, which suppresses dielectric fluctuations at low temperature. Under AC driving, the peak shifts with frequency, resembling relaxor ferroelectrics despite the absence of quenched disorder or polar nanoregions. Our results establish a disorder-free route to relaxor-like dielectric response and identify topological defect correlations as an organizing principle for thermodynamic anomalies, providing a mechanism distinct from conventional critical behavior associated with symmetry breaking and divergent order-parameter fluctuations.

cond-mat.mtrl-sci

Scalability of Graph Neural Network Policies in Wireless Communication Networks

Graph Neural Networks (GNNs) offer scalable solutions for wireless resource allocation, yet existing formal performance guarantees across varying scales for spatial and sparse graphs do not directly apply to these settings. Scalability theories rely on dense graphon limits or continuous manifold approximations, both of which fail under the sparse, bounded-degree regimes and Euclidean metric constraints of physical wireless environments. This paper establishes a theoretical framework for GNN transferability over sparse Random Geometric Graphs (RGGs), capturing distance-dependent channel decay and spatial interference. We model sparse RGGs as spatial perturbations of regular Deterministic Grid Graphs (DGGs) and employ spatial windowing operators to compare networks across differing scales. Assuming Lipschitz continuity of GNN architectures and signal stationarity, we prove scalability over DGGs and bound same-scale transferability between DGGs and RGGs. Combining these results establishes formal scalability bounds across sparse RGG topologies. Finally, we extend this framework to conflict graph models, deriving equivalent scalability guarantees for link-level resource allocation policies. We verify our results for scalability with two experiment settings: power allocation and wireless link scheduling. The simulations show GNNs exhibit the expected scalable behavior, and analyze the relevance of our theoretical assumptions in practical deployment.

eess.SP

Persuasion Index: A Theory-Guided Framework for Persuasion Analysis

Identifying persuasive rhetorical cues is critical across domains, from detecting information manipulation and improving AI safety to advancing public health communication. We propose the Persuasion Index (PI), a taxonomy of 15 dimensions grounded in persuasion theories from psychology and communication, and one transparent implementation using 55 sub-features built from lexicons and rule-based detectors. The taxonomy is modular: individual detectors can be replaced while preserving the theoretical structure. We evaluate PI on four public datasets for English argumentative text that vary in domain, style, and outcome measures, and show that PI provides a shared feature space for interpreting rhetorical patterns associated with persuasion-related outcomes. Linear models show that PI features carry meaningful predictive signal while remaining computationally lightweight. Dimension-level analyses reveal recurring associations between PI dimensions and persuasion outcomes across datasets, while also highlighting topic- and stance-specific variation. We release PI as an open-source package and web interface for principled and auditable analysis of human and AI-mediated communication.

cs.CL

RareLens: Towards End-to-End Rare Disease Care via Aligning Divergent Large Language Model Reasoning

Rare diseases represent one of the most challenging settings for clinical decision-making, where heterogeneous presentations, sparse evidence and limited expertise create persistent uncertainty throughout the care pathway. Although artificial intelligence could help, existing systems largely address isolated tasks, particularly diagnosis, and usually rely on downstream investigations rather than information available at initial presentation. Here we show that clinical AI performance under uncertainty can be improved not by scaling a single model, but by exploiting the diversity of multiple imperfect reasoning systems. Across heterogeneous large language models, we identify divergent reasoning trajectories with complementary error patterns and develop RareLens, which learns to reconcile these perspectives into actionable decisions across four stages of rare disease care: risk screening, diagnosis, treatment planning and prognosis prediction. Built on RarelensBench, a real-world dataset of 157,525 cases spanning all 33 Orphanet categories and more than 7,000 conditions, RareLens outperformed every frontier model tested, including GPT-5, DeepSeek-R1, Claude-3.7-Sonnet and Gemini-2.5-Pro, across all stages. It achieved an area under the curve of 0.917 for screening and top-1 accuracies of 65.5% and 89.8% for diagnosis and treatment. In an external evaluation involving 1,287 cases and 23 physicians, autonomous RareLens and physicians assisted by RareLens both outperformed unaided physicians, while demonstrating that effective human-AI collaboration requires more than simply providing model outputs. These findings establish divergent model reasoning as an exploitable source of information and suggest a general strategy for building AI systems that operate reliably under high clinical uncertainty.

cs.AI

Dual-Faraday-laser-pumped cesium beam clock with $7.7\times 10^{-13}/\sqrtτ$ frequency stability

Compact cesium beam clocks are major frequency references for deployable timing systems. However, further improvement of their short-term frequency stability is limited by the clock signal-to-noise ratio (SNR). Although two-laser optical pumping can increase the effective atomic utilization, the achievable clock SNR has long been limited by laser-induced frequency-to-amplitude noise conversion. Here, we demonstrate a compact dual-Faraday-laser-pumped (DFP) Cs beam clock enabled by a low-frequency-noise atom-referenced laser architecture. The intracavity Faraday anomalous dispersion optical filter provides inherent alignment to the Cs D$_2$ resonances, while modulation transfer spectroscopy offers suppressed frequency noise and drift. The resulting laser system supports robust turnkey operation with a Lorentzian linewidth of 2.12 kHz. The DFP Cs clock achieves a clock SNR of 46,365 in a 1-Hz bandwidth and a fractional Allan deviation of $7.7\times 10^{-13}/\sqrtτ$ , with Hadamard deviation reaching $7.7\times 10^{-15}$ at 10,000 s. This work pushes the fractional frequency stability of a compact Cs beam clock into the $10^{-13}/\sqrtτ$ regime, providing a pathway toward high-performance Cs frequency references for field-deployable precision timing, navigation, and synchronization.

physics.atom-ph

Long-Horizon Wireless Link Scheduling with State-Augmented Graph Neural Networks

We address optimal link scheduling in large-scale wireless networks. The goal is to schedule transmissions over a time horizon so that to maximize sum rate while ensuring that average rates of each customer attain a minimum rate requirement. To this end, we formulate a constrained optimization problem and solve it using Lagrangian duality. Common primal-dual approaches lead to time invariant policies. Our constraint requires all links transmit a fraction of the time while avoiding interference, which calls for time-varying policies across time slots. We propose an iterative algorithm to sample optimal sequences of schedules and dual variables. The scheduling decisions are parameterized using a Graph Neural Network. We incorporate state-augmentation techniques to learn said parameterization, introducing dual variables as dynamic inputs to the policy. This augmentation enables the GNN to adapt scheduling decisions over time, balancing constraint satisfaction with performance maximization. We validate our approach through extensive numerical simulations, benchmarking against several baselines and considering varying constraint levels.

eess.SP

OpFlow: Learning Opportunity-Conditioned Choice Potentials for Robust OD Flow Prediction

Origin-destination (OD) flow prediction is central to urban analytics, yet deep models trained on raw counts remain vulnerable to distribution shift. The core problem is that raw count supervision cannot distinguish transferable choice mechanisms from environment-specific shortcuts. Raw OD count mixes two objects: how much demand an origin produces and how that demand is allocated across destinations. We argue that the transferable object is the exposure-to-choice law that maps spatial conditions to relative destination preferences. We propose OpFlow, a mechanism-constrained framework that learns row-centered choice potentials and reconstructs flows by combining the induced allocation with a separately calibrated origin scale. Under distribution shift, spatial exposures and the induced allocations are allowed to vary; what transfers is the conditional map from exposure states to relative choice potentials. Theoretically, we characterize the identifiable row-centered potential and show that classical spatial interaction laws are restricted log-potential cases. Controlled synthetic shifts and a real-world experiment show OpFlow improves robustness under environment shifts.

cs.LG

Graph Learning Should Move Beyond Restrictive Views of Spectral and Message-Passing GNNs

Graph neural networks (GNNs) are commonly divided into message-passing neural networks (MPNNs) and spectral GNNs, reflecting two largely separate research traditions in machine learning and signal processing. While MPNNs have a precise definition, there is no widely accepted criterion for what makes a mapping a spectral GNN. Most existing work restricts spectral GNNs to layered architectures based on linear spectral filters. Under this restriction, we show that spectral and spatial GNNs have largely equivalent expressive power. To promote progress in the field, we propose a precise definition of spectral GNNs based on eigenbasis symmetries, in contrast to the definition of MPNNs via neighborhood permutation symmetries. We further argue that the two perspectives offer complementary strengths. MPNNs provide a natural language for discrete structure and expressivity analysis through tools from logic and graph isomorphism, while the spectral perspective offers principled tools for understanding smoothing, bottlenecks, stability, and community structure. Overall, we argue that progress in graph learning will be accelerated by clarifying the similarities and differences between these perspectives and by moving toward a unified theoretical framework.

cs.LG

Limit Analysis of Graph Neural Networks with Wireless Conflict Graphs

Graph Neural Networks (GNNs) have emerged as a powerful tool for wireless resource allocation that leverages the underlying graph structure of communication networks. Their transferability property enables models trained on small-scale graphs to generalize to large-scale deployments with little performance deterioration, a desirable property for currently growing networks. Wireless networks are sparse regimes, where a single node is connected to a small number of other users. This work establishes theoretical results for transferability of GNNs over graphs derived from sparse Random Geometric Graphs (RGGs). In particular, we focus on conflict graphs of RGGs used to model interference among links. Our approach considers the closeness between RGGs and Deterministic Grid Graphs (DGG) to establish bounds in the performance loss when a model is transferred across scales. We validate our theoretical findings through the problem of link scheduling, demonstrating that our learned policies consistently outperform existing benchmarks at scale. Finally, we examine the impact of our theoretical assumptions on empirical performance.

cs.LG

Graph Semi-Supervised Learning for Point Classification on Data Manifolds

We propose a graph semi-supervised learning framework for classification tasks on data manifolds. Motivated by the manifold hypothesis, we model data as points sampled from a low-dimensional manifold $\mathcal{M} \subset \mathbb{R}^F$. The manifold is approximated in an unsupervised manner using a variational autoencoder (VAE), where the trained encoder maps data to embeddings that represent their coordinates in $\mathbb{R}^F$. A geometric graph is constructed with Gaussian-weighted edges inversely proportional to distances in the embedding space, transforming the point classification problem into a semi-supervised node classification task on the graph. This task is solved using a graph neural network (GNN). Our main contribution is a theoretical analysis of the statistical generalization properties of this data-to-manifold-to-graph pipeline. We show that, under uniform sampling from $\mathcal{M}$, the generalization gap of the semi-supervised task diminishes with increasing graph size, up to the GNN training error. Leveraging a training procedure which resamples a slightly larger graph at regular intervals during training, we then show that the generalization gap can be reduced even further, vanishing asymptotically. Finally, we validate our findings with numerical experiments on image classification benchmarks, demonstrating the empirical effectiveness of our approach.

cs.LG

Graph Neural Networks in Large Scale Wireless Communication Networks: Scalability Across Random Geometric Graphs

The growing complexity of wireless systems has accelerated the move from traditional methods to learning-based solutions. Graph Neural Networks (GNNs) are especially well-suited here, since wireless networks can be naturally represented as graphs. A key property of GNNs is transferability: models trained on one graph often generalize to much larger graphs with little performance loss. While empirical studies have shown that GNN-based wireless policies transfer effectively, existing theoretical guarantees do not capture this phenomenon. Most works focus on dense graphs where node degrees scale with network size, an assumption that fails in wireless systems. In this work, we provide a formal theoretical foundation for transferability on Random Geometric Graphs (RGGs), a sparse and widely used model of wireless networks. We further validate our results through numerical experiments on power allocation, a fundamental resource management task.

eess.SP

A Manifold Perspective on the Statistical Generalization of Graph Neural Networks

Graph Neural Networks (GNNs) extend convolutional neural networks to operate on graphs. Despite their impressive performances in various graph learning tasks, the theoretical understanding of their generalization capability is still lacking. Previous GNN generalization bounds ignore the underlying graph structures, often leading to bounds that increase with the number of nodes -- a behavior contrary to the one experienced in practice. In this paper, we take a manifold perspective to establish the statistical generalization theory of GNNs on graphs sampled from a manifold in the spectral domain. As demonstrated empirically, we prove that the generalization bounds of GNNs decrease linearly with the size of the graphs in the logarithmic scale, and increase linearly with the spectral continuity constants of the filter functions. Notably, our theory explains both node-level and graph-level tasks. Our result has two implications: i) guaranteeing the generalization of GNNs to unseen data over manifolds; ii) providing insights into the practical design of GNNs, i.e., restrictions on the discriminability of GNNs are necessary to obtain a better generalization performance. We demonstrate our generalization bounds of GNNs using synthetic and multiple real-world datasets.

cs.LG

Generalization of Geometric Graph Neural Networks with Lipschitz Loss Functions

In this paper, we study the generalization capabilities of geometric graph neural networks (GNNs). We consider GNNs over a geometric graph constructed from a finite set of randomly sampled points over an embedded manifold with topological information captured. We prove a generalization gap between the optimal empirical risk and the optimal statistical risk of this GNN, which decreases with the number of sampled points from the manifold and increases with the dimension of the underlying manifold. This generalization gap ensures that the GNN trained on a graph on a set of sampled points can be utilized to process other unseen graphs constructed from the same underlying manifold. The most important observation is that the generalization capability can be realized with one large graph instead of being limited to the size of the graph as in previous results. The generalization gap is derived based on the non-asymptotic convergence result of a GNN on the sampled graph to the underlying manifold neural networks (MNNs). We verify this theoretical result with experiments on multiple real-world datasets.

eess.SP

Wireless Link Scheduling with State-Augmented Graph Neural Networks

We consider the problem of optimal link scheduling in large-scale wireless ad hoc networks. We specifically aim for the maximum long-term average performance, subject to a minimum transmission requirement for each link to ensure fairness. With a graph structure utilized to represent the conflicts of links, we formulate a constrained optimization problem to learn the scheduling policy, which is parameterized with a graph neural network (GNN). To address the challenge of long-term performance, we use the state-augmentation technique. In particular, by augmenting the Lagrangian dual variables as dynamic inputs to the scheduling policy, the GNN can be trained to gradually adapt the scheduling decisions to achieve the minimum transmission requirements. We verify the efficacy of our proposed policy through numerical simulations and compare its performance with several baselines in various network settings.

eess.SP

Investigating the $4D_{3/2}|3,\pm2\rangle$--$4D_{5/2}|3,\pm2\rangle$ transition in Nb$^{4+}$ for a THz atomic clock

In this work, the $4D_{3/2}|3,\pm2\rangle \rightarrow 4D_{5/2}|3,\pm2\rangle$ transition in the Nb$^{4+}$ ion is identified as a promising candidate for a terahertz (THz) atomic clock, with the transition frequency occurring at 56.0224 THz. This transition is primarily driven by the magnetic dipole decay channel, which can easily be accessed by a laser. We focus on the stable $^{93}$Nb isotope, which has 100\% natural abundance and a nuclear spin of $I=9/2$ for experimental advantage. Our data analysis allows us to estimate potential systematic shifts in the proposed clock system, including those due to blackbody radiation, electric quadrupole, second-order Zeeman, and second-order Doppler {shifts}. {The scheme presented in this study can help suppress the AC Stark and electric quadrupole shifts in the clock frequency measurement.} {All these analyses} suggest that the proposed THz atomic clock using Nb$^{4+}$ could be valuable in both quantum thermometry and frequency metrology.

physics.atom-ph

Velocity-comb modulation transfer spectroscopy

Sub-Doppler laser spectroscopy is a crucial technique for laser frequency stabilization, playing a significant role in atomic physics, precision measurement, and quantum communication. However, recent efforts to improve frequency stability appear to have reached a bottleneck, as they primarily focus on external technical approaches while neglecting the fundamental issue of low atomic utilization (< 1%), caused by only near-zero transverse velocity atoms involved in the transition. Here, we propose a velocity-comb modulation transfer spectroscopy (MTS) solution that takes advantage of the velocity-selective resonance effect of multi-frequency comb lasers to enhance the utilization of non-zero-velocity atoms. In the probe-pump configuration, each pair of counter-propagating lasers interacts with atoms from different transverse velocity-comb groups, independently contributing to the spectral amplitude and signal-to-noise ratio. Preliminary proof-of-principle results show that the frequency stability of the triple-frequency laser is optimized by nearly a factor of \sqrt{3} compared to the single-frequency laser, consistent with theoretical expectations. With more frequency comb components, MTS-stabilized lasers are expected to achieve order-of-magnitude breakthroughs in frequency stability, taking an important step toward next-generation compact optical clocks. This unique method can also be widely applied to any quantum system with a wide velocity distribution, inspiring innovative advances in numerous fields with a fresh perspective.

physics.atom-ph

A Novel Adaptive Fine-Tuning Algorithm for Multimodal Models: Self-Optimizing Classification and Selection of High-Quality Datasets in Remote Sensing

We propose an adaptive fine-tuning algorithm for multimodal large models. The core steps of this algorithm involve two stages of truncation. First, the vast amount of data is projected into a semantic vector space, and the MiniBatchKMeans algorithm is used for automated clustering. This classification ensures that the data within each cluster exhibit high semantic similarity. Next, we process the data in each cluster, calculating the translational difference between the original and perturbed data in the multimodal large model's vector space. This difference serves as a generalization metric for the data. Based on this metric, we select the data with high generalization potential for training. We applied this algorithm to train the InternLM-XComposer2-VL-7B model on two 3090 GPUs using one-third of the GeoChat multimodal remote sensing dataset. The results demonstrate that our algorithm outperforms the state-of-the-art baselines. various baselines. The model trained on our optimally chosen one-third dataset, based on experimental validation, exhibited only 1% reduction in performance across various remote sensing metrics compared to the model trained on the full dataset. This approach significantly preserved general-purpose capabilities while reducing training time by 68.2%. Furthermore, the model achieved scores of 89.86 and 77.19 on the UCMerced and AID evaluation datasets, respectively, surpassing the GeoChat dataset by 5.43 and 5.16 points. It only showed a 0.91-point average decrease on the LRBEN evaluation dataset.

cs.CV