arXiv ScienceSearch

arXiv subjects

An He

Publications and source records attributed to An He.

13 recordsLinked to original sources

Beyond Suspicious Steps: Ontological Trust in Long-Horizon Agents

Long-horizon agents increasingly operate across many steps, tools, and observa- tions. In this setting, the relevant oversight question is not only whether each action is locally valid, but whether the evolving trajectory still corresponds to the task the user authorized. Drift can accumulate quietly: an agent may call the right tool with plausible arguments at every step, while its prefix moves toward a broader role, an adjacent objective, or evidence the user never supplied. Existing monitors mostly check local compliance, deliver final-trace verdicts, or score generic risk; they do not directly estimate this prefix-level relation. We introduce ontological trust, a task-conditioned property of trajectory prefixes, and instantiate it as RGE, an online monitor that decomposes trust along Role, Goal, and Evidence. RGE uses LLMs only to derive structured task and step representations; trust-state updates, projec- tions, and intervention decisions are deterministic, so the output is a replayable and auditable trust trajectory rather than a single end-to-end judge verdict. We construct a cross-domain trajectory corpus from OSWorld, FinanceBench, and EICU-AC, covering benign executions, prefix-paired drift, and pseudo-consistency failures. On this corpus, RGE outperforms adapted rule-, judge-, and shield-style baselines on prefix-paired drift detection. With the two larger estimator models, it exceeds 93% Drift F1 on every benchmark while keeping benign coverage at or above 95.8%. Pseudo-consistency is harder: detection depends on whether task completion is externally visible, a structural limit we characterize empirically.

cs.AI

Timestamp-Aware Spatio-Temporal Graph Contrastive Learning for Network Intrusion Detection

Given their effectiveness in modeling the relational structure among network traffic flows, graph neural networks (GNNs) have been widely adopted in network intrusion detection systems (NIDSs). However, most existing GNN-based NIDS approaches focus on the relational structure of traffic flows, and treat them as temporally independent, which limits their ability to cope with evolving attack behaviors. Moreover, their reliance on supervised or semi-supervised learning often restricts generalization to unseen attacks. To address these limitations, we propose a novel self-supervised GNN-based framework. To the best of our knowledge, the proposed model is among the first self-supervised GNN-based NIDS models to explicitly leverage real timestamps, which provides faithful temporal dependencies for representation learning. We first construct a series of temporal graphs from network traffic flows according to their timestamps, and then employ an E-GraphSAGE and LSTM based encoder to fully extract temporal information and spatial dependencies of network traffic, without introducing time-costly attention mechanisms. A multi-view graph contrastive learning (GCL) scheme is introduced, where temporal, spatial, and feature contrasts are jointly performed to capture temporal continuity, preserve structural consistency, and improve the generalization and robustness of the learned representations, respectively. In addition, a gradient-norm-based adaptive weighting strategy is designed to optimize the contrastive loss weights. Experimental results on four representative NIDS datasets with real timestamps demonstrate that our method significantly outperforms existing self-supervised approaches and achieves performance comparable to the supervised state-of-the-art GNN method, while maintaining high computational efficiency.

cs.CR

Deep Photonic Reservoir Computing with On-chip Nonlinearity

Reservoir computing, renowned for its low training cost, has emerged as a promising lightweight paradigm for efficient spatiotemporal processing,it remains challenging to realize deep photonic reservoir computing (DPRC) systems, due to the lack of scalable on-chip nonlinearity. Here, we introduce a versatile time delayed DPRC framework that natively supports deep and concurrent spatiotemporal processing entirely in the optical domain. At its core, the system leverages free carrier dynamics in silicon microring resonators to provide the fundamental nonlinearity and short term memory, and these nonlinear nodes are interconnected through true time delay lines that establish shared long-term memory. Benefiting from intrinsic physical nonlinearity and multi-timescale fading memory, this simple yet effective architecture demonstrates remarkable high dimensional representation capabilities. On the NTU RGB D benchmark, the parameter efficient DPRC system achieves superior action recognition accuracies compared to mainstream deep learning models, while requiring only a single shot regression training procedure. We further verify a prototype DPRC chip that excels across diverse dataset classification and time series prediction tasks. It enables a streamlined all optical pipeline between hierarchical layers, delivering a consistent computational density of 334.25 TOPs/mm2, independent of the reservoir depth and three orders of magnitude higher than conventional approaches. Moreover, its performance scales with near-zero hardware overhead by utilizing additional wavelength channels. This DPRC network is highly scalable on a silicon photonic platform, with flexible extension to hundreds of deep reservoir layers and parallel channels, paving the way toward intelligent optoelectronic systems for advanced real time processing and parallel decision making.

physics.optics

DiffSemanticFusion: Semantic Raster BEV Fusion for Autonomous Driving via Online HD Map Diffusion

Autonomous driving requires accurate scene understanding, including road geometry, traffic agents, and their semantic relationships. In online HD map generation scenarios, raster-based representations are well-suited to vision models but lack geometric precision, while graph-based representations retain structural detail but become unstable without precise maps. To harness the complementary strengths of both, we propose DiffSemanticFusion -- a fusion framework for multimodal trajectory prediction and planning. Our approach reasons over a semantic raster-fused BEV space, enhanced by a map diffusion module that improves both the stability and expressiveness of online HD map representations. We validate our framework on two downstream tasks: trajectory prediction and planning-oriented end-to-end autonomous driving. Experiments on real-world autonomous driving benchmarks, nuScenes and NAVSIM, demonstrate improved performance over several state-of-the-art methods. For the prediction task on nuScenes, we integrate DiffSemanticFusion with the online HD map informed QCNet, achieving a 5.1\% performance improvement. For end-to-end autonomous driving in NAVSIM, DiffSemanticFusion achieves state-of-the-art results, with a 15\% performance gain in NavHard scenarios. In addition, extensive ablation and sensitivity studies show that our map diffusion module can be seamlessly integrated into other vector-based approaches to enhance performance. All artifacts are available at https://github.com/SunZhigang7/DiffSemanticFusion.

cs.CV

Reconfigurable non-Abelian geometric phase in hybrid integrated photonics

The non-Abelian geometric phase possesses the capability of enabling robust and fault-resilient unitary transformations, making it a cornerstone of holonomic quantum computation. This "all-geometric" approach has successfully advanced the manipulation of electrons in condensed matter physics and has sparked growing interest in its implementation within photonics, an area that has traditionally relied on sensitive dynamic phases. However, a major limitation of the topologically protected and inherently robust geometric phase is its lack of reconfigurability. In contrast, mainstream optical computing schemes demand high reconfigurability to compensate for fabrication errors and to support diverse computational tasks. Here, we demonstrate a reconfigurable non-Abelian geometric phase based on the non-volatile phase-change material Sb$_2$Se$_3$. By switching between its crystalline and amorphous states, the number of degenerate subspaces can be actively adjusted. Thus, multilevel second-order matrices and reconfigurable third-order matrices with 3-bit control is realized. For larger reconfigurable rotation angles, tunable braiding operations are also demonstrated. Furthermore, high-dimensional reconfigurable braiding shows promising potential for applications in optical switching. Our results pave the way for the all-geometric-phase-based approach in optical computing.

physics.optics

Observation of generic U(m) non-Abelian holonomy in photonics

Non-Abelian geometric phases form the foundation of fault-tolerant holonomic quantum computation. An "all-geometric" approach leveraging these phases enables robust unitary operations in condensed matter systems. Photonics, with rich degrees of freedom, offer a highly promising platform for non-Abelian holonomy. Yet, achieving universal unitary transformations in photonic holonomy remain elusive. Intrinsic positive real couplings in dissipationless photonic waveguides restrict holonomy to special orthogonal matrices, falling short of universal quantum gates or arbitrary linear operations. Here, we introduce artificial gauge fields (AGFs) to enable complex-valued couplings, expanding photonic holonomy to the full unitary group. We realize generic U(2) transformations and synthesize higher dimensional U(m) operations (up to U(4)) in integrated photonics. Our results open doors toward the transformative "all-geometric-phase" approach in photonic computing in both classical and quantum realms.

physics.optics

High computational density nanophotonic media for machine learning inference

Efficient machine learning inference is essential for the rapid adoption of artificial intelligence across various domains.On-chip optical computing has emerged as a transformative solution for accelerating machine learning tasks, owing to its ultra-low power consumption. However, enhancing the computational density of on-chip optical systems remains a significant challenge, primarily due to the difficulties in miniaturizing and integrating key optical interference components.In this work, we harness the potential of fabrication-constrained scattering optical computing within nanophotonic media to address these limitations.Central to our approach is the use of fabrication-aware inverse design techniques, which enable the realization of manufacturable on-chip scattering structures under practical constraints.This results in an ultra-compact optical neural computing architecture with an area of just 64 um2,representing a remarkable three orders of magnitude reduction in footprint compared to traditional optical neural networks. Our prototype, tested on the Iris flower dataset, achieved an experimental accuracy of 86.7%, closely matching the simulation benchmark.This breakthrough showcases a promising pathway toward ultra-dense, energy-efficient optical processors for scalable machine learning inference, significantly reducing both the hardware footprint, latency, and power consumption of next-generation AI applications.

physics.optics

GHz spiking neuromorphic photonic chip with in-situ training

Neuromorphic photonic computing represents a paradigm shift for next-generation machine intelligence, yet critical gaps persist in emulating the brain's event-driven, asynchronous dynamics,a fundamental barrier to unlocking its full potential. Here, we report a milestone advancement of a photonic spiking neural network (PSNN) chip, the first to achieve full-stack brain-inspired computing on a complementary metal oxide semiconductor-compatible silicon platform. The PSNN features transformative innovations of gigahertz-scale nonlinear spiking dynamics,in situ learning capacity with supervised synaptic plasticity, and informative event representations with retina-inspired spike encoding, resolving the long-standing challenges in spatiotemporal data integration and energy-efficient dynamic processing. By leveraging its frame-free, event-driven working manner,the neuromorphic optoelectronic system achieves 80% accuracy on the KTH video recognition dataset while operating at ~100x faster processing speeds than conventional frame-based approaches. This work represents a leap for neuromorphic computing in a scalable photonic platform with low latency and high throughput, paving the way for advanced applications in real-time dynamic vision processing and adaptive decision-making, such as autonomous vehicles and robotic navigation.

physics.optics

Case studies on time-dependent Ginzburg-Landau simulations for superconducting applications

The macroscopic electromagnetic properties of type II superconductors are primarily influenced by the behavior of microscopic superconducting flux quantum units. Time-dependent Ginzburg-Landau (TDGL) equations provide an elegant and powerful tool for describing and examining both the statics and dynamics of these superconducting entities. They have been instrumental in replicating and elucidating numerous experimental results over the past decades.This paper provides a comprehensive overview of the progress in TDGL simulations, focusing on three key aspects of superconductor applications. The initial section delves into vortex rectification in superconductors described within the TDGL framework. We specifically highlight the superconducting diode effect achieved through asymmetric pinning landscapes and the reversible manipulation of vortex ratchets with dynamic pinning landscapes. The subsequent section reviews the achievements of TDGL simulations concerning the critical current density of superconductors, emphasizing the optimization of pinning sites, particularly vortex pinning and dynamics in polycrystalline Nb$_3$Sn with grain boundaries. The third part concentrates on numerical modeling of vortex penetration and dynamics in superconducting radio frequency (SRF) cavities, including a discussion of superconductor insulator superconductor multilayer structures. In the last section, we present key findings, insights, and perspectives derived from the discussed simulations.

cond-mat.supr-con

Applying Self-supervised Learning to Network Intrusion Detection for Network Flows with Graph Neural Network

Graph Neural Networks (GNNs) have garnered intensive attention for Network Intrusion Detection System (NIDS) due to their suitability for representing the network traffic flows. However, most present GNN-based methods for NIDS are supervised or semi-supervised. Network flows need to be manually annotated as supervisory labels, a process that is time-consuming or even impossible, making NIDS difficult to adapt to potentially complex attacks, especially in large-scale real-world scenarios. The existing GNN-based self-supervised methods focus on the binary classification of network flow as benign or not, and thus fail to reveal the types of attack in practice. This paper studies the application of GNNs to identify the specific types of network flows in an unsupervised manner. We first design an encoder to obtain graph embedding, that introduces the graph attention mechanism and considers the edge information as the only essential factor. Then, a self-supervised method based on graph contrastive learning is proposed. The method samples center nodes, and for each center node, generates subgraph by it and its direct neighbor nodes, and corresponding contrastive subgraph from the interpolated graph, and finally constructs positive and negative samples from subgraphs. Furthermore, a structured contrastive loss function based on edge features and graph local topology is introduced. To the best of our knowledge, it is the first GNN-based self-supervised method for the multiclass classification of network flows in NIDS. Detailed experiments conducted on four real-world databases (NF-Bot-IoT, NF-Bot-IoT-v2, NF-CSE-CIC-IDS2018, and NF-CSE-CIC-IDS2018-v2) systematically compare our model with the state-of-the-art supervised and self-supervised models, illustrating the considerable potential of our method. Our code is accessible through https://github.com/renj-xu/NEGSC.

cs.LG

Monolithic Integration of Embedded III-V Lasers on SOI

Silicon photonic integration has gained great success in many application fields owing to the excellent optical device properties and complementary metal-oxide semiconductor (CMOS) compatibility. Realizing monolithic integration of III-V lasers and silicon photonic components on single silicon wafer is recognized as a long-standing obstacle for ultra-dense photonic integration, which can provide considerable economical, energy efficient and foundry-scalable on-chip light sources, that has not been reported yet. Here, we demonstrate embedded InAs/GaAs quantum dot (QD) lasers directly grown on trenched silicon-on-insulator (SOI) substrate, enabling monolithic integration with butt-coupled silicon waveguides. By utilizing the patterned grating structures inside pre-defined SOI trenches and unique epitaxial method via molecular beam epitaxy (MBE), high-performance embedded InAs QD lasers with out-coupled silicon waveguide are achieved on such template. By resolving the epitaxy and fabrication challenges in such monolithic integrated architecture, embedded III-V lasers on SOI with continuous-wave lasing up to 85 oC are obtained. The maximum output power of 6.8 mW can be measured from the end tip of the butt-coupled silicon waveguides, with estimated coupling efficiency of approximately -7.35 dB. The results presented here provide a scalable and low-cost epitaxial method for realization of on-chip light sources directly coupling to the silicon photonic components for future high-density photonic integration.

physics.optics

Fast dynamic aperture optimization with reversal integration

A fast method for dynamic aperture (DA) optimization of storage rings has been developed through the use of reversal integration. While chaotic dynamical systems have exact time-reversal symmetry, numerical forward integration differs from its reversal due to scaled cumulative round-off errors. The difference, intrinsically associated with the Lyapunov exponent, is a generic indicator of chaos because it represents the sensitivity of chaotic motion to an initial condition. A chaos indicator of the charged particle motion is then obtained by comparing the forward integrations of particle trajectories with corresponding reversals, a.k.a. "backward integrations." The indicator was confirmed to be observable through short-term particle tracking simulations. Therefore, adopting it as an objective function could speed up optimization. The DA of the National Synchrotron Light Source II storage ring, and another test diffraction-limited light source ring, were optimized using this method for the purpose of demonstration.

physics.acc-ph

Ultra-Compact Coupling Structures for Heterogeneously Integrated Silicon Lasers

Due to the inherent in-direct bandgap nature of Silicon, heterogeneous integration of semiconductor lasers on Silicon on Insulator (SOI) is crucial for next-generation on-chip optical interconnects. Compact, high-efficient and high-tolerant couplers between III-V light source and silicon chips have been the challenge for photonic integrated circuit (PIC). Here, we redesign the taper adiabatic coupler with the total coupling length of only 4 {\mu}m, and propose another two novel slot coupler and bridge-SWG coupler with both coupling length of 7 {\mu}m, to heterogeneously integrate III-V lasers and silicon chips. We study theoretically the optical mode coupling process through the redesigned taper coupler, the final coupling results match well with the simulation in 3D-FDTD. The three compact couplers represent fundamental TE mode coupling efficiencies all over 90%, even 95.7% for bridge-SWG coupler, to the best of our knowledge, are also the shortest coupling structures (7 um). Moreover, these coupling structures also possess excellent fabrication tolerance.

physics.optics