arXiv ScienceSearch

arXiv subjects

Yinhao Xu

Publications and source records attributed to Yinhao Xu.

7 recordsLinked to original sources

Autonomous Chaotic Time Series Prediction using Physical Neuromorphic Networks

Physical reservoir computing (PRC) with neuromorphic networks offers a promising approach to brain-inspired information processing, exploiting emergent nonlinear dynamics of physical neural networks as a computational resource. This study demonstrates fully autonomous closed-loop prediction of the Mackey--Glass (MG) chaotic time series using a simulated neuromorphic nanowire network as the physical reservoir. Two strategies are evaluated: the virtual node (VN) method, which expands the feature space by temporal multiplexing of reservoir states, and a non-VN approach that uses all physical node readouts directly without temporal multiplexing. Results are reported for two values of the MG time delay parameter, $\tau = 18$ and $\tau = 21$, the latter representing a more complex chaotic regime not previously evaluated for this class of physical reservoir. Over a short prediction horizon of $T = 100$ timesteps, the VN approach achieves autonomous prediction accuracies of $90.4$% and $89.7$% at $\tau = 18$ and $\tau = 21$, respectively, while the non-VN approach achieves $81.5$% and $76.2$%. Long-horizon analysis over $T = 500$ timesteps shows that both approaches reproduce the qualitative attractor structure and dominant spectral content of the true MG signal, with trajectories remaining bounded throughout. These results suggest that the intrinsic dynamics of neuromorphic nanowire networks are sufficient to support meaningful autonomous chaotic time series prediction without virtual node augmentation, and that performance may improve further as physical network sizes scale to the millions of nodes achievable in hardware. As this study uses simulated networks, extrapolation to physically fabricated large-scale arrays remains to be validated experimentally.

cond-mat.dis-nn

Quantifying Error Tolerance in Synthetic Data: An Atomic-level Operand vs. Operator Perturbation Study

Synthetic data generation has become a cornerstone for advancing large language models. However, the lack of the quantitative analysis for error tolerance became a critical bottleneck. Consequently, current filtering strategies fluctuate between two extremes: they are either overly aggressive, risking the exclusion of potentially valuable samples, or overly permissive, failing to eliminate erroneous samples effectively. To bridge this gap, this paper introduces Atomic Tree Operation Modeling (ATOM), a framework that decomposes data into functional units ($f(x)\rightarrow y$). ATOM distinguishes benign Operand $x$ perturbations from fatal Operator $f$ perturbations. The former are needlessly discarded by aggressive filtering, while the latter slip through permissive filtering. Our experiments reveal a double dissociation: models are robust to operand perturbations but collapse under operator perturbations. By prioritizing operator over aggressive operand precision, our ATOM-synthesized data outperforms rigorous baselines (e.g., +3.1% gain over LIMA), suggesting that operator diversity matters more than operand precision. Our code is available at https://github.com/Lut-hub/ATOM.

cs.CL

LongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction Synthesis

Synthesizing long-context supervised fine-tuning (SFT) data is a scalable way to enhance the long-context understanding of large language models (LLMs), yet existing approaches share three limitations: narrow task coverage, insufficient instruction difficulty, and a lack of faithfulness supervision. We propose \textbf{LongCrafter}, a structured synthesis framework that couples a hierarchical task taxonomy with an evidence-grounded pipeline. The taxonomy organizes long-context understanding into local/shallow and global/deep levels and yields 32 fine-grained task types that serve as a global generative prior. Guided by this taxonomy, LongCrafter constructs task-aligned long contexts, decomposes them into explicit evidence graphs that model cross-paragraph dependencies, and generates instruction--response pairs strictly grounded in the located evidence spans, ensuring both controllable difficulty and faithful, traceable reasoning. Models fine-tuned on LongCrafter data outperform all SFT baselines and even the official post-trained models on LongBench, LongBench~v2, and LooGLE across both Qwen2.5-7B and LLaMA-3.1-8B, with the largest gains on high-difficulty tasks. Further analysis shows that LongCrafter data is more diverse and better spread across difficulty levels, and that the trained models locate evidence robustly regardless of position, effectively mitigating the ``lost in the middle'' problem.

cs.CL

Intrinsic Neuro-Synaptic Spiking Dynamics and Resonance in Memristive Networks

Self-organizing memristive networks are physical circuits that dynamically reconfigure their circuitry in response to external input signals. Their adaptive behavior arises from intrinsic neuro-synaptic dynamics combined with a heterogeneous network topology. In this work, we demonstrate that such networks naturally generate neuronal population spiking dynamics similar to those observed in biological neuronal systems. This study investigates the intrinsic and emergent dynamics of memristive networks mathematically and numerically for both DC and AC input signals. Nonlinear spike-like features are maximized when the frequency of the input driving signal matches the network's intrinsic dynamical timescale, where nonlinear resonance is observed. Furthermore, the optimal frequency for computation is found to be the maximal frequency before the onset of resonance.

cond-mat.dis-nn

Learning Chaotic Dynamics with Neuromorphic Network Dynamics

This study investigates how dynamical systems may be learned and modelled with a neuromorphic network which is itself a dynamical system. The neuromorphic network used in this study is based on a complex electrical circuit comprised of memristive elements that produce neuro-synaptic nonlinear responses to input electrical signals. To determine how computation may be performed using the physics of the underlying system, the neuromorphic network was simulated and evaluated on autonomous prediction of a multivariate chaotic time series, implemented with a reservoir computing framework. Through manipulating only input electrodes and voltages, optimal nonlinear dynamical responses were found when input voltages maximise the number of memristive components whose internal dynamics explore the entire dynamical range of the memristor model. Increasing the network coverage with the input electrodes was found to suppress other nonlinear responses that are less conducive to learning. These results provide valuable insights into how a physical neuromorphic network device can be feasibly optimised for learning complex dynamical systems using only external control parameters.

cond-mat.dis-nn

Dynamic Reservoir Computing with Physical Neuromorphic Networks

Reservoir Computing (RC) with physical systems requires an understanding of the underlying structure and internal dynamics of the specific physical reservoir. In this study, physical nano-electronic networks with neuromorphic dynamics are investigated for their use as physical reservoirs in an RC framework. These neuromorphic networks operate as dynamic reservoirs, with node activities in general coupled to the edge dynamics through nonlinear nano-electronic circuit elements, and the reservoir outputs influenced by the underlying network connectivity structure. This study finds that networks with varying degrees of sparsity generate more useful nonlinear temporal outputs for dynamic RC compared to dense networks. Dynamic RC is also tested on an autonomous multivariate chaotic time series prediction task with networks of varying densities, which revealed the importance of network sparsity in maintaining network activity and overall dynamics, that in turn enabled the learning of the chaotic Lorenz63 system's attractor behavior.

cs.ET

Multi-Scale Feature Fusion Transformer Network for End-to-End Single Channel Speech Separation

Recently studies on time-domain audio separation networks (TasNets) have made a great stride in speech separation. One of the most representative TasNets is a network with a dual-path segmentation approach. However, the original model called DPRNN used a fixed feature dimension and unchanged segment size throughout all layers of the network. In this paper, we propose a multi-scale feature fusion transformer network (MSFFT-Net) based on the conventional dual-path structure for single-channel speech separation. Unlike the conventional dual-path structure where only one processing path exists, adopting several iterative blocks with alternative intra-chunk and inter-chunk operations to capture local and global context information, the proposed MSFFT-Net has multiple parallel processing paths where the feature information can be exchanged between multiple parallel processing paths. Experiments show that our proposed networks based on multi-scale feature fusion structure have achieved better results than the original dual-path model on the benchmark dataset-WSJ0-2mix, where the SI-SNRi score of MSFFT-3P is 20.7dB (1.47% improvement), and MSFFT-2P is 21.0dB (3.45% improvement), which achieves SOTA on WSJ0-2mix without any data augmentation method.

cs.SD