A general representation form of system data with fundamental lemma as a special case
This note introduces a general data representation of dynamic systems.
arXiv subjects
Publications and source records attributed to Linlin Li.
This note introduces a general data representation of dynamic systems.
Galaxy interactions can enhance star formation, but how star formation evolves during galaxy interactions remains poorly constrained by observations. We combine pair separation, morphological disturbance, and recent star formation history to study this evolution. We measure morphological disturbance with the shape asymmetry parameter ($A_{\rm shape}$) using deep images from the DESI Legacy Surveys, and use Galaxy Zoo DESI classifications as an independent test. We infer recent star formation histories by comparing SFR and $\mathrm{EW}(\mathrm{H}\alpha)$, which trace current star formation, with $\mathrm{EW}(\mathrm{H}\delta_A)$ and $D_n4000$, which are sensitive to stellar populations formed over longer timescales. At small projected separations, strongly disturbed galaxies show the strongest enhancements in SFR and $\mathrm{EW}(\mathrm{H}\alpha)$, indicating that close encounters trigger strong star formation. At intermediate separations ($d_{\rm p}\sim100\,\mathrm{kpc}$), their current star formation is only moderately enhanced, but their $\mathrm{EW}(\mathrm{H}\delta_A)$ enhancement is the strongest. This indicates a larger contribution from intermediate-age stars formed during stronger star formation in the past $\sim0.1$--$1\,\mathrm{Gyr}$. The TNG100 analysis shows that this pattern results from rapid changes in SFR around close passage. SFR rises sharply during the encounter and declines as the galaxies move apart, while tidal disturbances remain visible. As short-lived massive stars disappear and intermediate-age A-type stars begin to dominate the spectrum, the $\mathrm{EW}(\mathrm{H}\delta_A)$ enhancement peaks later than the SFR enhancement. Our results provide important observational evidence for the evolution of SFR along the merger sequence.
Accurate 3D neuron segmentation in fluorescence microscopy is critical for neuroscience. However, the sparse and elongated morphology of neurons poses significant challenges to existing segmentation methods. These methods struggle to preserve both local details and global topology, leading to fragmented results. To address this, we propose NeuroRefiner, a multi-agent system that formalizes the human expert workflow involving iterative global observation and local editing. Specifically, NeuroRefiner comprises three collaborative agents dedicated to diagnosing topological errors, generating correction instructions, and validating refinement quality. To facilitate agent instruction-guided segmentation refinement, we propose TopoRefineNet, a dedicated 3D U-Net-based tool that leverages cross-modality feature fusion to generate refined masks. Through multi-round agent reasoning and voxel-level editing, NeuroRefiner produces topologically more accurate segmentations with enhanced interpretability. Experiments on the BigNeuron, CWMBS, and ZBFWB datasets demonstrate that NeuroRefiner outperforms state-of-the-art methods, notably achieving a 3.02% improvement in F1 score on the challenging ZBFWB dataset.
This paper deals with system representations in finite-sample signal subspaces and their application to data-driven fault detection. The first part addresses concepts of finite-sample image and kernel system representations and, associated with them, image and residual subspaces of finite-sample signals. On this basis, the equivalence between the fundamental lemma and finite-sample image subspace is demonstrated. While the image representation models the nominal system dynamics, the residual representation describes uncertainties in the input-output data and is essential for fault detection. This result extends the fundamental lemma and builds the basis for exploring data-driven fault detection. In the second part, a data-driven projection-based fault detection approach is developed. By means of a singular value decomposition, orthogonal projections onto the image and residual subspaces are realized in the context of a low-rank matrix approximation, leading to projection-based residual generation and evaluation. Finally, analysis of detection performance in the framework of matrix perturbation theory and comparison with existing data-driven fault detection methods are explored.
Using integral field spectroscopy from SDSS-IV MaNGA, we investigate the radial distributions of star formation rate (SFR) and gas-phase metallicity in spiral galaxies that reside in spiral-elliptical (S+E) pairs. Spirals in S+E pairs show suppressed central star formation and elevated metallicities, whereas spirals in spiral-spiral pairs exhibit centrally enhanced star formation and reduced metallicities. The degree of SFR suppression and metallicity enhancement in S+E pairs depends on the masses of the pair members. Spirals with more massive elliptical companions experience stronger star-formation suppression and larger increases in metallicity, while lower-mass spirals show more pronounced metallicity enhancement. In addition, within S+E systems, galaxies with asymmetric gas velocity fields display enhanced SFR and higher metallicities, whereas those with symmetric velocity fields exhibit clear central suppression. Based on these results, we infer that in S+E pairs, the spiral galaxy experiences suppressed gas accretion once it enters the hot circumgalactic medium of its early-type companion, which leads to the observed decline in star-formation activity. When a close encounter takes place, tidal perturbations can compress the remaining cold gas and trigger enhanced star formation, producing rapid chemical enrichment and the associated increase in metallicity.
We investigate the bar fraction in galaxy pairs from the SDSS to assess how galaxy interactions affect bar structures. Compared to isolated galaxies, close pairs exhibit a significantly reduced bar fraction at projected separations within 25 kpc. This reduction is driven almost entirely by systems showing clear merger or disturbance signatures, indicating that tidal interactions suppress bars. The decline is dominated by a decrease in weak bars, while the fraction of strong bars remains largely unchanged. Bar suppression is primarily associated with major mergers and is strongest in massive host galaxies. A weaker but statistically significant suppression is detected in minor mergers only for massive galaxies with small bulges. In contrast, no significant dependence of bar suppression on the relative orientation between pair members is found. These findings provide observational evidence that tidal perturbations in major mergers play a key role in regulating bar evolution.
This paper deals with analysis, simultaneous detection of faults and attacks, fault-tolerant control and attack-resilient of cyber-physical control systems. In our recent work, it has been observed that an attack detector driven by an input residual signal is capable of reliably detecting attacks. In particular, observing system dynamics from the perspective of the system input-output signal space reveals that attacks and system uncertainties act on different system subspaces. These results motivate our exploration of secure and safe cyber-physical control systems in the unified framework of control and detection. The unified framework is proposed to handle control and detection issues uniformly and in subspaces of system input-output data. Its mathematical and control-theoretic basis is system coprime factorizations with Bezout identity at its core. We firstly explore those methods and schemes of the unified framework, which serve as the major control-theoretic tool in our work. It is followed by re-visiting and examining established attack detection and resilient control schemes. The major part of our work is the endeavours to develop a control-theoretic paradigm, in which analysis, simultaneous detection of faults and attacks, fault-tolerant and attack-resilient control of cyber-physical control systems are addressed in a unified manner.
Peri-implant inflammation in orthodontic mini-implant may lead to patient discomfort and treatment failure. This study aims to evaluate the effects of diode laser application on the health of mini-implant, preventing peri-implantitis and promoting healing. A randomized controlled trial was conducted involving 30 orthodontic patients (12 males and 18 females, aged 18-32) who had mini-implants implanted on both sides of the maxilla for anterior teeth retraction. One side of each patient was assigned to either an experimental group receiving diode laser irradiation (650 nm, 25 mW) at specific postoperative intervals or a control group receiving simulated radiation. Clinical assessments included plaque index, modified sulcus bleeding index, probing depth, and incidence of peri-implant mucositis and implant mobility, measured at 1, 4, and 12 weeks post-implantation. Additionally, interleukin-1 beta (IL-1\b{eta}) levels in peri-implant fluid were analyzed via enzyme-linked immunosorbent assay (ELISA). Results indicated that the experimental group exhibited significantly lower plaque indices, sulcus bleeding indices, and probing depths (p < 0.05) compared to the control group. Moreover, the experimental group had fewer cases of peri-implant mucositis (p < 0.05), while differences in implant stability were not statistically significant (p > 0.05). IL-1\b{eta} levels were consistently lower in the experimental group throughout the study duration (p < 0.05). In conclusion, adjunctive diode laser therapy appears to enhance peri-implant health and reduce complications associated with orthodontic mini-implants, suggesting a promising direction for improving patient outcomes in orthodontics. Future research should explore long-term effects and the mechanisms underlying these benefits.
Objective: Time-domain diffuse optical imaging (DOI) requires accurate forward models for photon propagation in scattering media. However, existing simulators lack comprehensive experimental validation, especially for non-contact configurations with oblique illumination. This study rigorously evaluates three widely used open-source simulators, including MMC, NIRFASTer, and Toast++, using time-resolved experimental data. Approach: All simulations employed a unified mesh and point-source illumination. Virtual source correction was applied to FEM solvers for oblique incidence. A time-resolved DOI system with a 32 $\times$ 32 single-photon avalanche diode (SPAD) array acquired transmission-mode data from 16 standardized phantoms with varying absorption coefficient $\mu_a$ and reduced scattering coefficient $\mu_s'$. The simulation results were quantified across five metrics: spatial-domain (SD) precision, time-domain (TD) precision, oblique beam accuracy, computational speed, and mesh-density independence. Results: Among three simulators, MMC achieves superior accuracy in SD and TD metrics, and shows robustness across all optical properties. NIRFASTer and Toast++ demonstrate comparable overall performance. In general, MMC is optimal for accuracy-critical TD-DOI applications, while NIRFASTer and Toast++ suit scenarios prioritizing speed with sufficiently large $\mu_s'$. Besides, virtual source correction is essential for non-contact FEM modeling, which reduced average errors by > 34% in large-angle scenarios. Significance: This work provides benchmarked guidelines for simulator selection during the development phase of next-generation TD-DOI systems. Our work represents the first study to systematically validate TD simulators against SPAD array-based data under clinically relevant non-contact conditions, bridging a critical gap in biomedical optical simulation standards.
Fluorescence Molecular Tomography (FMT) is a promising technique for non-invasive 3D visualization of fluorescent probes, but its reconstruction remains challenging due to the inherent ill-posedness and reliance on inaccurate or often-unknown tissue optical properties. While deep learning methods have shown promise, their supervised nature limits generalization beyond training data. To address these problems, we propose $\mu$NeuFMT, a self-supervised FMT reconstruction framework that integrates implicit neural-based scene representation with explicit physical modeling of photon propagation. Its key innovation lies in jointly optimize both the fluorescence distribution and the optical properties ($\mu$) during reconstruction, eliminating the need for precise prior knowledge of tissue optics or pre-conditioned training data. We demonstrate that $\mu$NeuFMT robustly recovers accurate fluorophore distributions and optical coefficients even with severely erroneous initial values (0.5$\times$ to 2$\times$ of ground truth). Extensive numerical, phantom, and in vivo validations show that $\mu$NeuFMT outperforms conventional and supervised deep learning approaches across diverse heterogeneous scenarios. Our work establishes a new paradigm for robust and accurate FMT reconstruction, paving the way for more reliable molecular imaging in complex clinically related scenarios, such as fluorescence guided surgery.
The morphology of ionized gas velocity maps provides a direct probe of the internal gas kinematics of galaxies. Using integral field spectroscopy from SDSS-IV MaNGA, we analyze a sample of 528 low-inclination, regular disk galaxies to investigate the correlations between velocity map morphology, star formation rate, and gas-phase metallicity. We quantify velocity map morphology using harmonic expansion and adopt two complementary diagnostics: the global kinematic asymmetry, which traces non-axisymmetric perturbations, and the first-order term ratio, which captures axisymmetric radial motions. We find that galaxies with higher kinematic asymmetry are more likely to deviate from the scaling relations, typically lying either above or below the star formation main sequence and systematically below the mass-metallicity relation. In contrast, the first-order term ratio shows only a correlation with gas-phase metallicity in the low-mass range and no significant dependence on star formation rate. Moreover, galaxies below the mass-metallicity relation generally exhibit higher HI gas fractions. These results suggest that external gas accretion is the primary driver of the observed phenomena: inflowing metal-poor gas increases velocity map asymmetry in disk galaxies, dilutes the metallicity, and triggers enhanced star formation. Feedback-driven outflows, bar- and spiral-driven inflows, and galaxy mergers may also contribute, but likely play a secondary role.
Large Language Models struggle with memory demands from the growing Key-Value (KV) cache as context lengths increase. Existing compression methods homogenize head dimensions or rely on attention-guided token pruning, often sacrificing accuracy or introducing computational overhead. We propose FourierAttention, a training-free framework that exploits the heterogeneous roles of transformer head dimensions: lower dimensions prioritize local context, while upper ones capture long-range dependencies. By projecting the long-context-insensitive dimensions onto orthogonal Fourier bases, FourierAttention approximates their temporal evolution with fixed-length spectral coefficients. Evaluations on LLaMA models show that FourierAttention achieves the best long-context accuracy on LongBench and Needle-In-A-Haystack (NIAH). Besides, a custom Triton kernel, FlashFourierAttention, is designed to optimize memory via streamlined read-write operations, enabling efficient deployment without performance compromise.
Long context is an important topic in Natural Language Processing (NLP), running through the development of NLP architectures, and offers immense opportunities for Large Language Models (LLMs), giving LLMs the lifelong learning potential akin to humans. Unfortunately, the pursuit of a long context is accompanied by numerous obstacles. Nevertheless, long context remains a core competitive advantage for LLMs. In the past two years, the context length of LLMs has achieved a breakthrough extension to millions of tokens. Moreover, research on long-context LLMs has expanded beyond length extrapolation to a comprehensive focus on architecture, infrastructure, training, and evaluation technologies. Inspired by the symphonic poem, Thus Spake Zarathustra, we draw an analogy between the journey of extending the context of LLM and the attempts of humans to transcend their mortality. In this survey, we will illustrate how LLM struggles between the tremendous need for a longer context and its equal need to accept the fact that it is ultimately finite. To achieve this, we give a global picture of the lifecycle of long-context LLMs from four perspectives: architecture, infrastructure, training, and evaluation, showcasing the full spectrum of long-context technologies. At the end of this survey, we will present 10 unanswered questions currently faced by long-context LLMs. We hope this survey can serve as a systematic introduction to research on long-context LLMs. Video: https://www.bilibili.com/video/BV11h9AYoEYj. Github: https://github.com/OpenMOSS/Thus-Spake-Long-Context-LLM.
Recent advancements in model architectures and length extrapolation techniques have significantly extended the context length of large language models (LLMs), paving the way for their application in increasingly complex tasks. However, despite the growing capabilities of long-context LLMs, the safety issues in long-context scenarios remain underexplored. While safety alignment in short context has been widely studied, the safety concerns of long-context LLMs have not been adequately addressed. In this work, we introduce \textbf{LongSafety}, a comprehensive safety alignment dataset for long-context LLMs, containing 10 tasks and 17k samples, with an average length of 40.9k tokens. Our experiments demonstrate that training with LongSafety can enhance long-context safety performance while enhancing short-context safety and preserving general capabilities. Furthermore, we demonstrate that long-context safety does not equal long-context alignment with short-context safety data and LongSafety has generalizing capabilities in context length and long-context safety scenarios.
This paper is concerned with the detection, resilient and fault-tolerant control issues for cyber-physical systems. To this end, the impairment of system dynamics caused by the defined types of cyber-attacks and process faults is analyzed. Then, the relation of the system input and output signals with the residual subspaces spanned by both the process and the controller is studied. Considering the limit capacity of standard observer-based detection and feedback control schemes in detecting and handling the cyber-attacks, a modified configuration for cyber-physical systems is developed by transmitting the combinations of the input and output residuals instead of the input and output signals, which is facile for dealing with both the process faults and cyber-attacks. It is followed by the integrated design of fault and attack detection, resilient and fault-tolerant control schemes. To enhance the detectability of cyber-attacks, the potential stealthy attack mechanisms on deteriorating the tracking behavior and feedback control performance are developed from the attackers' point of view, and the associated detection schemes for such stealthy attacks are proposed from the defenders' point of view. A case study on the robotino system is utilized to demonstrate the proposed resilient cyber-physical configuration.
Recently, significant efforts have been devoted to enhancing the long-context capabilities of Large Language Models (LLMs), particularly in long-context reasoning. To facilitate this research, we propose \textbf{DetectiveQA}, a dataset specifically designed for narrative reasoning within long contexts. We leverage detective novels, averaging over 100k tokens, to create a dataset containing 1200 human-annotated questions in both Chinese and English, each paired with corresponding reference reasoning steps. Furthermore, we introduce a step-wise reasoning metric, which enhances the evaluation of LLMs' reasoning processes. We validate our approach and evaluate the mainstream LLMs, including GPT-4, Claude, and LLaMA, revealing persistent long-context reasoning challenges and demonstrating their evidence-retrieval challenges. Our findings offer valuable insights into the study of long-context reasoning and lay the base for more rigorous evaluations.
The long-context capability of the Large Language Models (LLM) has made significant breakthroughs, but the maximum supported context length in length extrapolation remains a critical bottleneck limiting their practical applications. The constraint of context length in LLMs arises from the self-attention mechanism, which cannot effectively and efficiently capture the semantic relationships within infinitely long contexts via the limited pre-trained positional information and attention scope. In this work, we propose ReAttention, a training-free approach enabling LLM based on the self-attention mechanism to support an infinite context with a finite attention scope under sufficient memory resources. ReAttention performs the position-agnostic top-$k$ attention before the ordinary position-aware self-attention, freeing LLMs from the length extrapolation issue. We validate the performance of ReAttention on the LongBench, L-Eval, and InfiniteBench and demonstrate that it is on par with traditional methods. Furthermore, we also apply ReAttention on mainstream LLMs, including LLaMA3.1-8B and Mistral-v0.3-7B, enabling them to support context lengths of at least 1M and even expanding the context length of LLaMA3.2-3B-chat by 128$\times$ to 4M without any further training in Needle-In-A-Haystack tests. We also improve the efficiency of ReAttention with Triton and achieve an efficient extrapolation without additional overhead. The code is available at https://github.com/OpenMOSS/ReAttention.
IW And-type dwarf novae are anomalous Z Cam stars featured with outbursts happening during standstill states, which are not expected in the standard disk instability model. The physical mechanisms for these variations remain unclear. In this study, we report the discovery of a new candidate IW And-type dwarf nova J0652+2436, identified with its frequent outbursts from the slowly rising standstill states. Luckily, the TESS observations during a long standstill state and the earlier K2 observations give a chance to find the orbital and negative superhump period in the light curve of J0652+2436, allowing the measurement of its mass ratio of 0.366. This mass ratio is marginally possible for the tidal instability to set in according to previous SPH simulations. Thus, we propose that the outbursts in J0652+2436 are likely to be caused by the growing accretion disk during standstills, in favor of the previous hypothesis of the mechanisms lying in all IW And stars. We conclude that J0652+2436 might be the first IW And star with both a precessing tilted disk and tidal instability, which will be an important laboratory for studying the accretion disk dynamics and help understand IW And phenomenon.