arXiv ScienceSearch

arXiv subjects

Xiaobo Li

Publications and source records attributed to Xiaobo Li.

At least 19 recordsLinked to original sources

A computable representation of the physical laboratory enables verifiable workflows

Making science computable requires representations of both scientific knowledge and the physical world in which scientific claims are tested. A computable representation of the physical laboratory is established through typed research objects, capability-bound operations and a compositional workflow algebra. It provides the physical-world counterpart to machine-readable knowledge, expressing workflows as programs over evolving laboratory states with explicit dependencies, decisions, iteration and concurrency. The representation was implemented in a modular agentic robotic laboratory by binding formal operations to executable Function Skills. For diverse scientific intents, capability-relative workflows were generated, while stateful simulation propagated object transformations and verified operation preconditions and laboratory constraints before dispatch. The proposed representation and its engineering framework jointly establish a general computational interface between agent reasoning and capability-bound physical transformations, providing a foundation for end-to-end autonomous scientific discovery.

cs.AI

Preview-Based Relative-Motion Control of an Insertion Tool for Neural-Thread Placement in Pulsating Tissue

Flexible neural electrode threads must be placed at a prescribed depth while the cortical surface moves with cardiac and respiratory pulsation. A controller tracking a fixed point in the laboratory frame cannot distinguish commanded insertion from tissue motion; the error appears as both a depth offset and relative tip--tissue velocity during contact. This paper formulates thread insertion in tissue-relative coordinates: a harmonic observer predicts delayed cortical-surface motion over the control horizon, a constrained MPC regulates the tip relative to that prediction while limiting actuator effort and lateral relative velocity, and an augmented disturbance state removes the steady offset from persistent contact force and model mismatch. In a 1-DOF MuJoCo benchmark, the controller reaches RMS relative-placement errors of 12.0\um\ free-space and 1.9\um\ in contact, versus 18.3/176.8\um\ for delayed-feedback impedance and 286.1/275.5\um\ for laboratory-frame PD -- the lower contact offset costs more peak contact force (3.43 vs.\ 2.00~mN), since it drives to commanded depth rather than yielding to tissue. A 3-DOF extension reduces lateral shear velocity from 1.34 to 0.50~mm/s at 2.1\um\ lateral placement error, and a feasibility-restoring soft-slack formulation keeps the shear constraint solvable under degraded sensing where a matched hard-constraint controller fails. A two-vertex Lyapunov certificate for the finite-horizon gain holds over $-40\%/{+}50\%$ reflected-mass mismatch, and the 1-DOF QP solves in under 0.4~ms at the 95th percentile. These results are a simulation-based control benchmark, not a clinical safety claim: the modeled tip is a rigid contact point, and flexible-thread mechanics, a validated force constraint, biological damage thresholds, and hardware-realistic sensing and timing remain necessary before deployment.

eess.SY

Stress-testing large language model agents in a robotic chemistry laboratory

AI is evaluated through knowledge, reasoning and plan generation, yet scientific agency requires reliable physical action and adaptation to evidence. Here, we use a robotic chemistry laboratory as a physical-world testbed to make scientific agency measurable. Its 45 modular workstations exposed as machine-readable skills enabled 4,608 trials. Only 3.3% of trials produced expert-assessed executable workflows under laboratory constraints; even the best system achieved 28.1%. Long-horizon planning remained a challenge: only three executable workflows exceeded 30 operations, although the longest contained 44. Across five rounds, experimental feedback prompted local adjustments but no workflow-level replanning or analytical-method redesign. By making physical executability and evidence-driven replanning measurable, our study provides an evidence-based assessment of deployment readiness and a diagnostic framework to guide closed-loop improvements towards physically grounded autonomous research.

cs.AI

OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations

Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-Positives in marketing slots, where position advantage masks mediocre content quality, leading to biased recommendation ecosystems. We propose a framework called Orthogonal Disentanglement of Access habits (OrDA) to purify interest signals. OrDA utilizes a dual-tower structure with a gated allocation layer to adaptively route features and minimize interference. To ensure rigorous separation, we employ orthogonal regularization to constrain the latent interest and habit manifolds to be geometrically perpendicular. OrDA performs causal intervention (do-calculus) during inference to rank items solely by purified interest scores. Empirical online evaluations on large-scale datasets demonstrate that OrDA effectively eliminates access-habit bias, outperforming state-of-the-art methods in predictive accuracy. Online AB test 5.64% shows user click-through rates (UCTR) improvement on the Zhima homepage marketing block, Zhima rent-floor recommendation.

cs.LG

Labimus: A Simulation and Benchmark for Humanoid Dexterous Manipulation in Chemical Laboratory

Laboratory automation has made remarkable progress through robotic platforms and AI-driven scientific reasoning. However, many laboratory operations (e.g., solid--solid transfer) remain inherently dynamic and require real-time adaptation to different materials and experimental conditions. Such precision-critical manipulations are difficult to standardize, motivating the use of humanoid robots with dexterous hands. Despite this opportunity, no existing benchmark evaluates humanoid manipulation in precision-critical laboratory environments. We present Labimus, to our knowledge, the first benchmark for humanoid dexterous manipulation in organic chemistry laboratories. Labimus reconstructs over 30 functionally faithful assets from real organic chemistry workstations through real-to-sim modeling, collectively covering the core operations of routine organic chemistry experiments. The benchmark integrates articulated laboratory instruments, particle-based powder physics, and closed-loop instrument readouts, enabling a complete manipulation-to-measurement pipeline. It further defines six atomic operations and a seven-step solid-weighing workflow derived from real laboratory standard operating procedures. We introduce a precision-aware evaluation protocol designed to jointly measure task completion, experimental precision, and long-horizon execution. We benchmark three representative policies under procedural layouts and environmental perturbations. Results reveal a precision gap: policies that successfully complete laboratory tasks can still fail to satisfy the quantitative tolerances required by experimental protocols. Our benchmark exposes a fundamental disconnect between task completion and experimental validity, providing a new testbed for developing reliable humanoid robots for scientific laboratories.

cs.RO

Observational Technological Innovations and Future Development of the Lijiang Coronagraph

As a core ground-based coronal observation facility in China's low-latitude high-altitude regions, the Lijiang Coronagraph leverages the natural advantages of Lijiang Astronomical Observation Station, including its 3200 m altitude and low atmospheric turbulence. It has undergone a full development process, from introduction via Chinese-Japanese cooperation to independent innovation and iteration. This paper systematically summarizes its core technological innovations: upgrade of the automatic operating system, integration of the dual-band observation system, stray light suppression based on image differencing before and after cleaning, and high-precision image calibration and registration. These advances have significantly improved observation efficiency and data quality, laying a solid foundation for high-quality observations. Scientifically, the data reveal that 1.1 solar radii is a highly correlated region between coronal green line brightness and magnetic field intensity. The study also confirms a strong correlation between the coronal green line and the SDO/AIA 21.1 nm extreme ultraviolet band (correlation coefficient: 0.89-0.99), supporting early warning research on Coronal Mass Ejections (CMEs). These results provide key data for verifying coronal heating mechanisms and exploring the origin of the slow solar wind. The experience from the Lijiang Coronagraph not only lays a foundation for China's next-generation large-aperture coronagraphs, but also accelerates progress in low coronal observation capabilities, enabling the country to build internationally competitive capabilities in this field. The system is also an important part of the global coronal observation network.

astro-ph.SR

Comparative analysis of missing data imputation methods for CSST survey: Impact on photometric redshift estimation performance

Improving the accuracy of photometric redshifts (photo-$z$) is essential for reliable statistical studies of cosmology and galaxy evolution. However, missing photometric bands are a common observational challenge that can significantly degrade photo-$z$ estimation accuracy. In this work, we present a systematic evaluation of data imputation methods aimed at improving photo-$z$ performance. We benchmark a range of representative machine learning (ML) and deep learning (DL) architectures, identifying k-nearest neighbors (KNN) and the attention-based SAITS model as the leading performers. These models are then applied to China Space Station Survey Telescope (CSST) mock data to assess their performance under realistic observational conditions. Our results show that KNN yields the highest accuracy under idealized missing completely at random (MCAR) conditions with complete training sets, whereas robustness tests reveal that SAITS significantly outperforms KNN when training data is incomplete or when applied to realistic mixed-mechanism scenarios. We find that domain consistency between training and testing missingness patterns is a prerequisite for optimal performance, highlighting the risks of domain shift in supervised regression tasks. Furthermore, our analysis demonstrates that while general imputation models are highly effective for MCAR and missing at random (MAR) data, they are detrimental when applied to missing not at random (MNAR) data arising from flux limits, as statistical models fail to capture the physical information inherent in these non-detections. Consequently, we advocate for more sophisticated architectures capable of disentangling stochastic missingness from physical non-detections to address these distinct mechanisms individually.

astro-ph.GA

LoHGNet: Infrared Small Target Detection through Lorentz Geometric Encoding with High-Order Relation Learning

Infrared small target detection (IRSTD) remains challenging due to the scarcity of useful target cues and the presence of severe background clutter. Most current methods rely on conventional feature learning and local interaction modeling, where features are represented in Euclidean space. However, such designs may still be limited in describing the subtle differences of weak targets and the contextual relations between targets and backgrounds. To address these limitations, we propose LoHGNet, an IRSTD network that integrates Lorentz geometric encoding with high-order relation learning. By introducing Lorentz manifold based feature learning, LoHGNet offers a different feature representation from conventional IRSTD methods and provides new discriminative cues for IRSTD. Specifically, a Lorentz encoding branch is constructed with the Geometric Attention Guided Lorentz Residual Convolution Module (GA-LRCM) to perform feature modeling under hyperbolic geometric constraints and enhance the hierarchical geometric representation capability of weak targets. Subsequently, the hyperbolic features are mapped into the Euclidean tangent space through logarithmic mapping, and a High-Order Relation Learning Module (HORL) is designed to model the high-order contextual dependencies between targets and backgrounds via hypergraph construction, thereby improving target discrimination in complex backgrounds. Experimental results on three datasets demonstrate that the proposed LoHGNet achieves competitive performance in both detection accuracy and adaptability to complex scenes. The code will be available at https://github.com/Kingwin97.

cs.CV

IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based criteria rather than underlying biological mechanisms, thereby limiting early detection and the development of biologically-informed interventions. Emerging studies have begun to investigate the neurobiological signatures of BED, yet their findings are often difficult to generalize due to the reliance on hypothesis-driven parametric models, single-modality analyses, and limited data diversity. Therefore, there is a critical need for advanced data-driven frameworks capable of modeling multimodal data to uncover generalizable and biologically meaningful signatures of BED. In this study, we propose the Interpretable Modality-Aware Mixture-of-Experts (IMA-MoE), a novel architecture designed to integrate heterogeneous neuroimaging, behavioral, hormonal, and demographic measures within a unified predictive framework. By encoding each measure as a distinct token, IMA-MoE enables flexible modeling of cross-modal dependencies while preserving modality-specific characteristics. We further introduce a token-importance mechanism to enhance interpretability by quantifying the contribution of each measure to model predictions. Evaluated on the large-scale Adolescent Brain Cognitive Development (ABCD) dataset, IMA-MoE demonstrates superior performance in differentiating BED from healthy controls compared with baseline methods, while revealing sex-specific predictive patterns, with hormonal measures contributing more prominently to prediction in females. Collectively, these findings highlight the promise of interpretable, data-driven multimodal modeling in advancing biologically-informed characterization of BED and facilitating more precise and personalized interventions in neuropsychiatric disorders.

cs.CV

Conditional Evidence Reconstruction and Decomposition for Interpretable Multimodal Diagnosis

Neurobiological and neurodegenerative diseases are inherently multifactorial, arising from coupled influences spanning genetic susceptibility, brain alterations, and environmental and behavioral factors. Multimodal modeling has therefore been increasingly adopted for disease diagnosis by integrating complementary evidence across data sources. However, in both large-scale cohorts and real-world clinical workflows, modality coverage is often incomplete, making many multimodal models brittle when one or more modalities are unavailable. Existing approaches to incomplete multimodal diagnosis typically rely on group-wise or static priors, which may fail to capture subject-specific cross-modal dependencies; moreover, many models provide limited interpretability into which evidence sources drive the final decision. To address these limitations, we propose Conditional Evidence Reconstruction and Decomposition (CERD), a framework for interpretable multimodal diagnosis with incomplete modalities. CERD first reconstructs missing modality representations conditioned on each subject's observed inputs, then decomposes diagnostic evidence into shared cross-modal corroboration and modality-specific cues via logit-level attribution. Experiments on the Alzheimer's Disease Neuroimaging Initiative (ADNI) demonstrate that CERD outperforms competitive baselines under incomplete-modality settings while producing structured and clinically aligned evidence attributions for trustworthy decision support.

cs.CV

A Robust Geometric Distortion Solution for Main Survey Camera of CSST

The advancement in sensitivity and field of view of next-generation wide-field survey telescopes requires astrometric measurements with high precision, even in the presence of significant geometric distortions. To address this challenge, we develop a Weighted Polynomial Distortion Correction in 2-Phase (WPDC-2P) method. This approach enhances stellar cross-matching, incorporates distance-based weighting into the traditional polynomial fitting, and employs a look-up table to absorb the remaining distortion residuals. Validated on simulated data from the Main Survey Camera of the \emph{Chinese Space Station Survey Telescope} (CSST), incorporating geometric distortions up to approximately $200$ pixels, the method achieves astrometric standard deviation ranging from 0.013 to 0.107 pixels (0.03 pixels for the $g$-1 detector) across all 18 detectors. Under extreme crowding conditions (e.g., globular cluster NGC 2298), the astrometric precision for the $g$-1 detector reaches 0.05-pixel level within the central region ($r_d < 4000$), despite a centroiding precision of $\sim$0.04 pixels. When applied to the Beijing-Arizona Sky Survey data, for which the standard pipeline delivers an astrometric uncertainty of $\sim$20 mas, our method reduces the positional scatter to $ σ_{Δα}=5.494$ mas (0.01 pixels) and $ σ_{Δδ}=9.981$ mas (0.02 pixels) using only a weighted 3rd-order polynomial correction. The method has been integrated into the CSST data processing pipeline and is prepared for further refinement using on-orbit calibration data.

astro-ph.IM

Neural Functional Alignment Space: Brain-Referenced Representation of Artificial Neural Networks

We propose the Neural Functional Alignment Space (NFAS), a brain-referenced representational framework for characterizing artificial neural networks on equal functional grounds. NFAS departs from conventional alignment approaches that rely on layer-wise features or task-specific activations by modeling the intrinsic dynamical evolution of stimulus representations across network depth. Specifically, we model layer-wise embeddings as a depth-wise dynamical trajectory and apply Dynamic Mode Decomposition (DMD) to extract the stable mode. This representation is then projected into a biologically anchored coordinate system defined by distributed neural responses. We also introduce the Signal-to-Noise Consistency Index (SNCI) to quantify cross-model consistency at the modality level. Across 45 pretrained models spanning vision, audio, and language, NFAS reveals structured organization within this brain-referenced space, including modality-specific clustering and cross-modal convergence in integrative cortical systems. Our findings suggest that representation dynamics provide a principled basis for

cs.CV

A simple, flexible method for timing cross-calibration of space missions

The timing (cross-)calibration of astronomical instruments is often done by comparing pulsar times-of-arrival (TOAs) to a reference timing model. In high-energy astronomy, the choice of solar system ephemerides and source positions used to barycenter the photon arrival times has a significant impact on the procedure, requiring a full reprocessing of the data each time a new convention is used. Our method, developed as part of the activities of the International Astronomical Consortium for High Energy Calibration (IACHEC), adapts an existing pulsar solution to arbitrary JPL ephemerides and source positions by simulating geocentric TOAs and refitting timing models (implemented with PINT). We validate the procedure and apply it to thousands of observations of the Crab pulsar from 15 missions spanning 1996--2025, demonstrating inter-ephemeris TOA consistency at the $\lesssim5 μ$s level, using the DE200/FK5-based Jodrell Bank Monthly Ephemeris as a common reference. We release the TOAExtractor open-source tool and a TOA database to support future calibration and scientific studies. Instrument timing performance is broadly consistent with mission specifications; the X-ray-to-radio phase offset varies with energy and time at a level that is marginally consistent with the uncertainties of the radio ephemeris, motivating coordinated multiwavelength follow-up.

astro-ph.IM

Modeling Atmospheric Ion Escape from Kepler-1649 b and c over Time

Rocky planets orbiting M-dwarf stars are prime targets for atmospheric characterization, yet their long-term evolution under intense stellar winds and high-energy radiation remains poorly constrained. The Kepler-1649 system, hosting two terrestrial exoplanets orbiting an M5V star, provides a valuable laboratory for studying atmospheric evolution in the extreme environments typical of M-dwarf systems. In this Letter we show that both planets could have retained atmospheres over gigayear timescales. Using a multi-species magnetohydrodynamic model, we simulate atmospheric ion escape driven by stellar winds and extreme ultraviolet radiation from 0.8 to 4.0 Gyr. The results reveal a clear decline in total ion escape rates with stellar age, as captured by a nonparametric LOWESS regression, with O$^{+}$ comprising 98.3%-99.9% of the total loss. Escape rates at 4.0 Gyr are two to three orders of magnitude lower than during early epochs. At 0.8 Gyr, planet b exhibits 3.79$\times$ higher O$^{+}$ escape rates than planet c, whereas by 4.0 Gyr its O$^{+}$ escape rate becomes 39.5$\times$ lower. This reversal arises from a transition to sub-magnetosonic star-planet interactions, where the fast magnetosonic Mach number, $M_f$, falls below unity. Despite substantial early atmospheric erosion, both planets may have retained significant atmospheres, suggesting potential long-term habitability. These findings offer predictive insight into atmospheric retention in the Kepler-1649 system and inform future JWST observations of similar M-dwarf terrestrial exoplanets aimed at refining habitability assessments.

astro-ph.EP

Unveil A Peculiar Light Curve Pattern of Magnetar Burst with GECAM observations of SGR J1935+2154

Magnetar X-ray Burst (MXB) is usually composed of a single pulse or multiple pulses with rapid rise and brief duration mostly observed in hard X-ray (soft gamma-ray) band. Previous work studied the temporal behavior of some magnetar bursts and employed the Fast Rise Exponential Decay (FRED) model to fit pulses of MXB. However, whether there is other kind of pulse shape has not been explored. In this study, we systematically examined light curve of MXBs from SGR J1935+2154 detected by GECAM between 2021 and 2022. We find that there are different light curve morphologies. Especially, we discover a peculiar and new pattern, Exponential Rise and Cut-Off Decay (ERCOD), which is significantly different from FRED and could be well described by a mathematical function we proposed. We find that MXBs with ERCOD shape are generally longer in duration, brighter in the peak flux, and harder in spectrum. We note that the ERCOD shape is not unique to SGR J1935+2154 but also present in other magnetars. This new light curve pattern may imply a special burst and radiation mechanism of magnetar.

astro-ph.HE

Development and validation of an AI foundation model for endoscopic diagnosis of esophagogastric junction adenocarcinoma: a cohort and deep learning study

The early detection of esophagogastric junction adenocarcinoma (EGJA) is crucial for improving patient prognosis, yet its current diagnosis is highly operator-dependent. This paper aims to make the first attempt to develop an artificial intelligence (AI) foundation model-based method for both screening and staging diagnosis of EGJA using endoscopic images. In this cohort and learning study, we conducted a multicentre study across seven Chinese hospitals between December 28, 2016 and December 30, 2024. It comprises 12,302 images from 1,546 patients; 8,249 of them were employed for model training, while the remaining were divided into the held-out (112 patients, 914 images), external (230 patients, 1,539 images), and prospective (198 patients, 1,600 images) test sets for evaluation. The proposed model employs DINOv2 (a vision foundation model) and ResNet50 (a convolutional neural network) to extract features of global appearance and local details of endoscopic images for EGJA staging diagnosis. Our model demonstrates satisfactory performance for EGJA staging diagnosis across three test sets, achieving an accuracy of 0.9256, 0.8895, and 0.8956, respectively. In contrast, among representative AI models, the best one (ResNet50) achieves an accuracy of 0.9125, 0.8382, and 0.8519 on the three test sets, respectively; the expert endoscopists achieve an accuracy of 0.8147 on the held-out test set. Moreover, with the assistance of our model, the overall accuracy for the trainee, competent, and expert endoscopists improves from 0.7035, 0.7350, and 0.8147 to 0.8497, 0.8521, and 0.8696, respectively. To our knowledge, our model is the first application of foundation models for EGJA staging diagnosis and demonstrates great potential in both diagnostic accuracy and efficiency.

cs.CV

A molecular rotor driven by an electric field on graphene

We propose a scheme for driving a dipolar molecular rotor to rotate continuously by applying an external electric field: the dipolar rotor is fixed on a graphene sheet via a metal atom to facilitate the free rotation; it is in the meantime subjected to an electric field oriented parallel to the graphene sheet. We use computational modeling with density functional theory and Newtonian mechanics, similar to molecular dynamics simulations, to obtain the torque, angular velocity, and rotation period of the rotor. Our results show that the dipolar rotor designed here can rotate with a period of 2.96 ps by an alternating rectangular electric field with a strength of 0.5 V/Å. However, a cosine wave alternating electric field depending on time cannot drive the dipolar rotor to rotate regularly. Therefore, a cosine wave electric field depending on the rotation angle is suggested, as it can not only drive the rotor but also produce additional power. Machine learning molecular dynamics (MLMD) simulations further confirm that the rotor remains thermodynamically stable under an electric field. This work reveals the rotation mechanism of a dipolar molecular rotor in a transverse electric field, and we hope this work can open a new path for designing more diverse molecular machines in experiments.

physics.comp-ph

Timing and spectral studies of SRGA J144459.2$-$604207 with NICER, Einstein Probe, IXPE, NuSTAR, Insight-HXMT and INTEGRAL during its 2024 outburst

SRGA J144459.2$-$604207 is a newly confirmed accreting millisecond X-ray pulsar and type I X-ray burster. We present the broadband X-ray timing and spectral behaviors of SRGA J144459.2$-$604207 during its 2024 outburst. The data were collected from NICER, Einstein Probe, IXPE, Insight-HXMT, NuSTAR and INTEGRAL observations. X-ray pulsations have been detected for the 1.5--90 keV energy range throughout the `ON' phase of the outburst from MJD $\sim 60355-60385$. We refined the orbital and spin ephemerides assuming a circular orbit, and found that the pulsar was in a spin-up state during MJD $\sim$ 60361--60377 showing a significant spin-up rate $\dotν$ of $(3.15\pm 0.36)\times10^{-13}~{\rm Hz~s^{-1}}$. Around MJD $\sim 60377$ a swing was detected in the spin evolution accompanied by significantly enhanced pulsed emission. We studied the pulse profile morphology during the X-ray bursts as observed by Insight-HXMT, IXPE and NuSTAR. During the bursts, pulsations were detected across the 2--60 keV with shapes broadly consistent with those observed for the persistent emission. We found, however, that the `burst' pulse profiles exhibit significant phase offsets relative to the pre- and post-burst profiles. These offsets systematically decrease with increasing energy, $Δϕ\approx0.15$, 0.11 and 0.02 for IXPE, Insight-HXMT ME and HE in 2--8, 5--30 and 20--60 keV, respectively, and $Δϕ\approx 0.21$, 0.10 and 0.07 for NuSTAR in 3--10, 20--35 and 35--60 keV, respectively, compared to the pre- and post-burst profiles. We performed a joint spectral analysis of quasi-simultaneous NICER, NuSTAR, and Insight-HXMT data for two epochs. The resulting spectra from both observations were consistent and well-described by an absorbed thermal Comptonization model, nthcomp, plus relativistic reflection, relxillCp.

astro-ph.HE