arXiv ScienceSearch

arXiv subjects

Jianing Li

Publications and source records attributed to Jianing Li.

At least 19 recordsLinked to original sources

Sparse Identification for Automatic Large-Scale Screening: A Constraint-Aware Framework with Ultra Fast Decoding Algorithm

In the early stages of a pandemic, identification of a small number of infected individuals through large-scale screening is critical for pandemic control, yet remains challenging under limited reagents and testing capacity. Existing group testing methods suffer from either high computational complexity or low identification accuracy. Even worse, no available methods provide theoretically rigorous analysis for sparse identification with hard constraints caused by the sample usage constraint and the dilution effect existing ubiquitously in practical applications. In this article, we propose the Logic Screening method (LoSc), an ultra fast, accurate, and theoretically grounded framework for large-scale screening. LoSc introduces a novel decoding algorithm with a very simple selection strategy, achieving identification of all positives with only O(klogn) pooled tests. The decoding relies only on logical operations, enabling direct hardware implementation and yielding ultra fast computational implementation. Moreover, LoSc explicitly incorporates dilution and sample usage constraints into pooling designs, and establishes theoretical guarantees to guide optimal pooling configurations. Extensive simulations confirm the superior effectiveness, efficiency, and scalability. We believe LoSc offers a fast and reliable solution for automatic large-scale screening.

stat.ML

Graph-Aware Group Testing with Locally Clustered Infections

Group testing has been widely used to identify infected individuals with a limited number of tests, typically under the assumption of independent infections. Recent studies have exploited correlations among individuals, but often require additional information beyond the contact graph, such as community structures, interaction strengths, or detailed infection dynamics. Such information may be unavailable, incomplete or unreliable in practice. In this work, we assume that only the contact graph is known and develop a graph-aware group testing framework that exploits localized infection clustering in pooling design, fundamental limits of recovery, and decoding. Specifically, we propose an optimal transport-based pooling design that incorporates graph proximity and pooling constraints into a unified optimization framework. We prove that, under mild conditions, the proposed design eliminates uninfected individuals with higher probability than the Bernoulli pooling design, reducing the feasible search space for decoding. Then, we characterize the family of possible infected sets induced by localized infection clustering and derive necessary conditions on the number of tests required for exact recovery, revealing a lower testing requirement than that under the combinatorial prior. For decoding, we model the infection states of the population as a piecewise-constant graph signal and propose a graph total variation regularized decoder. We establish sufficient conditions for exact recovery under the Bernoulli pooling design in both noiseless and noisy settings, and show that O(Klog(n/K)) tests are sufficient under mild conditions in the noiseless case. Extensive simulations on synthetic and real-world networks demonstrate the effectiveness of the proposed framework and the benefit of exploiting graph-induced correlations in group testing.

eess.SP

Robust Recovery of Sparse Support in Constrained Group Testing

In the early stage of a pandemic, rapidly identifying a small number of infected individuals through large-scale screening is critical for pandemic control. Group testing has been widely used to improve testing efficiency and numerous studies have investigated the problem under noisy measurements, typically modeled as bit-flipping of test outcomes. However, these methods do not consider the constraints imposed by dilution, pool size, and the limit of detection (LOD), which can lead to false negatives when the viral load in a pool falls below the LOD. In addition, liquid dispensing errors, common in laboratory settings, affects diluted viral loads in a nonlinear manner. In this work, we introduce a novel measurement model that characterizes the process of sample pooling and dilution, incorporating LOD-induced binary quantization as well as liquid dispensing errors. For the case with known sparsity level, we propose a low-complexity decoding algorithm and provide theoretical guarantees for exact support recovery under both noiseless and noisy settings. For the case with unknown sparsity level, we develop a blind support recovery algorithm, along with a heuristic variant to enhance robustness, which can achieve exact support recovery with only O(klogn) measurements. Extensive simulations show that the proposed algorithms outperform existing combinatorial group testing algorithms, validating the effectiveness, efficiency and robustness in large-scale screening.

eess.SP

Unveiling QCD Criticality with Cross-Rapidity Net-Baryon Cumulants

Fluctuations of conserved charges are a primary tool in the search for the QCD critical endpoint, but their beam-energy dependence is complicated by global baryon-number conservation, whose influence changes as the experimental acceptance covers different fractions of the collision system. We propose using correlations of net-baryon fluctuations between two separated rapidity windows to exploit the distinct signatures of conservation and critical dynamics. Global conservation produces a negative cross-window correlation, whereas a common long-wavelength critical fluctuation produces a positive one. We show that, within a canonical independent-source framework, the conservation-induced background and the critical signal enter additively at leading order, allowing the leading conservation term to be estimated and subtracted. The resulting yield-scaled correlator follows the nonmonotonic enhancement of an Ising-mapped equilibrium correlation length along a freeze-out trajectory passing near a hypothetical critical endpoint, while wider rapidity windows reduce the response through thermal smearing. Cross-rapidity cumulants therefore provide a rapidity-differential strategy for reducing the leading conservation background and sharpening fluctuation-based searches for QCD criticality in beam-energy-scan experiments.

nucl-th

Hadronic rescattering effects on net-proton cumulants from functional renormalization group calculations

Net-proton cumulants in the Beam Energy Scan region of heavy-ion collisions are widely used to probe critical fluctuations associated with the conjectured critical endpoint of Quantum Chromodynamics (QCD). Most existing studies, however, concentrate on the initial-state or phase-transition contributions, while the impact of hadronic rescattering on these observables has not been fully quantified. To address this gap, we construct event-by-event proton and antiproton distributions from functional renormalization group (fRG) cumulants using the maximum entropy principle, and propagate the resulting particles through the hadronic transport model SMASH in a simplified spherical evolution setup. We systematically investigate how the hadronic cascade modifies net-proton cumulants at collision energies $\sqrt{s_{NN}}=3.0$, 3.9, 4.9, 7.2, and 7.7~GeV. In the canonical-ensemble framework, which enforces exact net-baryon number conservation, the higher-order cumulant signal---in particular the ratio $C_4/C_2$ at $\sqrt{s_{NN}}=4.9$~GeV---is strongly reduced during the early stage of the cascade; the suppression of $C_4/C_2$ reaches approximately $20\%$. The non-monotonic energy dependence inherited from the fRG input survives the hadronic evolution, but its magnitude is substantially modified. These results demonstrate that hadronic rescattering provides a non-negligible background effect that must be accounted for when extracting QCD critical-point signals from experimental data.

nucl-th

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either focus solely on perception or rely on specific domains such as maths or coding. Evaluation for reasoning capabilities that align with an open-world environment is still required, especially one that considers perception and reasoning jointly. To bridge this gap, we propose to evaluate LVLMs by exploiting visual illusions as a diagnostic tool. Visual illusions are phenomena in which the human visual system misinterprets objective signals, resulting in an understanding that deviates from reality. We constructed Illusion-Reasoning, a benchmark of illusion images collected from the real world, incorporating diverse annotated question-answer pairs. Based on Illusion-Reasoning, we show that the reasoning capabilities of a wide range of LVLMs are not as advanced as claimed. Our work provides new insights into LVLMs and offers future directions for optimisation. Our project is publicly available at https://github.com/zhaoliangjie55/EMNLP2026_Illusion.

cs.CL

Magic-wavelength matter-wave interferometry with optical clock states

Optical clocks and atom interferometers provide complementary ways to measure time, motion and gravity. Combining these capabilities requires matter-wave beam splitters that manipulate different clock states in the same way, so that optical internal energy becomes a controlled degree of freedom rather than a source of systematic phase shifts. Here we realized a dual matter-wave interferometer operating simultaneously on the two states of the $^{88}$Sr optical clock transition, $^1S_0$ and $^3P_0$. The interferometer is driven by Bragg pulses at the 813 nm magic wavelength, for which the two clock states experience the same optical coupling strength. This realizes a common matter-wave beam splitter for atoms whose internal energies differ by an optical excitation. With a sensitivity of 30 mrad, our measurement is consistent with a zero differential phase shift between the two clock-state Mach-Zehnder interferometers, translating to an absence of state-dependent acceleration in free fall at the level of $10^{-5}$. We further used the same interferometer to measure state-dependent optical dipole forces and determine a tune-out wavelength of the metastable $^3P_0$ state to be 478.95(8) nm. These results establish magic-wavelength clock-state interferometry as a platform for differential force sensing, excited-state polarizability metrology and future quantum-clock tests of gravity.

physics.atom-ph

Existence of generalized bent functions in the exceptional $q\equiv2\pmod4$, odd-dimensional case

We resolve an open problem of Kumar, Scholtz, and Welch (1985) by constructing generalized bent functions from $(\mathbb{Z}/q\mathbb{Z})^m$ to $\mathbb{Z}/q\mathbb{Z}$ in the exceptional case $q\equiv2\pmod4$ with $m$ odd, the case their paper left without a construction and which four decades of subsequent work had addressed only through nonexistence results. Concretely, for all integers $m\geq d\geq 2$ with $m=dr+2k$, $r\geq 1$, $k\geq 0$, we construct an explicit generalized bent function from $(\mathbb{Z}/q\mathbb{Z})^m$ to $\mathbb{Z}/q\mathbb{Z}$, where $q=2(2^d-1)$. We further show that, when $d\ge3$, these generalized bent functions have Fourier coefficients that are not roots of unity --- all of them when $r$ is odd --- which gives a negative answer to a recent question of Armario, Egan, Kharaghani, and Ó~Catháin about bent vectors for character tables.

math.CO

GraphQLer: Enhancing GraphQL Security with Context-Aware API Testing

GraphQL APIs power production systems across financial services, e-commerce, and social platforms, yet their most critical access-control vulnerabilities--Insecure Direct Object Reference (IDOR), Use-After-Free (UAF), and state-dependent injection--routinely escape automated security testing. Industry-standard scanners (ZAP) and the leading research fuzzer (EvoMaster) test operations in isolation and cannot compose the multi-step sequences these flaws require. We present GraphQLer, an open-source automated security testing framework built for production GraphQL APIs. GraphQLer constructs a typed dependency graph from live schema introspection and synthesizes vulnerability chains--ordered operation sequences targeting specific flaw classes. Three strategies cover the critical attack surface: topological SCC-traversal for general reachability, cross-user IDOR replay for broken access control, and CREATE -> DELETE -> READ synthesis for UAF. On a production financial API (FinServ), GraphQLer identified eight potential vulnerabilities--including denial-of-service vectors that exposed stack traces and sensitive implementation details--without prior documentation or authentication credentials. On a self-hosted Saleor instance pinned to the CVE-2022-39275 commit, GraphQLer reproduced all four broken-access-control mutations cited in the security advisory. On the 11 public APIs of the coverage set, GraphQLer achieves 85.52% mean PositiveCoverage versus 29.29% (EvoMaster) and 21.80% (ZAP); across the 21 evaluated APIs it detects all 5 confirmed IDOR endpoints, UAF behavior on two controlled schemas, and confirms XSS and SQLi on DVGA (an independent third-party oracle)--while all baselines detect zero chain-based vulnerabilities.

cs.CR

Neuromorphic Object Detection: An In-Depth Study and Future Directions

Conventional frame-based cameras face significant challenges in detecting objects under high-speed motion blur or in low-light environments. Neuromorphic cameras provide asynchronous visual streams with high temporal resolution and a wide dynamic range, offering a promising solution for object detection under challenging conditions. Despite the development of numerous models and the emergence of various applications in neuromorphic object detection, there is still a lack of deep understanding and standardized benchmarks to assess progress and address key challenges. In this paper, we provide a comprehensive survey and benchmark of existing neuromorphic object detection algorithms. Specifically, we first present a problem description, review the available datasets, and revisit the evaluation metrics. We then explore existing neuromorphic object detection approaches from various perspectives, including event representation, temporal modeling, multimodal fusion, asynchronous processing, low-latency processing, and energy-efficient computing. Furthermore, we evaluate a wide range of representative neuromorphic object detection models and offer detailed analyses of the comparative results. Finally, we discuss unresolved issues in neuromorphic object detection and propose potential future research directions. We hope this survey and benchmark will be a valuable resource for researchers and provide guidance for future advancements in neuromorphic object detection.

cs.CV

Towards Ultrafast Depth Sensing Via Active Event-based Stereo Vision

Conventional frame-based imaging for active stereo systems has encountered major challenges in fast-motion scenarios. However, how to design a novel paradigm for ultrafast depth sensing remains an open issue. In this paper, we propose a novel problem setting, namely active event-based stereo vision, which attempts to integrate binocular event cameras and an infrared 2D pattern projector for high-speed dense depth sensing. Technically, we first build a stereo camera prototype system and present a real-world dataset with over 21.5k spatiotemporal synchronized labels at 15 Hz, while also establishing a realistic synthetic dataset with stereo event streams and 23.8k synchronized labels at 20 Hz. Then, we propose ActiveEventNet+, a lightweight yet effective event-based stereo matching neural network that learns to generate high-quality dense disparity maps from stereo event streams with low latency. Our ActiveEventNet+ mainly involves three innovations: incorporating lightweight blocks into event-based stereo matching frameworks, designing a novel cost volume with dynamic interactions between stereo pairs, and presenting an effective temporal consistency architecture to fully use rich temporal cues in event streams. The results show that our ActiveEventNet+ outperforms state-of-the-art methods while significantly reducing computational complexity. Our solution offers superior depth sensing performance compared to conventional frame-based stereo cameras in high-speed scenes. In particular, the lightweight ActiveEventNet enables the prototype system to achieve real-time processing at speeds up to 150 FPS. We believe that this novel active event-based stereo vision paradigm can provide new insights into the design of future high-speed depth sensing camera systems. Our dataset and code can be available at https://github.com/jianing-li/active_event_based_stereo.

cs.RO

Explaining and Tuning Transformer-based LLMs in Arithmetic Tasks with Human Strategies

Transformer-based large language models (LLMs) continue to achieve state-of-the-art performance across various natural language processing tasks. However, their subpar performance on seemingly elementary problems, such as basic arithmetic, raises concerns about model reliability, safety, and ethical deployment. In this study, we demonstrate that the performance of a vanilla Transformer model trained on integer arithmetic tasks can be improved using methods effective for human learners. We begin by decomposing the arithmetic task into well-defined subtasks and conducting loss convergence order analysis together with ablation studies for each subtask. Our findings reveal that LLMs exhibit learning patterns similar to those of human learners, with a faster learning speed for simpler subtasks compared to more complex ones. In addition, we successfully improved the accuracy of LLMs by applying problem-solving strategies and cognitive empowerment methods shown to enhance the performance of human learners. This suggests that transformer-based LLMs may share cognitive processes with human learners in arithmetic. Lastly, we provide a comprehensive demonstration of our method's effectiveness, including significant accuracy improvement experiments, visualization verification, and explanation-based analysis to illuminate the intricacies of LLMs in arithmetic learning. In general, this work explores the potential similarities between transformer-based LLMs and human learners, supported by explainable AI (XAI) verifications, ultimately fostering trust in LLMs for critical and high-stakes applications.

cs.LG

Occ-VLM: Occupancy Grounded Vision Language Model for Indoor Scene Understanding

Recently, vision-language models (VLMs) have made significant progress in 3D scene understanding, driving advances in applications such as embodied intelligence and robotic vision. However, existing approaches typically either rely directly on explicit 3D inputs (e.g., point clouds or RGB-D sequences), or introduce an additional 3D geometry encoder to derive 3D-aware visual tokens from 2D images. Such designs structurally decouple 3D geometric perception from the rich 2D semantics learned via vision-language pre-training, hindering the development of a unified 3D vision-language representation. In this work, we propose Occ-VLM, a novel framework for 3D scene understanding that operates purely on posed RGB images and employs a single 2D vision encoder. Specifically, Occ-VLM reconstructs 3D scene occupancy as an auxiliary geometric prior, which is utilized to spatially associate foreground 2D tokens with 3D space. These tokens are then decoded by a Large Language Model (LLM) for unified scene understanding. Extensive experiments demonstrate that Occ-VLM achieves both accurate geometric perception and robust vision-language reasoning: it attains state-of-the-art performance on multi-view occupancy prediction, while performing on par with 3D-input VLMs on 3D Visual Question Answering (VQA) and 3D dense captioning benchmarks.

cs.CV

Agentic Discovery of Non-Canonical Antimicrobial Peptides with AMPGAN v3

Antimicrobial resistance causes to over a million deaths annually. Antimicrobial peptides (AMPs) are a promising solution, but generative AMP models are not yet ready to design peptides with non-natural amino acids and/or chemical modifications, which are essential for real-world peptide drugs. We present AMPGAN v3, a multi-objective conditional GAN that expands the generative vocabulary to D-amino acids and N/C-terminus modifications such as amidation. By separating adversarial and activity-aware supervision across two specialized discriminators, AMPGAN v3 substantially improves training stability and outperforms prior generative AMP models on external classifiers. We validated five candidates spanning three structural classes in vitro; two showed activity against Gram-positive strains, with the best candidate reaching MIC 8 μg/mL against B. subtilis. To support downstream curation, we further present PepCraft, a multi-agent framework for end-to-end AMP discovery in which a Planning Agent orchestrates specialized executors for generation, filtering, and verification. Its prioritization recommendations align with our in vitro outcomes. Together, these contributions let us examine, on a small but real scale, how generative and agentic AI compose in therapeutic peptide discovery. Code: https://github.com/marszzibros/AMPGANv3

q-bio.QM

Fermi-liquid view of viscosity in cold and dense nucleon matter

We develop a framework to calculate transport properties in cold, dense relativistic quasiparticle system within the Fermi-liquid theory at the mean-field level. Building on our previous study J. Li \emph{et al.} [Phys. Rev. C \textbf{111}, 044904 (2025)], we start from the linearized relativistic Boltzmann equation tailored to quasiparticles with medium-dependent dispersion relation and implement Landau matching conditions, proving that the bulk viscosity is manifestly nonnegative. A low-temperature expansion then yields leading-order ($T/μ^*$) expressions for the shear ($η$) and bulk ($ζ$) viscosities, where the behavior $ζ/η\propto (T/μ^*)^4$ in the degenerate regime is found to be robust against quasiparticle mass correction. We couple the kinetic framework to a Walecka-type mean-field equation of state and compute $η$ and $ζ$ for cold, dense nucleon matter. The transport properties of nucleonic matter in the degenerate regime can be relevant for intermediate beam-energy nuclear experiments.

nucl-th

NAPPure: Adversarial Purification for Robust Image Classification under Non-Additive Perturbations

Adversarial purification has achieved great success in combating adversarial image perturbations, which are usually assumed to be additive. However, non-additive adversarial perturbations such as blur, occlusion, and distortion are also common in the real world. Under such perturbations, existing adversarial purification methods are much less effective since they are designed to fit the additive nature. In this paper, we propose an extended adversarial purification framework named NAPPure, which can further handle non-additive perturbations. Specifically, we first establish the generation process of an adversarial image, and then disentangle the underlying clean image and perturbation parameters through likelihood maximization. Experiments on GTSRB and CIFAR-10 datasets show that NAPPure significantly boosts the robustness of image classification models against non-additive perturbations.

cs.CV

Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training

Reasoning post-training improves Large Language Models (LLMs) on complex tasks such as mathematics and coding, but its benefits across diverse multimodal tasks remains uncertain. The trend of releasing parallel "Instruct" and "Thinking" models by leading teams is both resource-intensive and user-unfriendly. Prior work finds that the gains from reasoning training are influenced by multiple factors, such as base model capabilities, task characteristics, and Chain-of-Thought (CoT) data quality. However, principled criteria for determining when reasoning post-training is beneficial and which data should support it are still lacking. In this paper, we propose Dual Tuning, a reasoning efficacy-driven data curation framework for multimodal LLMs training. Given a target task and a base model, Dual Tuning jointly evaluates whether the training data is beneficial and whether reasoning training with current CoT content yields positive gains over non-reasoning alternatives. We apply Dual Tuning across spatial, mathematical, and multi-disciplinary tasks, and further analyze how reinforcement learning and thinking patterns affect reasoning efficacy. The Dual Tuning results guide data curation by identifying data that benefit reasoning training, data better suited to direct-answer training, and data that are detrimental under both training modes. Our work provides quantitative criteria for selecting appropriate training data and matching post-training strategies.

cs.CL

SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition

Spatial cognition is fundamental to real-world multimodal intelligence, allowing models to effectively interact with the physical environment. While multimodal large language models (MLLMs) have made significant strides, existing benchmarks often oversimplify spatial cognition, reducing it to a single-dimensional metric, which fails to capture the hierarchical structure and interdependence of spatial abilities. To address this gap, we propose a hierarchical spatial cognition framework that decomposes spatial intelligence into five progressively complex levels from basic observation to high-level planning. Building upon this taxonomy, we construct SpatialBench, a large-scale, fine-grained benchmark covering 15 tasks aligned with these cognitive levels. To provide a unified evaluation across heterogeneous tasks, we further introduce a high-level capability-oriented metric that reliably assesses a model's overall spatial reasoning ability. Extensive experiments over massive MLLMs reveal distinct performance stratification across cognitive levels: models exhibit strong perceptual grounding yet remain limited in symbolic reasoning, causal inference, and planning. Additional human tests demonstrate that humans perform selective, goal-directed abstraction, while MLLMs tend to over-attend to surface details without coherent spatial intent. Our work establishes the first systematic framework for measuring hierarchical spatial cognition in MLLMs, laying the foundation for future spatially intelligent systems.

cs.AI