arXiv ScienceSearch

arXiv subjects

Tingting Zhang

Publications and source records attributed to Tingting Zhang.

At least 19 recordsLinked to original sources

A Central Disulfide Junction Drives Transient Network Formation in Elastin-Like Polypeptides, Enabling Low-Concentration Hydrogels

The self-assembly of associative triblock copolymers composed of a central hydrophilic elastin-like polypeptide (ELP) block and short fatty acid end groups (C16) was investigated in aqueous solution. In one system, the ELP contains 80 pentapeptide units (C16-80-C16), whereas in the other two C16-ELP40 chains were oxidatively coupled through their terminal cysteine residues to form a central disulfide bond, yielding C16-(40)2-C16. Despite their nearly identical molecular weights and compositions, the two polymers exhibit markedly different self-assembly behaviors. C16-80-C16 forms large hydrophobic aggregates that remain kinetically trapped and do not develop a dynamically connected network. In contrast, C16-(40)2-C16 forms very small associative nodes with an aggregation number of only $\sim$3 chains. These nodes coexist with larger clusters and become dynamically interconnected at higher concentrations, leading to transparent hydrogels at concentrations as low as 2.5 wt %. Oscillatory rheology reveals a transient Maxwell network governed by a single relaxation process associated with the reversible association of the C16 end groups. SAXS, light scattering, cryo-TEM, and molecular modeling consistently support a model in which the central disulfide junction promotes transient network formation.

cond-mat.soft

MinerU.Chem: A High-Precision System for Optical Chemical Structure and Reaction Recognition

In organic chemistry papers and patents, molecular structures, reaction schemes, and experimental conditions are often presented as molecular structure depictions, reaction diagrams, and complex tables or figures. Such information is difficult for general-purpose document parsing systems to directly convert into machine-readable data. This limits data production for organic chemistry knowledge base construction and for AI for Chemistry tasks such as reaction prediction, retrosynthesis, condition recommendation, molecular property prediction, and drug molecule design. This report introduces MinerU-Chem, a document parsing system for organic chemistry literature integrated into the MinerU online platform. Built on top of MinerU's general document parsing pipeline, MinerU-Chem adds five chemistry-specific modules: chemistry relevance filtering, molecular structure detection, molecule identifier extraction, molecular structure recognition, and reaction scheme parsing. Together, these modules convert organic-chemistry-related image regions in documents into a Molecule Summary List and a Reaction Summary List. For molecular structure recognition, MinerU-Chem uses CARBON (Complex Atomic Representation and Bonding Object Notation) as its core representation. CARBON enables recognition results to preserve both the visual layout of the original image and complex chemical semantics, while supporting the export of standard downstream formats such as MolFile and SMILES. On the SMILES-evaluable subset of MolRecBench-Wild (N=2,392), MinerU-Chem's molecular structure recognition module achieves a SMILES exact-match accuracy of 93.02%, outperforming the best evaluated comparison system, GPT-5.6-Sol (74.87%), by 18.15 percentage points. The system has been integrated into the MinerU online platform and is available at https://mineru.net/OpenSourceTools/Extractor .

cs.CV

Memory Layer: Train the In-Model Cache for Recommendation Models

Early ranking stages in recommendation systems precompute item embeddings and cache them in-model for scoring within strict latency constraints. Because this cache exists only at serving time, outside the training loop, training and serving use different item representations, a structural discrepancy that limits quality and adds operational fragility. We show that co-designing the training and serving paths removes this representation discrepancy at its source. We introduce the memory layer, an in-model key-value embedding cache co-trained with the model: the item tower writes embeddings during training and the model reads them at serving, one source of truth for item representations by construction. Always-on embeddings cover items not yet cached, so every item receives a prediction, and the design consolidates three separate trainer-to-predictor update paths into a single self-contained pipeline. Deployed in production on Instagram Reels, the memory layer raises prediction coverage from 96% to 100%, improves embedding freshness from $O(5\text{ min})$ to $O(20\text{ s})$, and narrows the training-serving Normalized Entropy (NE) gap by up to 86%, yielding over $2\times$ recall for the freshest content and a 5-6% cold start engagement lift. Because embeddings are produced during training, the system needs no separate bulk-evaluation or publish-time recomputation, cutting training-and-publish computational cost by 30% at neutral serving computational cost.

cs.IR

Towards Effcient Low Altitude Sensing: A Dual Heterogeneous Graph Learning Method for UAV Task Allocation

With the development of low altitude intelligent systems, multiple unmanned aerial vehicles (UAVs) can collaboratively execute more complex tasks. Conventional task allocation methods usually regard tasks and UAVs as isolated entities, making it difficult to capture task dependencies and UAV communication relationships. To address this issue, this paper proposes a dual heterogeneous graph learning based UAV task allocation method. A directed task graph is constructed to represent task dependencies and encode task resource requirements, while an undirected UAV communication graph is built to model communication relationships and encode UAV resource states. The task allocation problem is formulated as a structural matching problem between the task graph and the UAV communication graph. A graph attention network based feature extraction method is introduced to learn structural representations from both graphs through message passing. A cross attention mechanism is further integrated with proximal policy optimization to optimize the matching between task nodes and UAV nodes for task allocation. Simulation results demonstrate that the proposed method achieves a higher task completion rate and shorter task completion time than benchmark methods under different evaluation settings. Furthermore, a UAV sensing and computing application is developed on the AirSim simulation platform. A large language model is employed to convert natural language task requirements into a structured task graph for autonomous UAV task execution, demonstrating the potential of the proposed framework for natural language driven UAV mission planning and execution.

eess.SY

Handover-Aware Trajectory Planning for Cellular-Connected UAVs under STL Specifications and URLLC Constraints

This paper investigates handover-aware trajectory planning for cellular-connected UAVs executing mission-centric tasks under ultra-reliable low-latency communication (URLLC) constraints. Signal temporal logic (STL) provides a formal specification layer for translating mission semantics into time-bounded trajectory requirements, while finite-blocklength URLLC feasibility characterizes reliable command-and-control (C2) links with serving base stations (BSs). We formulate a joint planning problem that optimizes the UAV trajectory, STL mission satisfaction, serving-BS association, and handover behavior. To solve this mixed discrete-continuous problem, we adopt and integrate a Logic Network Flow (LNF) based STL reformulation with B\'ezier-parameterized motion, disk-shaped URLLC service regions, and binary association variables, so that the resulting mixed-integer quadratically constrained formulation can be solved by standard branch-and-bound solvers. Numerical simulations over a library of STL missions show that the proposed planner can execute different mission specifications under the same cellular map while maintaining URLLC feasibility. The results further reveal how mission timing, handover-aware association, and finite-blocklength stringency jointly affect trajectory shape, serving margin, and computational complexity.

eess.SY

Quantum percolation based dynamic propagation connectivity for critical-area identification in transport networks

Transport networks often lose functionality through gradual degradation in link operating conditions before topological disconnection occurs. Link-centred and binary percolation measures identify important facilities or connectivity failures, but they provide limited information on which spatial areas cause the largest loss of network-wide propagation capability. This paper develops a Dynamic Propagation Connectivity (DPC) metric based on quantum percolation for critical-area identification in transport networks. Time-varying link travel times are converted into continuous propagation strengths, which define a Hermitian propagation operator at each observation time. Candidate regions are then evaluated by a regional degradation experiment that measures the resulting loss of DPC. The method is applied to a benchmark Sioux Falls network and six Florida road networks during the post-Hurricane Irma disruption and recovery period, using 1,281 five-minute observation times. The benchmark confirms that the regional DPC score identifies a predefined structurally critical corridor. In the Florida networks, the identified critical areas differ from regions selected by link count, local degradation, edge betweenness, algebraic connectivity, and classical percolation. In Networks 1 to 4, DPC and classical percolation rankings have negative Spearman correlations, showing that continuous propagation degradation and binary fragmentation reveal different vulnerability patterns. Robustness tests under alternative travel time scaling, degradation strength, and grid size show stable results, with mean rank agreement between 0.84 and 0.96. The findings extend transport resilience analysis based on percolation from binary connectivity loss to continuous propagation degradation and provide a spatial diagnostic tool for regional monitoring, emergency planning, and recovery prioritisation.

quant-ph

Optimal Consumption and Retirement Time under Shortfall Risk Measure

This paper studies the optimal portfolio, consumption, and endogenous early retirement problem within a benchmark tracking framework by incorporating a new relative performance evaluation. In this framework, the investor maximizes expected lifetime consumption utility while managing the maximum wealth shortfall relative to a benchmark, with shortfall-management costs that may differ before and after retirement. Mathematically, the problem is a hybrid stochastic control problem involving both regular controls and an optimal stopping time, in which the running maximum process records the investor's largest benchmark shortfall. We introduce an auxiliary reflected state process and establish an equivalent hybrid stochastic control problem. By proving the convex duality theorem, we technically transform the original problem into a two-dimensional pure optimal stopping problem with state reflection. This enables us to characterize the geometric structure of the stopping set and derive the feedback-form optimal retirement boundary, as well as optimal portfolio and consumption policies. Analytical examples and numerical simulations reveal a two-stage structure with more conservative investment and more aggressive consumption after retirement. Driven by the retirement option, the expected largest shortfall risk follows a pronounced U-shaped pattern with respect to wealth. Shortfall management costs, labor income, and leisure preference significantly influence retirement timing, investment, and consumption.

math.OC

FAM-Bench: A Multimodal Benchmark for Condition-Aware Food-as-Medicine Reasoning

Food-as-Medicine requires models to reason beyond what a dish is or what nutrition it contains: they must decide whether a concrete food choice is appropriate for a specific health condition. Existing food AI benchmarks primarily evaluate dish recognition, recipe understanding, nutrient estimation, or general nutrition question answering, leaving this health-aware decision layer largely untested. We introduce FAM-Bench, a multi-modal Food-as-Medicine benchmark with 2500 nutrition-expert-verified instances across 13 diet-related health conditions. The benchmark contains two complementary tasks: dish-level suitability assessment, where models judge whether a dish is suitable for a condition from its image and ingredient list, and comparative dish analysis, where models rank four candidate dishes by condition-specific suitability. Both tasks require integrating ingredient evidence, visual preparation cues, and clinical nutrition constraints, providing a standardized testbed for grounded health-aware reasoning in language and vision-language models.

cs.AI

PetroBench: A Benchmark for Large Language Models in Petroleum Engineering

Large Language Models are increasingly applied in the petroleum industry, highlighting the need for a domain-specific evaluation framework. This study develops a benchmark for LLMs in petroleum engineering, including a three-stage process of data preprocessing, quality filtering, and multi-model validation. Using expert review, a standardized question bank with strong domain relevance and discriminative capability was constructed. The benchmark covers production, reservoir, and drilling engineering, with 1,200 questions across multiple-choice, true or false, term definition, and short-answer formats. Eight mainstream LLMs were evaluated under a unified API environment. Results show that models performed better on subjective than objective questions, indicating weaknesses in factual knowledge discrimination. The highest accuracies for multiple-choice and true or false questions were 65.3% and 74.3%, respectively. Gemini-3-Pro, Kimi-K2.5, and Claude-Opus-4.6-Thinking achieved the best overall scores of 72%-74%. Models performed best in production engineering and weakest in reservoir engineering. Chinese models showed advantages in multiple-choice questions, while international models performed slightly better in short-answer questions. The benchmark provides a reproducible and practical reference for evaluating and deploying LLMs in petroleum engineering.

cs.AI

SCORP: Scene-Consistent Multi-agent Diffusion Planning with Stable Online Reinforcement Post-Training for Cooperative Driving

Cooperative driving is a safety- and efficiency-critical task that requires the coordination of diverse, interaction-realistic multi-agent trajectories. Although existing diffusion-based methods can capture multimodal behaviors from demonstrations, they often exhibit weak scene consistency and poor alignment with closed-loop cooperative objectives. This makes post-training necessary for further improvement, yet achieving stable online post-training in reactive multi-agent environments remains challenging. In this paper, we propose SCORP, a scene-consistent multi-agent diffusion planner with stable online reinforcement learning (RL) post-training for cooperative driving. For pre-training, we develop a scene-conditioned multi-agent denoising architecture that couples inter-agent self-attention with a dual-path conditioning mechanism: cross-attention provides direct scene-information injection, while AdaLN-Zero enables additional flexible and stable conditional modulation, thereby improving the scene consistency and road adherence of joint trajectories. For post-training, we formulate a two-layer Markov decision process (MDP) that explicitly integrates the reverse denoising chain with policy-environment interaction. We further co-design dense, well-shaped planning rewards and variance-gated group-relative policy optimization (VG-GRPO) to mitigate advantage collapse and gradient instability during closed-loop training. Extensive experiments show that SCORP outperforms strong open-source baselines on WOMD, with 10.47%-28.26% and 1.70%-7.22% improvements in core safety and efficiency metrics, respectively. Moreover, compared with alternative post-training methods, SCORP delivers significant and consistent gains in both driving safety and traffic efficiency, highlighting stable and sustained advances in closed-loop cooperative driving.

cs.RO

Evaluation Before Generation: A Paradigm for Robust Multimodal Sentiment Analysis with Missing Modalities

The missing modality problem poses a fundamental challenge in multimodal sentiment analysis, significantly degrading model accuracy and generalization in real world scenarios. Existing approaches primarily improve robustness through prompt learning and pre trained models. However, two limitations remain. First, the necessity of generating missing modalities lacks rigorous evaluation. Second, the structural dependencies among multimodal prompts and their global coherence are insufficiently explored. To address these issues, a Prompt based Missing Modality Adaptation framework is proposed. A Missing Modality Evaluator is introduced at the input stage to dynamically assess the importance of missing modalities using pretrained models and pseudo labels, thereby avoiding low quality data imputation. Building on this, a Modality invariant Prompt Disentanglement module decomposes shared prompts into modality specific private prompts to capture intrinsic local correlations and improve representation quality. In addition, a Dynamic Prompt Weighting module computes mutual information based weights from cross attention outputs to adaptively suppress interference from missing modalities. To enhance global consistency, a Multi level Prompt Dynamic Connection module integrates shared prompts with self attention outputs through residual connections, leveraging global prompt priors to strengthen key guidance features. Extensive experiments on three public benchmarks, including CMU MOSI, CMU MOSEI, and CH SIMS, demonstrate that the proposed framework achieves state of the art performance and stable results under diverse missing modality settings. The implementation is available at https://github.com/rongfei-chen/ProMMA

cs.CV

UAV Control and Communication Enabled Low-Altitude Economy: Challenges, Resilient Architecture and Co-design Strategies

The emerging low-altitude economy has catalyzed the large-scale deployment of unmanned aerial vehicles (UAVs), driving a paradigm shift in environment monitoring, logistics, and emergency response. However, operating within these environments presents notable challenges as pervasive coverage holes, unpredictable interference, and spectrum scarcity. To this end, this article present a communication and control co-design framework to enable a resilient architecture for cellular-connected UAVs. Specifically, we first characterize typical service applications and their stringent performance requirements, followed by a comprehensive analysis of the unique challenges. To bridge the gap between volatile wireless links and rigid flight stability, a three layered architecture is proposed, integrating pre-flight strategic planning, in-flight adaptive action, and system-level resource orchestration. Furthermore, we detail the key enabling technologies for communication and control co-design. Preliminary case studies are proposed to validate that the co-design framework significantly improve the resilience of cellular-connected UAV systems, providing a robust foundation for the evolution of intelligent low-altitude networks.

cs.NI

Data-driven identification of critical links in transport networks using quantum annealing

In urban transport systems, time-varying demand and network conditions cause the importance of infrastructure elements to evolve, requiring the identification of period-specific critical links to support systemlevel risk and resilience analysis. However, static or time-averaged network analyses struggle to capture the temporal variation of infrastructure importance at the city scale. To address this gap, this study proposes a time-dependent critical link identification framework for large-scale urban transport networks. The problem is formulated as a Quadratic Unconstrained Binary Optimisation (QUBO) model and solved using quantum annealing on D-Wave hardware. Empirical analysis using real-world traffic data reveals a strong temporal concentration of critical links. Rather than persistently influencing system performance, critical links emerge mainly within a small number of key time windows, during which even limited disruptions can lead to substantial network delay amplification. These findings demonstrate the value of time-dependent analysis for risk screening, stress testing, and resilience-oriented transport management.

math.OC

LLM-Enabled Low-Altitude UAV Natural Language Navigation via Signal Temporal Logic Specification Translation and Repair

Natural language (NL) navigation for low-altitude unmanned aerial vehicles (UAVs) offers an intelligent and convenient solution for low-altitude aerial services by enabling an intuitive interface for non-expert operators. However, deploying this capability in urban environments necessitates the precise grounding of underspecified instructions into safety-critical, dynamically feasible motion plans subject to spatiotemporal constraints. To address this challenge, we propose a unified framework that translates NL instructions into Signal Temporal Logic (STL) specifications and subsequently synthesizes trajectories via mixed-integer linear programming (MILP). Specifically, to generate executable STL formulas from free-form NL, we develop a reasoning-enhanced large language model (LLM) leveraging chain-of-thought (CoT) supervision and group-relative policy optimization (GRPO), which ensures high syntactic validity and semantic consistency. Furthermore, to resolve infeasibilities induced by stringent logical or spatial requirements, we introduce a specification repair mechanism. This module combines MILP-based diagnosis with LLM-guided semantic reasoning to selectively relax task constraints while strictly enforcing safety guarantees. Extensive simulations and real-world flight experiments demonstrate that the proposed closed-loop framework significantly improves NL-to-STL translation robustness, enabling safe, interpretable, and adaptable UAV navigation in complex scenarios.

cs.RO

Overlap-Summation-Based Pulse Shaping Transceiver for Affine Frequency Division Multiplexing

Affine frequency division multiplexing (AFDM) has recently emerged as a promising waveform for doubly-selective channels. A direct-windowing-based pulse shaping transceiver (PS-AFDM) was proposed to suppress the Doppler sidelobes, thus improving the accuracy of channel estimation. We observe that the legacy PS-AFDM significantly increases the condition number of the effective channel matrix when path delay and Doppler parameters are randomly distributed, such ill-conditioning leads to the degradation in the solution stability of channel equalization under noisy conditions, thus resulting in the degradation of bit error rate (BER). To address this issue, this letter applies the existing weighted overlap-summation (WOLA) transceiver to AFDM and proposes a novel channel-aware (CA) receive shaping window design, which simultaneously achieves accurate channel estimation and robust equalization performance, at the cost of time-domain prefix and receive window calculation overheads. The resulting scheme is termed CAWOLA-AFDM. Compared with the legacy WOLA scheme, which employs a fixed receive window, the proposed CAWOLA design exploits the non-instantaneous channel estimates of power gains and Doppler shifts to design a channel-tailored and closed-form receive shaping window, thereby further enhancing channel estimation accuracy while maintaining the channel condition number when a Nyquist prototype window is adopted. The proposed CAWOLA receive window design aims to provide the effective AFDM channel with a more compact DAFT-domain support. The source codes for the simulations are provided at https://github.com/SANIS-HITSZ/Waveform_AFDM.

eess.SP

Uncertainty-Aware 3D UAV Tracking Using Single-Anchor UWB Measurements

In this letter, we present an uncertainty-aware single-anchor Ultra-Wideband (UWB)-based 3D tracking framework. Specifically, a mobile Unmanned Aerial Vehicle (UAV) maintains a desired standoff distance to a moving target using range and 3D bearing measurements from a multi-antenna UWB anchor rigidly mounted on the UAV. To enhance the stability and safety under measurement degradation and motion uncertainty, we jointly design a robust factor-graph-based target localization method and a covariance-aware control Lyapunov function--control barrier function (CLF--CBF) tracking controller. This controller adaptively adjusts distance bounds and safety margins based on the posterior target covariance provided by the factor graph. The proposed system is evaluated through numerical simulations and real-world experiments carried out in a narrow indoor corridor environment.

eess.SY

RIS-Assisted Coordinated Multi-Point ISAC for Low-Altitude Sensing Coverage

The low-altitude economy (LAE) has emerged and developed in various fields, which has gained considerable interest. To ensure the security of LAE, it is essential to establish a proper sensing coverage scheme for monitoring the unauthorized targets. Introducing integrated sensing and communication (ISAC) into cellular networks is a promising solution that enables coordinated multiple base stations (BSs) to significantly enhance sensing performance and extend coverage. Meanwhile, deploying a reconfigurable intelligent surface (RIS) can mitigate signal blockages between BSs and low-altitude targets in urban areas. Therefore, this paper focuses on the low-altitude sensing coverage problem in RIS-assisted coordinated multi-point ISAC networks, where a RIS is employed to enable multiple BSs to sense a prescribed region while serving multiple communication users. A joint beamforming and phase shifts design is proposed to minimize the total transmit power while guaranteeing sensing signal-to-noise ratio and communication spectral efficiency. To tackle this non-convex optimization problem, an efficient algorithm is proposed by using the alternating optimization and semi-definite relaxation techniques. Numerical results demonstrate the superiority of our proposed scheme over the baseline schemes.

eess.SY

Handover-Aware URLLC UAV Trajectory Planning: A Continuous-Time Trajectory Optimization via Graphs of Convex Sets

In this paper, we study a cellular-connected unmanned aerial vehicle (UAV) which aims to fly between two predetermined locations while maintaining ultra-reliable low-latency communications (URLLC) for command-and-control (C2) links with terrestrial base stations (BSs). Long-range flights often trigger frequent inter-cell handovers, which may introduce delays and synchronization overhead. We jointly optimize the continuous trajectory and BS association to minimize handovers, path length, and flying time, subject to communication reliability and kinematic constraints. To address this problem, we reformulate it as an optimization based on the graph of convex sets (GCS). First, the URLLC requirement is translated into spatially feasible regions in the flight plane for each BS. And an intersection graph is constructed including the start and goal points. Each graph node is associated with a smooth and dynamically feasible trajectory segment. The trajectory is parameterized in space by B\'ezier curves and in time by a monotonic B\'ezier scaling, together with convex constraints that ensure continuity and enforce speed bounds. Next, we impose unit-flow constraints to enforce a single path, and by coupling the resulting binary edge-selection variables with the convex constraints, we obtain a mixed-integer convex program (MICP). Applying a convex relaxation and rounding to the mixed-integer convex program produces nearly globally optimal routes, and a final refinement yields smooth, dynamically feasible trajectories. Simulations verify that the method preserves URLLC connectivity while achieving a clear trade-off between fewer handovers and flight efficiency.

eess.SY