arXiv ScienceSearch

arXiv subjects

Wei Xie

Publications and source records attributed to Wei Xie.

At least 19 recordsLinked to original sources

Minor-Order Exponent Profiles of Oscillatory Matrices

For an $n\times n$ oscillatory matrix $A$, let $e_k(A)$ be the least positive integer $m$ for which every minor of order $k$ of $A^m$ is positive. We determine the profile $(e_1(A),\ldots,e_n(A))$. For every nonsingular totally nonnegative matrix, the column index sets of positive entries in each compound row form an interval in the componentwise order. The endpoint maps are order-preserving and compose under multiplication. This yields a two-corner criterion for positivity of all minors of a fixed order. Consequently, $e_k(A)$ is the larger of the first positivity times of two remote corner minors. We prove the sharp inequalities $e_k(A)\leq\max{k,n-k}$ for $1\leq k<n$ and $|e_{k+1}(A)-e_k(A)|\leq1$ for $1\leq k\leq n-2$. For $D=\operatorname{diag}(1,-1,1,-1,\ldots)$, the matrix $DA^{-1}D$ is oscillatory and satisfies $e_k(DA^{-1}D)=e_{n-k}(A)$ for $1\leq k<n$; the determinant exponent remains one. The ordered positions of the positive factors in an adjacent bidiagonal factorization determine the profile independently of their values. Each one-sided factor sequence can be represented by a single permutation, giving a finite characterization of all profiles and an integer unimodular realization of each. We also obtain a compatibility condition across minor orders: if $3\leq k\leq n-3$ and $e_k(A)\leq2$, then $e_j(A)\leq2+\lceil |j-k|/2\rceil$ for $2\leq j\leq n-2$. In particular, $e_k(A)\leq2$ implies $e_{k+2}(A)\leq3$ for $3\leq k\leq n-4$, and both the constant and this range are sharp.

math.CO

Electron-phonon coupled hydrodynamics in semimetal TaAs_2

Hydrodynamic corrections to diffusive transport can arise when momentum-conserving collisions between quasiparticles become prominent, and they have been documented for both electrons and phonons. An emerging frontier topic is coupled electron-phonon (e-ph) hydrodynamics. Here, through electrical and thermal transport measurements on TaAs2 crystals with different impurity levels, we document the emergence of an e-ph bifluid in the temperature window of 5 to 15 K. Within this range, the lattice thermal conductivity exhibits a faster-than-T^3 temperature dependence, as a consequence of non-monotonic and purity-dependent phonon mean free paths, a signature of phonon Poiseuille flow. However, strong e-ph coupling impedes the emergence of a ballistic regime. This is corroborated by the observation of quantum oscillations in the lattice thermal conductivity. Prominent phonon-mediated momentum exchange between electrons amplifies the violation of the Wiedemann-Franz law and yields a two-order-of-magnitude discrepancy between quantum and transport lifetimes, a signature of electron hydrodynamics in semimetals. Our results imply that in semimetals with optimized e-ph coupling, thanks to matching between the cryogenic phonon wave?length and the Fermi wavelength, momentum and energy flow between the electron and phonon reservoirs as frequently as within each reservoir.

cond-mat.mtrl-sci

PrimSynth: An Agentic Approach to Discover, Validate, and Synthesize Exploit Primitives for Linux Kernel Vulnerabilities

Linux kernel vulnerabilities are critical to downstream systems. Despite extensive research on automated kernel exploitation, a fundamental challenge remains the conceptual gap between abstract exploit strategies and concrete technical operations. To fill this gap, this paper introduces a systematic characterization that formalizes six classes of exploit primitives from logical capability to validatable effect. Then, an extended exploit strategy representation is proposed, which couples primitive upgrading strategies with primitive path code synthesis rules governing object constraints, temporal sequencing, environment prerequisites, and validation constraints. Building upon this foundation, this paper presents \textsc{PrimSynth}, a multi-agent framework that encapsulates these representations through coordinated agents to discover, validate, and synthesize exploit primitives for memory corruption vulnerabilities in the Linux kernel. These agents operate in an iterative closed loop until valid primitives are found, leveraging validation signals as evidence of exploitable state transitions to ground primitive synthesis decisions. An automated method for extracting and validating primitives is also proposed based on vulnerability-directed execution and a rebootable validation environment. \textsc{PrimSynth} is evaluated on 16 real-world Linux kernel CVEs spanning 5 vulnerability types. Experimental results show that PrimSynth achieves reliable primitive extraction, maintaining a 100% primitive match rate. For primitive synthesis, PrimSynth successfully synthesizes multi-primitive exploitation chains with 82.4% strategy synthesis rate (SSR) when the public PoC is available and a 61.3% SSR without the guidance of primitive hypotheses.

cs.CR

GlycoMAC: A Multiscale Metabolic-Glycosylation Framework for Predicting Glycosylation Across Conditions in Mammalian Cell Cultures

Antibody productivity and glycosylation quality in CHO cell cultures emerge from a dynamically evolving metabolic environment, yet existing models often work in isolation or at a single scale. Here, we present a multiscale mechanistic framework linking molecular, cellular, and process scales to predict how inputs shape bioprocess trajectories. The framework combines a single-cell kinetic model of metabolism and glycosylation with a stochastic population model that captures environment-dependent transitions among growth, production, and decline states. To characterize metabolic adaptation, we introduce the cumulative variation in oxygen uptake rate, a trajectory-based biomarker that quantifies the total metabolic adjustment experienced during culture. Unlike population-averaged approaches, the model propagates cell-resolved metabolic states (including ammonia-regulated Golgi pH, nucleotide sugar availability, manganese cofactors, and synthesis rates) into glycan processing. The framework was evaluated using CHO-K1 fed-batch cultures producing VRC01 IgG1 under targeted ammonia stress, matched control conditions, and a pyramid-feeding strategy with tighter control. It accurately reproduced trajectories of cell growth, metabolites, productivity, and harvest glycosylation, including increased G0F abundance and reduced galactosylation under ammonia stress. By mechanistically linking process conditions to cell-state dynamics and glycosylation outcomes, the framework provides a unified foundation for digital bioprocessing, predictive biomanufacturing, and advanced process control.

q-bio.CB

Asymptotic Entanglement Hiding under Stabilizer Restrictions

Entanglement is central to quantum information processing, while stabilizer operations underpin fault-tolerant quantum computation. We ask how much entanglement remains visible or distillable under stabilizer restrictions. We quantify stabilizer-visible entanglement by restricting the measured relative entropy of entanglement to stabilizer measurements, thereby obtaining converse bounds on entanglement distillation under stabilizer operations. We demonstrate magic-free asymptotic entanglement hiding: we construct explicit convex mixtures of pure stabilizer states on $N$ qutrits per party whose unrestricted visible entanglement and LOCC-distillable entanglement both grow as $Ω(N/\log N)$, while their stabilizer-visible and stabilizer-distillable entanglement vanish as $N\to\infty$. Thus, an unbounded amount of LOCC-distillable entanglement carried by stabilizer states can become asymptotically invisible and undistillable under stabilizer restrictions. We further prove that stabilizer-visible entanglement is $O(1)$ with high probability for Haar-random pure states despite extensive unrestricted visibility, and vanishes uniformly over entangled Werner states as the local dimension grows through odd primes. These results reveal a fundamental separation between entanglement and magic as resources, exposing intrinsic limits on entanglement extraction using stabilizer operations.

quant-ph

Entrywise Positivity Preservers on Green Matrices

We classify the entrywise functions that preserve positive semidefiniteness on discrete Green matrices \(G(p,q)=(p_{\min(i,j)}q_{\max(i,j)})\) with positive parameters, without requiring the resulting matrix to retain Green structure. For matrices of all orders, the preservers are the zero function and the functions \(f(t)=\int_{[0,\infty)}t^α\,dμ(α)\), where \(μ\) is a nonzero finite positive measure and the integral is finite for every \(t>0\). Requiring the resulting matrix to be totally nonnegative reduces the nonzero preservers to \(f(t)=ct^α\), where \(c>0\) and \(α\ge0\). These power functions also preserve positive semidefinite Green structure, while strict Green structure is preserved precisely when \(α>0\). No regularity assumption is needed for these classifications. We also characterize continuously differentiable functions that are entrywise Loewner monotone on every fixed-\(q\) Green family: this holds precisely when \(f'\) is a positive mixture of nonnegative real powers, with the zero measure allowed.

math.RA

Entrywise Loewner Preservers on Min and Max Matrix Cones

Let \[ A_{\min}(x)=\bigl(x_{\min(i,j)}\bigr)_{i,j=1}^n, \qquad A_{\max}(x)=\bigl(x_{\max(i,j)}\bigr)_{i,j=1}^n \] be the Min and Max matrices generated by a real sequence \(x=(x_1,\ldots,x_n)\). Using their classical cone parametrizations, we give exact characterizations of entrywise maps preserving positive semidefiniteness, total nonnegativity, Loewner order, and Loewner convexity. Our main results concern the Loewner structure.Without assuming continuity or differentiability, entrywise Loewner-order preservation is equivalent to \(f\) being nondecreasing and convex. On the Min and Max cones, requiring the Loewner-convexity inequality on arbitrary pairs is rigid and forces \(f\) to be affine. Under the standard convention of restricting the inequality to Loewner-comparable pairs, the condition automatically forces \(f\in C^1([0,\infty))\) and is equivalent to convexity of both \(f\) and \(f'\). The analogous statement holds for Loewner concavity. In each case, the condition is already detected in dimension two. We also show that \(f:[0,\infty)\to\mathbb R\) preserves positive semidefiniteness entrywise on all positive semidefinite Min or Max matrices if and only if \(f\) is nonnegative and nondecreasing; the same condition characterizes total-nonnegativity preservation. Finally, we determine the power-function ranges and Loewner-order automorphisms, characterize the entrywise preservers of the strict Min and Max classes, and relate these strict cones to inverse \(M\)-matrices and oscillatory matrices.

math.FA

Variance Reduction Based Experience Replay for Policy Optimization

Effective reinforcement learning (RL) for complex stochastic systems requires leveraging historical data to improve sample efficiency and accelerate policy optimization. However, classical experience replay treats all past observations uniformly and fails to account for their varying contributions to learning. To address this limitation, we propose Variance Reduction Experience Replay (VRER), a principled framework that selectively reuses informative samples to reduce the variance of policy gradient estimates. VRER is algorithm-agnostic and can be integrated with existing policy optimization methods, yielding the sample-efficient off-policy algorithm, Policy Gradient with VRER (PG-VRER). To provide rigorous theoretical guarantees, we develop a novel analysis framework for experience replay that explicitly accounts for dependencies induced by Markovian dynamics and behavior-policy interactions. Using this framework, we establish finite-time convergence guarantees for PG-VRER and characterize a fundamental bias-variance trade-off: reusing older samples reduces gradient variance but may introduce greater estimation bias. Extensive experiments show that VRER consistently accelerates learning and outperforms state-of-the-art policy optimization algorithms

stat.ML

Heavy-quark transport across the QCD crossover driven by a lattice-constrained in-medium potential

We present a self-consistent framework for heavy-quark transport in the quark-gluon plasma across the QCD crossover region. By synthesizing perturbative and nonperturbative interactions into a unified interaction kernel, we circumvent the traditional reliance on arbitrary soft-hard momentum separation scales. The interaction is governed by an in-medium effective potential, incorporating short-range Yukawa screening and long-range confining string contributions, both rigorously constrained by the latest lattice QCD data. Our results reveal that the nonperturbative string tension is indispensable for capturing the extreme opacity of the medium near the critical temperature $T_c$. Specifically, our model predicts a spatial diffusion coefficient of $2πT D_s \approx 0.5 \sim 1.7$, demonstrating a striking quantitative agreement with the recent lattice QCD extractions. Ultimately, our results provide a robust dynamical interpretation of the strong heavy-quark coupling near the QCD crossover and offer a unified framework for describing heavy-flavor transport in hot and dense QCD matter.

hep-ph

Generalized Mermin Inequalities for Benchmarking Large-Scale GHZ States

Multipartite Bell tests provide a correlation-only route to benchmarking quantum processors, but their application at large scales is hindered by the rapid decay of many-body correlators under noise and exponentially many terms in conventional Bell expressions. Here we address these scalability obstacles by introducing a finite-setting generalized Mermin family of state-tailored Bell inequalities with analytic certification bounds, in which the measurement-setting number $m$ provides an additional certification dimension complementary to the system size $n$. We show that, for the powers-of-two setting choices considered here, increasing $m$ leaves the ideal normalized multipartite quantum value unchanged while lowering the relevant classical bounds, thereby strengthening the Bell-violation ratios and yielding an improved noise-robustness scaling compared to the standard Mermin inequality. We test this construction experimentally on a programmable superconducting processor by preparing Greenberger-Horne-Zeilinger (GHZ) states of up to 80 qubits. Using randomized sampling for direct Bell-operator estimation, we observe Bell ratios that grow exponentially with system size, certify a nonlocality depth of 14, and show that increasing $m$ strengthens both the Bell ratio and depth certification. All results are obtained solely from measured correlators and analytical bounds, without readout correction, tomography, or model-based mitigation. Generalized Mermin inequalities therefore provide a sharper Bell benchmark for noisy large-scale GHZ states.

quant-ph

Salience Induction against Multi-Hop RAG Agents: Threat and Defense

Agentic retrieval-augmented generation (RAG) systems increasingly retrieve external evidence and orchestrate tools for knowledge-intensive applications. In Multi-Hop question answering, agents chain facts across documents. Existing defenses focus on content poisoning, which injects false facts, and prompt injection, which embeds directives. We identify a third attack surface: the salience channel, through which fact position, emphasis, framing, and semantic proximity can redirect reasoning even when all retrieved claims are true and no instructions are present. We formalize Salience Induction as truth-preserving edits that redirect Multi-Hop attribute binding while leaving the retrieval trace semantically intact. We define six Salience-Editing operator classes and build an iterative proposer-verifier pipeline under factual and stealth constraints. We also introduce SalientWiki-MH, a decoy-annotated Multi-Hop benchmark. Evaluations across five frontier model families (GPT, Claude, Gemini, DeepSeek, and Qwen) and three agent architectures (ReAct, Reflexion, and tool-calling) show broad generalization. Under a 30% edit budget, Salience Induction achieves an 83.3% attack success rate; the strongest evaluated baseline defense leaves 75.7% post-defense ASR. Untargeted rewriting further reduces attacks only by degrading neutral task success. Our lightweight input-side defense, Salience Normalization, reduces attack success to 15.3% under standard attacks and 23.6% under an adaptive attack. These results show that truthfulness and instruction filtering alone are insufficient: robust agentic RAG also requires defenses against salience-relevance decoupling.

cs.CR

Perturbative and nonperturbative properties of heavy quark transport in a thermal SU(3) gluon plasma

We investigate the perturbative and nonperturbative aspects of heavy quark transport in a thermal SU(3) gluon plasma. Based on the soft-hard factorized model, we extend the original perturbative framework to the near-critical temperature region, where nonperturbative effects become significant. The transition behavior of the semi-quark-gluon-plasma (semi-QGP) is described via a temperature-dependent background field incorporated in the background field effective theory. By implementing this approach, we quantitatively evaluate the collisional energy loss and momentum diffusion coefficients of charm and bottom quarks as functions of the incoming energy and medium temperature. Our results show a distinct suppression of both the energy loss and the diffusion coefficients relative to conventional perturbative estimates, especially near the critical temperature. This suppression originates from the emergence of a temperature-dependent color background field, which effectively reduces the color charge screening of the medium. These findings provide important theoretical insight into the phenomenology of heavy-flavor probes, offering a unified theoretical framework applicable across both high- and low-momentum regimes.

hep-ph

Intracellular Measurement-Informed Multiscale Modeling for Scalable iPSC Manufacturing

Scalable manufacturing of human induced pluripotent stem cells (iPSCs) is essential for industrial-scale production of cell therapies and regenerative medicines. However, the 3D aggregate cultures used in manufacturing exhibit substantial spatial and metabolic heterogeneity compared with the relatively homogeneous monolayer systems used in laboratory studies, complicating mechanistic understanding and predictive metabolic modeling across culture scales. To address this challenge, we developed a modular multiscale mechanistic foundation model that links molecular, cellular, and macroscopic processes while accounting for spatial and metabolic heterogeneity. The framework integrates extracellular culture dynamics, intracellular metabolic fluxes, and cellular redox states by extending a previously established monolayer kinetic network and coupling it with a biological systems-of-systems (Bio-SoS) multiscale model for aggregate cultures, incorporating explicit redox interactions. Systematic monolayer and aggregate experiments (including multiple isotopic tracers, extracellular metabolite profiling, and two-photon optical redox imaging) were used to improve and validate the model. This integrated framework unifies heterogeneous datasets across culture configurations and enables mechanistic interpretation of metabolic and redox responses across heterogeneous culture scales, providing a quantitative foundation for scalable iPSC biomanufacturing.

q-bio.CB

Adaptive multiscale model reduction for linear elasticity equation in perforated domains

In this paper, we develop a Constraint Energy Minimizing Generalized Multiscale Finite Element Method (CEM-GMsFEM) for solving linear elasticity problems in heterogeneous perforated domains. The presence of numerous perforations introduces multiple scales into the computational domain, making direct fine-grid simulations computationally expensive. The proposed method follows the standard offline--online decomposition of CEM-GMsFEM. In the offline stage, local spectral problems are solved on coarse elements to construct auxiliary spaces, and localized energy-minimizing basis functions are then computed on oversampled regions to capture fine-scale geometric information induced by the perforations. In the online stage, residual-driven basis functions are constructed in enlarged coarse neighborhoods to incorporate source-term information and improve the accuracy of the multiscale approximation adaptively. We establish convergence results for both the offline and online stages. In particular, we derive error estimates for the localized multiscale approximation and prove the convergence of the adaptive online enrichment algorithm. Moreover, we show that the oversampling regions used in the online stage can be determined locally, leading to a reduction in computational cost while maintaining convergence properties. Numerical experiments on perforated media with different geometric configurations demonstrate the accuracy and efficiency of the proposed method.

math.NA

What drives performance in molecular MPNNs? An operator-level factorial benchmark

Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it difficult to identify how specific message-passing operators affect performance. We present an operator-level factorial benchmark that decomposes 2D molecular MPNNs into the three families of message-seed initialization, node-edge fusion, and node update operators. The resulting 84 configurations are benchmarked on ten MoleculeNet datasets under a shared experimental setup and statistical analysis protocol. Across this controlled design, performance variation is associated primarily with message construction rather than update complexity. Message-seed initialization shows significant family-level effects for both regression and classification, node-edge fusion shows a significant family-level effect for regression with descriptive advantages for concatenation-based mixing, and the update family shows no statistically supported effect for either endpoint family. A representation probe into the Quinethazone molecule further demonstrates that concatenation-based mixing can better differentiate chemically distinct heteroatoms and withstand oversmoothing than Hadamard gating. Representative configurations selected separately for classification and regression recover competitive performance relative to established molecular graph neural network (GNN) baselines, ranking numerically best on eight of ten benchmark datasets. These empirical results are interpreted through concise mechanistic analyses of representative node-edge fusion and update operators. Our findings provide empirical design heuristics for molecular MPNNs by turning model design from a search over monolithic architectures into a targeted assessment of where and how chemical information enters the message-passing pipeline.

cond-mat.mtrl-sci

Improved visual-information-driven model for crowd simulation and its modular application

Crowd movement simulation is crucial for pedestrian safety management and facility design. Data-driven models offer the potential to improve realism and predictive accuracy, but most are developed for a single scenario, limiting their flexibility. We propose a data-driven crowd simulation model that incorporates refined visual-information extraction and explicit exit cues, aiming to improve flexibility across multiple scenarios by more effectively capturing core navigational features. The model is tested on four fundamental modules (bottleneck, corridor, corner, and T-junction) and further evaluated in a composite scenario using a modular approach. Results show that our model performs well across these scenarios, aligning with pedestrian movement in real-world experiments, and outperforms the classical knowledge-driven model in these scenarios. The research outcomes can provide inspiration for the development of data-driven crowd simulation models and advance the application of data-driven approaches.

cs.CY

A collider as a quantum computer

Scattering processes in high-energy physics are inherently quantum mechanical, yet are typically analyzed at the level of final states, where entanglement appears as a property of the outcome rather than a consequence of the underlying dynamics. We reformulate scattering at the level of the process itself by representing helicity transition matrices as quantum circuits. Once the kinematic configuration and scattering channel are fixed, the problem reduces to a finite-dimensional quantum map, making a circuit description natural. Within this framework, an example of the process $e^+e^-\to μ^+μ^-$ is shown, which decomposes into unitary and non-unitary components, corresponding to coherent mixing and postselection effects. This representation reorganizes the amplitude into distinct operational elements, providing a perspective in which collider processes can be viewed as constrained quantum circuits and their entanglement structure can be understood in terms of the underlying circuit dynamics, opening the door to analyzing their properties using the language of quantum information.

hep-ph

League of LLMs: A Benchmark-Free Paradigm for Mutual Evaluation of Large Language Models

Although large language models (LLMs) have shown exceptional capabilities across a wide range of tasks, reliable evaluation remains a critical challenge due to data contamination, opaque operation, and subjective preferences. To address these issues, we propose League of LLMs (LOL), a novel benchmark-free evaluation paradigm that organizes multiple LLMs into a self-governed league for multi-round mutual evaluation. LOL integrates four core criteria (dynamic, transparent, objective, and professional) to mitigate key limitations of existing paradigms. Experiments on eight mainstream LLMs in mathematics and programming demonstrate that LOL can effectively distinguish LLM capabilities while maintaining high internal ranking stability (Top-$k$ consistency $= 70.7\%$). Beyond ranking, LOL reveals empirical findings that are difficult for traditional paradigms to capture. For instance, ``memorization-based answering'' behaviors are observed in some models, and higher in-family scores are found in the OpenAI model family ($Δ= 9$, $p < 0.05$). Finally, we make our framework and code publicly available as a valuable complement to the current LLM evaluation ecosystem.

cs.AI