arXiv ScienceSearch

arXiv subjects

Cheng Meng

Publications and source records attributed to Cheng Meng.

At least 19 recordsLinked to original sources

Beyond Straightness: Non-Crossing Flow Matching via Quantile AlignTree Coupling

The performance of Flow Matching largely depends on the quality of the coupling between the source and target distributions. However, independent coupling often leads to path crossings and local velocity ambiguity, while OT-based couplings typically incur high construction costs. To address this challenge, we propose Quantile AlignTree Flow Matching (QAT-FM), an efficient structured coupling strategy that constructs a hierarchical coupling between a Gaussian prior and the target data distribution via a quantile-aligned tree structure. QAT-FM constructs the coupling in $\mathcal{O}(Nd\log N)$ time and supports per-pair source sampling with $\mathcal{O}(d)$ complexity, enabling scalable training for large-scale high-dimensional generative tasks. Theoretically, we prove that the QAT coupling satisfies marginal consistency, induces non-crossing linear interpolation paths, and consistently improves path separation at intermediate times compared with independent coupling, thereby alleviating local velocity ambiguity. QAT-FM further extends naturally to conditional generation, enabling structured conditional coupling while preserving global Gaussian alignment. Experiments across diverse benchmark datasets demonstrate that QAT-FM achieves competitive generative performance while substantially reducing coupling construction cost.

cs.LG

Hilbert-Kunz multiplicity of quadrics decreases

In this paper we prove that the Green ring of $\mathbb{Z}/p^e\mathbb{Z}$ is the $e$-fold tensor product of the Green ring of $\mathbb{Z}/p\mathbb{Z}$, and this isomorphism is given by $p$-adic expansions of integers. As an application of this isomorphism, we compute the Hilbert-Kunz function and Hilbert-Kunz multiplicity of Fermat quadrics. Then we use Gelfand transform of the Green ring to give an analytical expression of this Hilbert-Kunz multiplicity, and prove that it decreases with its characteristic, thus giving a positive answer to a conjecture of Yoshida.

math.AC

From Context to Rules: Toward Unified Detection Rule Generation

Existing methods for detection rule generation are tightly coupled to specific input-output combinations, requiring dedicated pipelines for each. We formalize this problem as a unified mapping f:C*L->R and characterize optimal rules through semantic distance. We propose UniRule, an agentic RAG framework built on dual semantic projection spaces: detection intent and detection logic. This design enables retrieval and generation across arbitrary contexts and target languages within a single system. Experiments across 12 scenarios (3 languages, 4 context types, 12,000 pairwise comparisons) show that UniRule significantly outperforms pure LLM generation with a Bradley-Terry coefficient of 0.52, validating semantic projection as an effective abstraction for unified rule generation. Together, the formalization, method, and evaluation provide an initial framework for studying detection rule generation as a unified task.

cs.CR

Asymptotic behavior of modular representations over abelian $p$-groups

In this paper, we prove some results on the asymptotic behavior arising in modular representation theory over abelian $p$-groups. First, we embed the representation ring of a cyclic $p$-group into a real algebra of functions. Second, we calculate the asymptotic order of the dimension of the core of $n$-th tensor power of a direct sum of syzygies and cosyzygies of the trivial module, which is of the form $C\gamma^nn^\alpha$. This result leads to a negative answer to a question by Benson and Symonds, that is, the dimension of the core of $M^{\otimes n}$ for certain $\Omega$-algebraic module $M$ is not eventually recursive. Third, we give a systematic way of computing the core series of $\Omega$-algebraic modules. Finally, we show the existence of a transcendental core series, which comes from iterated syzygy modules of the trivial representation.

math.RT

HiMAP: Hilbert Mass-Addressed Parameterization for Multivariate Barycenters and Fre\'chet Regression

Learning from multivariate distribution-valued data requires both averaging distributions and predicting them from covariates. Positive barycenters and Fr\'echet regression impose different closure requirements. Regression weights can be negative, yet every fitted value must remain a probability law. We introduce the Hilbert mass-addressed parameterization (HiMAP), which uses balanced recursive median partitions to assign the same probability mass to each binary address across distributions. The common mass address yields two complementary representations. Q-HiMAP averages physical paths and defines a closed-form barycenter for nonnegative weights. Tree-logit HiMAP (TL-HiMAP) maps root geometry and relative split positions to Hilbert coordinates, where every finite affine combination decodes uniquely to a probability law. We establish exact mass coding, an encoder--decoder inverse, completeness, Wasserstein continuity, dense model coverage, and deterministic error decompositions for fixed-size particle outputs. The TL coordinates also give closed-form global and local Fr\'echet regression and allow each response to be encoded once for repeated prediction. Experiments with certified barycenters, simulated distribution regression, and two NHANES survey cycles show accurate barycenter recovery, algebraically exact TL-coordinate composition with finite particle reconstruction error, competitive predictive loss, and efficient repeated prediction.

stat.ME

Active Intelligence in Video Avatars via Closed-loop World Modeling

Current video avatar generation methods excel at identity preservation and motion alignment but lack genuine agency, they cannot autonomously pursue long-term goals through adaptive environmental interaction. We address this by introducing L-IVA (Long-horizon Interactive Visual Avatar), a task and benchmark for evaluating goal-directed planning in stochastic generative environments, and ORCA (Online Reasoning and Cognitive Architecture), the first framework enabling active intelligence in video avatars. ORCA embodies Internal World Model (IWM) capabilities through two key innovations: (1) a closed-loop OTAR cycle (Observe-Think-Act-Reflect) that maintains robust state tracking under generative uncertainty by continuously verifying predicted outcomes against actual generations, and (2) a hierarchical dual-system architecture where System 2 performs strategic reasoning with state prediction while System 1 translates abstract plans into precise, model-specific action captions. By formulating avatar control as a POMDP and implementing continuous belief updating with outcome verification, ORCA enables autonomous multi-step task completion in open-domain scenarios. Extensive experiments demonstrate that ORCA significantly outperforms open-loop and non-reflective baselines in task success rate and behavioral coherence, validating our IWM-inspired design for advancing video avatar intelligence from passive animation to active, goal-oriented behavior.

cs.CV

Core-elements Subsampling for Alternating Least Squares

In this paper, we propose a novel element-wise subset selection method for the alternating least squares (ALS) algorithm, focusing on low-rank matrix factorization involving matrices with missing values, as commonly encountered in recommender systems. While ALS is widely used for providing personalized recommendations based on user-item interaction data, its high computational cost, stemming from repeated regression operations, poses significant challenges for large-scale datasets. To enhance the efficiency of ALS, we propose a core-elements subsampling method that selects a representative subset of data and leverages sparse matrix operations to approximate ALS estimations efficiently. We establish theoretical guarantees for the approximation and convergence of the proposed approach, showing that it achieves similar accuracy with significantly reduced computational time compared to full-data ALS. Extensive simulations and real-world applications demonstrate the effectiveness of our method in various scenarios, emphasizing its potential in large-scale recommendation systems.

stat.ME

An adaptive subsampling method for large-sample feature screening

We consider the sure independence screening (SIS) method, a standard feature screening approach that aims to eliminate non-informative features in ultrahigh-dimensional datasets. Although effective, SIS incurs a computational cost of order $O(np)$ for a predictor matrix of size $n\times p$, which can be prohibitively expensive when both n and p are considerable. Motivated by the multi-armed bandit (MAB) problem, we propose a more computationally efficient feature screening algorithm that reduces the cost to $O(\sqrt{n}p)$. The core idea is to progressively increase the subsample size and eliminate variables with small empirical marginal Pearson correlations, thereby avoiding unnecessary computation on unpromising features. We develop a new interpretable statistical theoretical analysis that characterizes how the subsample size affects screening accuracy, thereby revealing the balance between computational efficiency and statistical reliability. Moreover, we show that the proposed method retains the sure screening property under mild regularity conditions. Extensive numerical experiments on synthetic and real-world datasets show that BanditSIS achieves screening and prediction performance comparable to SIS while substantially reducing computational time. Our method offers a scalable and adaptive alternative to SIS, particularly well-suited for large-sample, high-dimensional applications where computational efficiency is critical.

stat.ML

Finiteness and infiniteness of gradings of Noetherian rings

In this paper we show that for a torsion-free abelian group $G$, $\operatorname{rank}_\mathbb{Z}G<\infty$ if and only if there exists a Noetherian $G$-graded ring $R$ such that the set $\{R_g \neq 0\}$ generates the group $G$. For every $G$ of finite rank, we construct a $G$-graded ring $R$ such that $R_g \neq 0$ for all $g \in G$. We prove such rings give examples of PIDs which are not ED. We also use the relations between the graded division ring and the group cohomology to prove some vanishing and nonvanishing results for second group cohomology. Finally, we prove that the Hilbert series of a finitely generated $G$-graded $R$-module is well-defined when $R_0$ is Artinian, and this Hilbert series times some Laurent polynomial is equal to a Laurent polynomial.

math.AC

Explicit Stillman bounds for all degrees

In 2016 Ananyan and Hochster proved Stillman's conjecture by showing the existence of a uniform upper bound on the length of an $R_\eta$-sequence containing fixed $n$ forms of degree at most $d$ in polynomial rings over a field. This result yields many other uniform bounds including bounds on the projective dimension of the ideals generated by $n$ forms of degree at most $d$. Explicit values of these bounds for forms of degree $5$ and higher are not yet known. This article constructs such explicit bounds, one of which is an upper bound for the projective dimension of all homogeneous ideals, in polynomial rings over a field, generated by $n$ forms of degree at most $d$. In the settings of the Eisenbud-Goto conjecture, we derive an explicit bound of the Castelnuovo-Mumford regularity of a nondegenerate prime ideal $P$ in a polynomial ring $S$ in terms of the multiplicity of $S/P$.

math.AC

Limits of $F$-invariants and Riemann-Stieltjes integral

This paper proves several results on $F$-invariants of Fermat hypersurfaces, including the proof of an inequality on the Hilbert-Kunz multiplicity of Fermat quadric hypersurfaces conjectured by Watanabe and Yoshida, the asymptotic behavior of the Hilbert-Kunz multiplicity for Fermat cubic hypersurfaces, and a strict inequality of the $F$-signature of a Fermat hypersurface whose degree is equal to its dimension. To address the above problems, this paper introduces a numerical invariant for local rings of characteristic $p$ called multivariate $h$-function. It is a real function of several variables that recovers both the Hilbert-Kunz multiplicity and the $F$-signature of hypersurface rings. We prove the above results by developing integral formulas for the $h$-function of hypersurfaces defined by polynomials of the form $\phi(f_1,\ldots,f_s)$ in terms of the Riemann-Stieltjes integral, where $\phi$ is a polynomial and $f_i$'s are polynomials in independent sets of variables, and explore how taking derivatives and taking limit of the characteristic interact with the integrals.

math.AC

Gaussian Herding across Pens: An Optimal Transport Perspective on Global Gaussian Reduction for 3DGS

3D Gaussian Splatting (3DGS) has emerged as a powerful technique for radiance field rendering, but it typically requires millions of redundant Gaussian primitives, overwhelming memory and rendering budgets. Existing compaction approaches address this by pruning Gaussians based on heuristic importance scores, without global fidelity guarantee. To bridge this gap, we propose a novel optimal transport perspective that casts 3DGS compaction as global Gaussian mixture reduction. Specifically, we first minimize the composite transport divergence over a KD-tree partition to produce a compact geometric representation, and then decouple appearance from geometry by fine-tuning color and opacity attributes with far fewer Gaussian primitives. Experiments on benchmark datasets show that our method (i) yields negligible loss in rendering quality (PSNR, SSIM, LPIPS) compared to vanilla 3DGS with only 10% Gaussians; and (ii) consistently outperforms state-of-the-art 3DGS compaction techniques. Notably, our method is applicable to any stage of vanilla or accelerated 3DGS pipelines, providing an efficient and agnostic pathway to lightweight neural rendering. The code is publicly available at https://github.com/DrunkenPoet/GHAP

cs.CV

Instantiating Standards: Enabling Standard-Driven Text TTP Extraction with Evolvable Memory

Extracting MITRE ATT\&CK Tactics, Techniques, and Procedures (TTPs) from natural language threat reports is crucial yet challenging. Existing methods primarily focus on performance metrics using data-driven approaches, often neglecting mechanisms to ensure faithful adherence to the official standard. This deficiency compromises reliability and consistency of TTP assignments, creating intelligence silos and contradictory threat assessments across organizations. To address this, we introduce a novel framework that converts abstract standard definitions into actionable, contextualized knowledge. Our method utilizes Large Language Model (LLM) to generate, update, and apply this knowledge. This framework populates an evolvable memory with dual-layer situational knowledge instances derived from labeled examples and official definitions. The first layer identifies situational contexts (e.g., "Communication with C2 using encoded subdomains"), while the second layer captures distinctive features that differentiate similar techniques (e.g., distinguishing T1132 "Data Encoding" from T1071 "Application Layer Protocol" based on whether the focus is on encoding methods or protocol usage). This structured approach provides a transparent basis for explainable TTP assignments and enhanced human oversight, while also helping to standardize other TTP extraction systems. Experiments show our framework (using Qwen2.5-32B) boosts Technique F1 scores by 11\% over GPT-4o. Qualitative analysis confirms superior standardization, enhanced transparency, and improved explainability in real-world threat intelligence scenarios. To the best of our knowledge, this is the first work that uses the LLM to generate, update, and apply the a new knowledge for TTP extraction.

cs.CR

Restrictions on Hilbert coefficients give depths of graded domains

In this paper, we prove that if $P$ is a homogeneous prime ideal inside a standard graded polynomial ring $S$ with $\dim(S/P)=d$, and for $s \leq d$, adjoining $s$ general linear forms to the prime ideal changes the $(d-s)$-th Hilbert coefficient by 1, then $\text{depth}(S/P)=s-1$. This criterion also tells us about possible restrictions on the generic initial ideal of a prime ideal inside a polynomial ring.

math.AC

Asymptotic colengths for families of ideals: an analytic approach

This article focuses on the existence of asymptotic colengths for families of $\fm_{R}$-primary ideals in a Noetherian local ring $(R,\fm)$. In any characteristic, we generalize graded families to weakly graded families of ideals, and in prime characteristic, we explore various families such as weakly $p$-families and weakly inverse $p$-families. The main contribution of this paper is providing a unified analytic method to prove the existence of limits. Additionally, we establish Brunn-Minkowski type inequalities, positivity results, and volume = multiplicity formulas for these families of ideals.

math.AC

$h$-function, Hilbert-Kunz density function and Frobenius-Poincar\'e function

Given ideals $I,J$ of a noetherian local ring $(R, \mathfrak m)$ such that $I+J$ is $\mathfrak m$-primary and a finitely generated $R$-module $M$, we associate an invariant of $(M,R,I,J)$ called the $h$-function. Our results on $h$-functions allow extensions of the theories of Frobenius-Poincar\'e functions and Hilbert-Kunz density functions from the known graded case to the local case, answering a question of V.Trivedi. When $J$ is $\mathfrak m$-primary, we describe the support of the corresponding density function in terms of other invariants of $(R, I,J)$. We show that the support captures the $F$-threshold: $c^J(I)$, under mild assumptions, extending results of V. Trivedi and Watanabe. The $h$-function encodes Hilbert-Samuel, Hilbert-Kunz multiplicity and $F$-threshold of the ideal pair involved. Using this feature of $h$-functions, we provide an equivalent formulation of a conjecture of Huneke, Musta\c{t}\u{a}, Takagi, Watanabe; recover a result of Smirnov and Betancourt; give a new proof of a result answering Watanabe-Yoshida's question comparing Hilbert-Kunz and Hilbert-Samuel multiplicity and establish lower bounds on $F$-thresholds. We also point out that a conjecture of Smirnov-Betancourt as stated is false and suggest a correction which we relate to the conjecture of Huneke et al. We develop the theory of $h$-functions in a more general setting which yields a density function for $F$-signature. A key to many results on $h$-functions is a `convexity technique' that we introduce, which in particular proves differentiability of Hilbert-Kunz density functions almost everywhere on $(0,\infty)$, thus contributing to another question of Trivedi.

math.AC

Leverage classifier: Another look at support vector machine

Support vector machine (SVM) is a popular classifier known for accuracy, flexibility, and robustness. However, its intensive computation has hindered its application to large-scale datasets. In this paper, we propose a new optimal leverage classifier based on linear SVM under a nonseparable setting. Our classifier aims to select an informative subset of the training sample to reduce data size, enabling efficient computation while maintaining high accuracy. We take a novel view of SVM under the general subsampling framework and rigorously investigate the statistical properties. We propose a two-step subsampling procedure consisting of a pilot estimation of the optimal subsampling probabilities and a subsampling step to construct the classifier. We develop a new Bahadur representation of the SVM coefficients and derive unconditional asymptotic distribution and optimal subsampling probabilities without giving the full sample. Numerical results demonstrate that our classifiers outperform the existing methods in terms of estimation, computation, and prediction.

stat.ME

Sampling-Based Methods for Multi-Block Optimization Problems over Transport Polytopes

This paper focuses on multi-block optimization problems over transport polytopes, which underlie various applications including strongly correlated quantum physics and machine learning. Conventional block coordinate descent-type methods for the general multi-block problems store and operate on the matrix variables directly, resulting in formidable expenditure for large-scale settings. On the other hand, optimal transport problems, as a special case, have attracted extensive attention and numerical techniques that waive the use of the full matrices have recently emerged. However, it remains nontrivial to apply these techniques to the multi-block, possibly nonconvex problems with theoretical guarantees. In this work, we leverage the benefits of both sides and develop novel sampling-based block coordinate descent-type methods, which are equipped with either entropy regularization or Kullback-Leibler divergence. Each iteration of these methods solves subproblems restricted on the sampled degrees of freedom. Consequently, they involve only sparse matrices, which amounts to considerable complexity reductions. We explicitly characterize the sampling-induced errors and establish convergence and asymptotic properties for the methods equipped with the entropy regularization. Numerical experiments on typical strongly correlated electron systems corroborate their superior scalability over the methods utilizing full matrices. The advantage also enables the first visualization of approximate optimal transport maps between electron positions in three-dimensional contexts.

math.OC