arXiv ScienceSearch

arXiv subjects

Yuchen Ding

Publications and source records attributed to Yuchen Ding.

At least 19 recordsLinked to original sources

From Intent to Evidence: Policy-Steered Multi-Strategy Retrieval for Long-Video Agents

Existing long-video agents acquire evidence through one uniform behavior, ignoring whether the required evidence is concentrated, requires broad occurrence coverage, or must discriminate competing hypotheses---which can cause failure before substantive reasoning begins. Prescribing a fine-grained solution procedure for every question is not a satisfactory remedy, as it restricts autonomous exploration. We propose VESTA, a training-free long-video agent organized as a route-conditioned acquire--verify--consolidate loop. Before exploration, an intent router infers an evidence-acquisition policy---focused, recall, or contrastive retrieval over a shared visual--speech scene index---together with an evidence-accounting policy that configures the evidence view maintained during exploration. Policy-steered retrieval yields provisional references that multimodal evidence operations convert into observations, while the Reasoner remains free to verify them, re-query using intermediate findings, or inspect regions outside the retrieved set. A temporal evidence ledger consolidates observations into an adaptive, compressed view of temporal location, provenance, coverage, conflicts, verification outcomes, and hypothesis support, exposing missing and unresolved evidence to guide subsequent acquisition; finalization prioritizes verified observations. On Video-MME-v2, VESTA improves average accuracy by 2.7 points over VideoARM and gains across all six reported metrics. On LongVideoBench, EgoSchema, and LVBench under shared query-time models, it improves by 6.9 points on the LongVideoBench long subset and 1.5 on LVBench, and matches VideoARM on EgoSchema.

cs.CV

The GECKOS survey: Assembly history of the lenticular galaxy NGC 3957

We analyse the assembly history of the edge-on lenticular galaxy NGC 3957 using deep integral-field spectroscopic MUSE data from the GECKOS survey. By applying a dust-corrected Multi-Gaussian Expansion and a population-orbit superposition model, we disentangle the galaxy's stellar kinematics, age, and metallicity. We dynamically decompose the galaxy and identify three distinct components: a dynamically-cold main disc, a compact Nuclear Stellar Disc (NSD), and a hot component. The NSD emerges as the youngest and most metal-rich component ($t = 6.9 \pm 0.4$ Gyr; $[Z/H] = 0.49 \pm 0.06$ dex), implying that the stellar bar is a long-lived structure that formed at least $\sim 7$ Gyr ago. The main stellar disc is dynamically cold ($\sigma_z \sim 20-30$ km/s), precluding any significant mergers over the last $\sim 8$ Gyr, and exhibits a strong positive age gradient (younger inside, older outside) beyond the bar radius. Synthesising these dynamical fossil records, NGC 3957 likely evolved as a `faded spiral' in a small-to-medium group environment. Its outer disc might passively fade due to mild gas starvation, while the bar fuelled prolonged central star formation. Comparison with S0s in the Fornax cluster reveals that this combination of internal secular evolution and mild starvation produces `outside-in' fading signatures that could mimic the environmental stripping typically seen in dense clusters.

astro-ph.GA

Aggregating Visual Information with Optimal Transport for VideoLM Token Compression

Video language models process videos as dense visual-token sequences with substantial representational redundancy. Compressing these sequences is therefore essential for reducing the visual-token burden on language-model decoding. The central challenge is to preserve visual information dispersed across frames under such compression. To this end, we introduce Aggregating Visual Information with Optimal Transport (AVIOT), which casts video token compression as transporting a dense empirical measure of frame observations onto a compact target measure. The resulting source-to-target coupling induces a distribution over source observations for each target support, directly specifying how the compressed video representation is constructed. We further adapt this construction along task and spatial axes. Question conditioning modulates the transport cost between source frames and target supports, while influencing how many supports are allocated to each temporal segment, thereby directing representation capacity toward question-relevant content. At multiple spatial granularities, AVIOT computes region-specific temporal transport plans and adaptively fuses the representations they yield, allowing different regions within the same compact representation to draw from different moments. Evaluations across varying compression ratios show that AVIOT matches or outperforms the uncompressed baseline on multiple video-understanding benchmarks while retaining strong performance at higher compression ratios.

cs.CV

An improved upper bound on the Ruzsa number

Let $R_m$ be the least positive integer $r$ such that there exists a set $A\subseteq \mathbb{Z}_{m}$ with $A+A=\mathbb{Z}_m$ for which the number of ordered solutions of $n=x+y$ with $x,y\in A$ is at most $r$ for every $n\in \mathbb{Z}_m$. In this note we prove that $R_m\leqslant 128$ for every positive integer $m$, improving the previous bound $R_m\leqslant 192$.

math.NT

An improved lower bound for odd integers not of the form $p+2^a+2^b$

Let $x$ be sufficiently large and \[ N(x)=\big|\bigl\{n\le x:n\ \text{is odd and }n\ne p+2^a+2^b \textrm{ with } p \text{ a prime and } a,b\in \mathbb{N}\bigr\}\big|. \] Motivated by Crocker's result \[ N(x)\gg \log\log x, \] Erd\H os repeatedly asked whether there is an absolute constant $c_0$ such that $N(x)>c_0x$. Pan \cite{Pan} proved in 2011 that \[ N(x)\gg x\exp\!\left( -C_0\frac{\log\log\log\log x}{\log\log\log x}\log x \right), \] where $C_0>0$ is an absolute constant. We improve on Pan's result by showing that, given any $\eta>0$, for all sufficiently large $x$, \[ N(x)\gg_\eta x\exp\left(-(4+\eta)\frac{\log\log\log x}{\log\log x}\log x\right). \]

math.NT

A Romanoff-type theorem for $P_2$+{$a^a$: a$\ge$ 1}

Let $\Omega(n)$ denote the number of prime factors of $n$, counted with multiplicity, and put $P_2$={$m$ $\ge$ 1:$\Omega(m)$ $\le$ 2}. We prove that the sumset $P_2$+{$a^a$: a$\ge$ 1} has positive lower density. The proof uses the Romanoff second moment method, in the spirit of Li and Pan's theorem on $P_2$+$2^{\mathcal P}$. The main new ingredient is the following average estimate for the singular factor \[ \frac{1}{K(K-1)} \sum_{\substack{1\le a,b\le K\\a\ne b}} \prod_{p\mid a^a-b^b}\left(1+\frac{\kappa}{p}\right) \le C_\kappa \] for some constant $C_\kappa>0$, which is valid for all $K \ge 2$ and any fixed $\kappa>0$. This estimate controls the average arithmetic correlation among the shifts $a^a$ and allows the Romanoff argument to be carried out.

math.NT

Dissecting the 3D chemo-dynamical structures of NGC 1381: a galaxy hosting an ancient slow bar with an accreted bulge and thick disc

We applied the barred population-orbit superposition method developed in \citet{Jin2025a,Jin2025b} to construct 3D chemo-dynamical models for the barred S0 galaxy NGC~1381 in the Fornax cluster. Based on the stellar orbits in the models, we decomposed NGC~1381 into six components: (1) a dynamically warm nuclear disc with $f_{\rm nucl}\sim5\%$; (2) a rigidly rotating, BP/X-shaped bar with $f_{\rm bar}\sim30\%$; (3) a dynamically hot, spheroidal bulge with $f_{\rm bulge}\sim17\%$; (4) a dynamically cold thin disc with $f_{\rm thin}\sim28\%$; (5) a vertically extended thick disc with $f_{\rm thick}\sim16\%$; and (6) a dynamically hot, spatially diffuse stellar halo with $f_{\rm halo}\sim5\%$. The nuclear disc, bar, and thin disc are metal-rich ($[Z/\rm H]\gtrsim0$), $\alpha$-poor ($\rm[Mg/Fe]\lesssim0.2$), and old ($\sim13\rm\,Gyr$), corresponding to in situ formation in the early Universe. The bulge, thick disc, and stellar halo are metal-poor ($[Z/\rm H]\lesssim0$), $\alpha$-rich ($\rm[Mg/Fe]\gtrsim0.2$), and younger than or comparable in age to the in situ components, suggesting their relations with ex situ formation contributed by minor mergers. The flat metallicity and [Mg/Fe] gradients in the thick disc and stellar halo indicate they are dominated by a similar population of ex situ stars. In contrast, the bulge exhibits a negative metallicity gradient ($\nabla[Z/\rm H]_{bulge}<0$) pointing to a more complex formation history: the bulge could be either predominantly ex situ or contain a non-negligible mixture of in situ and ex situ stars. Our modelling also reveals the presence of a slow bar ($\mathcal{R}=2.40_{-0.27}^{+0.54}$), with a bar pattern speed of $\rm\Omega_p=34_{-7}^{+4}\,km\,s^{-1}\,kpc^{-1}$, a bar length of $R_{\rm bar}=2.24_{-0.22}^{+0.43}\rm\,kpc$, and a corotation radius of $R_{\rm CR}=5.38_{-0.28}^{+1.59}\rm\,kpc$, which is consistent with its ancient formation time.

astro-ph.GA

Restricted partition functions and additive complements

Let $\mathbb{N}$ be the set of positive integers. For subsets $\mathcal{A},\mathcal{M}\subseteq \mathbb{N}$ and $n\in \mathbb{N}$, let $p(n,\mathcal{A},\mathcal{M})$ denote the number of representations of $n$ in the form $$ n=\sum_{a\in \mathcal{A}}m_a a, $$ where $m_a\in \mathcal{M}\cup \{0\}$ for all $a\in \mathcal{A}$, and only finitely many $m_a$ are nonzero. We prove that there exist two infinite sets $\mathcal{A}=\{a_n\}_{n=1}^{\infty}$ and $\mathcal{M}$ of positive integers such that $$ \lim_{n\to\infty}\frac{\log a_{n+1}-\log a_n}{\log n}=+\infty, $$ $p(n,\mathcal{A},\mathcal{M})>0$ for every $n\in\mathbb{N}$, and $p$ has polynomial growth. More generally, we prove a construction that associates restricted partition functions of polynomial growth with additive complements satisfying a simple counting condition. This answers a 2016 question of Dai and Chen in the affirmative.

math.NT

Dense finite Sidon sets on arithmetic progressions

Let $S\subset \{1,2,\ldots,n\}$ be a Sidon set with $|S|=n^{1/2}+O(n^{1/2-\delta})$ for some fixed $\delta>0$. This article provides the following expected asymptotic formula $$ \sum_{\substack{a\in S\\ a\equiv r\pmod{m}}} a^\ell =\frac{1}{m(\ell+1)}n^{\ell+1/2} +o\left(n^{\ell+1/2}\right), $$ where $m\geq 1$, $0\leq r<m$, and $\ell\geq 0$ are three integers. This removes the additional hypothesis in a previous residue class asymptotic formula by the author. The proof uses the Fourier uniformity of extremal Sidon sets due to Ortega and Prendiville.

math.NT

An AI Proof of 18-Variable Undecidability for Diophantine Equations over $\mathbb Z[i]$

This paper presents an AI proof that there is no algorithm deciding whether a polynomial equation over the Gaussian integers in $18$ unknowns has a solution. The proof improves the $20$-unknown theorem of Matiyasevich and Sun. It follows their rationality criterion and integer test, but saves two variables: the clearing-denominator variable is avoided by imposing two integer conditions, and the remaining nonzero condition is absorbed into the relation-combining lemma by the one-variable gadget $(2R+1)(3R+1)$.

math.NT

Adjacent comparison bounds and extremal sets for Ruzsa numbers

Let $m$ be a positive integer and $\mathbb{Z}_m$ the residue class ring modulo $m$. The Ruzsa number $R_m$ is defined to be the least integer $r$ such that there is a subset $\mathcal{A}$ of $\mathbb{Z}_m$ satisfying $ 1\le \sigma_{\mathcal{A}}(n)\le r $ for any $n\in \mathbb{Z}_m$, where $$ \sigma_{\mathcal{A}}(n) =\#\big\{(a,a')\in\mathcal A^2: a+a'\equiv n\pmod{m}\big\}. $$ Motivated by a 2024 conjecture of Ding and Zhao, we prove $ | R_{m+1}-R_m|\le 144. $ Let $\mathcal{A}$ be a subset of $\mathbb{Z}_m$ satisfying $1\le \sigma_{\mathcal{A}}(n)\le R_m$ for any $n\in \mathbb{Z}_m$. We also give nontrivial bounds for the size of $\mathcal{A}$. Additionally, we provide exact values of $R_m$ for all $m\le 100$, which substantially extends the table of values given by S\'andor and Yang in 2017. Finally, we pose several related problems and prove some partial results.

math.NT

On a conjecture on Romanoff type sumsets

In this note, we generalize a 1950 result of P. Erd\H os on upper bounds of $k$-th moment of Romanoff type representation functions. As an application, we give a conditional proof of a recent conjecture of Y.-G. Chen on Romanoff type sumsets under the assumption of the Hardy-Littlewood conjecture.

math.NT

Short intervals for the Romanoff-type sumset

Let $X$ be large and let $\mathcal{P}$ denote the set of primes. Fix positive real parameters $r_1,\dots,r_s$ and a parameter $\lambda\geqslant 1$ determined by a balancing relation, and let $\mathcal{A}_{\lambda}(X)\subset[1,2X]$ be the associated lacunary set generated by sums of powers of $2$ with polynomially growing exponents. Set $\mathcal{S}_{\lambda}:=\mathcal{P}+\mathcal{A}_{\lambda}(X)$. Fix $\varepsilon>0$, choose $\theta$ with $2/15+\varepsilon<\theta<0.99$, and set $h=X^{\theta}$. We prove that for all but $O_{\varepsilon}\left(X\exp\left(-c_{\varepsilon}(\log X)^{1/4}\right)\right)$ values of $x\in[X,2X]$, the short interval $(x,x+h]$ contains $\asymp_{\varepsilon} h$ integers of the form $p+a$, where $p$ is prime and $a\in\mathcal{A}_{\lambda}(X)$.

math.NT

Note on unique representation bases

Answering affirmatively a 2007 problem of Chen, the first author proved that there is a unique representation basis $A$ of $\mathbb{Z}$ and a constant $c>0$ such that $$ A(-x,x)\ge c\sqrt{x} $$ for infinitely many positive integers $x$, where $A(-x,x)=\big|A\cap[-x, x]\big|$. Let $c_{\mathscr{A}}$ be the least upper bound for such $c$. It was proved in the former article by the first author that $\sqrt{2}/2\le c_{\mathscr{A}}\le \sqrt{2}$. In this note, the prior result is improved to $c_{\mathscr{A}}\ge 1$.

math.NT

ERNIE 5.0 Technical Report

In this report, we introduce ERNIE 5.0, a natively autoregressive foundation model desinged for unified multimodal understanding and generation across text, image, video, and audio. All modalities are trained from scratch under a unified next-group-of-tokens prediction objective, based on an ultra-sparse mixture-of-experts (MoE) architecture with modality-agnostic expert routing. To address practical challenges in large-scale deployment under diverse resource constraints, ERNIE 5.0 adopts a novel elastic training paradigm. Within a single pre-training run, the model learns a family of sub-models with varying depths, expert capacities, and routing sparsity, enabling flexible trade-offs among performance, model size, and inference latency in memory- or time-constrained scenarios. Moreover, we systematically address the challenges of scaling reinforcement learning to unified foundation models, thereby guaranteeing efficient and stable post-training under ultra-sparse MoE architectures and diverse multimodal settings. Extensive experiments demonstrate that ERNIE 5.0 achieves strong and balanced performance across multiple modalities. To the best of our knowledge, among publicly disclosed models, ERNIE 5.0 represents the first production-scale realization of a trillion-parameter unified autoregressive model that supports both multimodal understanding and generation. To facilitate further research, we present detailed visualizations of modality-agnostic expert routing in the unified model, alongside comprehensive empirical analysis of elastic training, aiming to offer profound insights to the community.

cs.CL

Cross representations of additive complements of $r$-th powers

Let $\mathbb{N}$ be the set of natural numbers and $\mathcal{S}_r=\big\{1^r, 2^r, 3^r,\cdots\big\}$ the set of $r$-th powers, where $r\ge 2$ is a natural number. Let $\mathcal{W}_r$ be an additive complement of $\mathcal{S}_r$ and $$ f_r(n)=\#\big\{(w,m^r)\in \mathcal{W}_r\times \mathcal{S}_r: n=w+m^r\big\}. $$ Motivated by a 1993 conjecture of Cilleruelo, we show that $$ \sum_{n\le N}f_r(n)-N\gg_r N^{1-\frac{1}{r}}. $$ Previously, the bound was only proved for $r=2$. In the case $r=2$, the lower bound above can be made more explicit as $$ \sum_{n\le N}f_2(n)-N\gg N^{3/4-o(1)}, $$ which improves the previous bound $N^{1/2}$ due to Ding, Sun, Wang and Xia.

math.NT

Disc growth and vertical heating of lenticular galaxies in the Fornax cluster

We present a detailed analysis of the vertical and radial structure of mono-age stellar populations in three edge-on lenticular galaxies (FCC 153, FCC 170, and FCC 177) in the Fornax cluster, using deep MUSE observations. By measuring the half-mass radius (R$_{50}$) and half-mass height (z$_{50}$) across 1 Gyr-wide age bins, we trace the spatial evolution of stellar populations over cosmic time. All galaxies exhibit a remarkably constant disc thickness for all stars younger than ~6 Gyr, suggesting minimal secular heating and limited impact from environmental processes such as tidal shocking or harassment. Evidence of past mergers (8-10 Gyr ago) is found in the increase of z$_{50}$ for older populations. We find that accreted (metal-poor) stars have been deposited in quite thick configurations, but that the interactions only moderately thickened pre-existing stars in the galaxies, and only caused mild flaring in the outer regions of the discs. The radial structure of the discs varies across galaxies, but in all cases we find that the radial extent of mono-age populations remains constant or grows over the past 8 Gyr. This leads us to argue that within the radial range we consider, strangulation, rather than ram-pressure stripping, is the dominant quenching mechanism in those galaxies. Our results highlight the usefulness of analysing the structure of mono-age population to uncover the mechanisms driving galaxy evolution, and we anticipate broader insights from the GECKOS survey, studying 36 nearby edge-on disc galaxies.

astro-ph.GA

Note on shifted primes with large prime factors

We denote by $P^+(n)$ the largest prime factor of the integer $n$. In 1935, Erd\H os studied the quantity $T_c(x)$ defined by $$ T_c(x)=\big|\big\{p\le x: P^+(p-1)\ge p^c\big\}\big|, $$ and he proved $$ \limsup_{x\rightarrow \infty}\frac{T_c(x)}{\pi(x)}\rightarrow 0, \quad \text{as~}c\rightarrow 1. $$ Recently, Ding gave a quantitative form of Erd\H os' result, showing that $$ \limsup_{x\rightarrow \infty}\frac{T_c(x)}{\pi(x)}\le 8\big(c^{-1}-1\big). $$ holds for $8/9< c<1$. In this paper, we improve Ding's upper bound to $$ \limsup_{x\rightarrow \infty}\frac{T_c(x)}{\pi(x)}\le -\frac{7}{2}\log c $$ for $e^{-\frac{2}{7}}< c<1$.

math.NT