arXiv ScienceSearch

arXiv subjects

Yunfeng Zhang

Publications and source records attributed to Yunfeng Zhang.

At least 19 recordsLinked to original sources

Action with Visual Primitives

Vision-Language-Action (VLA) models have emerged as a promising paradigm for generalist robotic manipulation. A common design in current architectures maps language instructions and visual observations to actions in a single forward pass. While conceptually simple, this formulation entangles instruction comprehension, spatial scene understanding, and motor control within a single learning objective. As a result, the action expert must implicitly relearn cognitive and perceptual capabilities already present in the pretrained VLM, which can limit both learning efficiency and generalization. We introduce AVP (Action with Visual Primitives), an end-to-end architecture that implements this visual-primitive-centric interface: the VLM infers the next-stage target and emits visual-primitive tokens that condition a flow-matching action expert, with supervision derived from end-effector kinematics. Real-robot experiments on general pick-and-place tasks show that AVP improves the success rate by 37.04% over pi_0.5 and outperforms other recent methods, with consistent gains in data efficiency, spatial-compositional generalization, and object-level transfer.

cs.RO

Pointwise character bounds for $\mathrm{SU}(3)$

We present a basic pointwise bound for the irreducible characters of $\mathrm{SU}(3)$ and, as an application, derive new $L^p$ bounds for these characters. Our approach is based on the descent of characters to singular sets and the cancellation in this formula.

math.RT

BAID: A Benchmark for Bias Assessment of AI Detectors

AI-generated text detectors have recently gained adoption in educational and professional contexts. Prior research has uncovered isolated cases of bias, particularly against English Language Learners (ELLs) however, there is a lack of systematic evaluation of such systems across broader sociolinguistic factors. In this work, we propose BAID, a comprehensive evaluation framework for AI detectors across various types of biases. As a part of the framework, we introduce over 200k samples spanning 7 major categories: demographics, age, educational grade level, dialect, formality, political leaning, and topic. We also generated synthetic versions of each sample with carefully crafted prompts to preserve the original content while reflecting subgroup-specific writing styles. Using this, we evaluate four open-source state-of-the-art AI text detectors and find consistent disparities in detection performance, particularly low recall rates for texts from underrepresented groups. Our contributions provide a scalable, transparent approach for auditing AI detectors and emphasize the need for bias-aware evaluation before these tools are deployed for public use.

cs.AI

Growth of Fourier--Lebesgue norms for mKdV

We demonstrate inflation of Fourier--Lebesgue norms for solutions to the focusing modified Korteweg--de Vries equation posed on the real line. For $p\neq 2$ and all $s\in \mathbb{R}$, we construct a sequence of solutions $u_n$ whose initial data $u_n(0)$ converges to zero in the Fourier--Lebesgue spaces $\mathcal F L^p_s(\mathbb{R})$, but whose evolutions at later times $t_n$ diverge to infinity.

math.AP

Restriction of eigenfunctions on products of spheres to submanifolds of maximal flats

Let $M$ be a product of rank-one symmetric spaces of compact type, each of dimension at least $3$. We establish sharp $L^p$ bounds for the restriction of Laplace--Beltrami eigenfunctions on $M$ to arbitrary submanifolds contained in a maximal flat, for all $p \ge 2$. The proof combines precise asymptotics of Jacobi polynomials and positivity of Fourier coefficients of spherical functions.

math.AP

Sharp bilinear eigenfunction estimate, $L^\infty_{x_2}L^p_{t,x_1}$-type Strichartz estimate, and energy-critical NLS

We establish sharp bilinear eigenfunction estimates for the Laplace-Beltrami operator on the standard three-sphere $\mathbb{S}^3$, eliminating the logarithmic loss that has persisted in the literature since the pioneering work of Burq, G\'erard, and Tzvetkov over twenty years ago. This completes the theory of multilinear eigenfunction estimates on the standard spheres. Our approach relies on viewing $\mathbb{S}^3$ as the compact Lie group $\mathrm{SU}(2)$ and exploiting its representation theory. Motivated by applications to the energy-critical nonlinear Schr\"odinger equation (NLS) on $\mathbb{R} \times \mathbb{S}^3$, we also prove a refined anisotropic Strichartz estimate on the cylindrical space $\mathbb{R}_{x_1} \times \mathbb{T}_{x_2}$ of $L^\infty_{x_2}L^4_{t,x_1}$-type, adapted to certain spectrally localized functions. The argument relies on multiple sharp measure estimates and a robust kernel decomposition method. Combining these two key ingredients, we derive a refined bilinear Strichartz estimate on $\mathbb{R} \times \mathbb{S}^3$, which in turn yields small-data global well-posedness for the above mentioned NLS in the energy space.

math.AP

Local well-posedness for nonlinear Schr\"odinger equations on compact product manifolds

We prove new local well-posedness results for nonlinear Schr\"odinger equations posed on a general product of spheres and tori, by the standard approach of multi-linear Strichartz estimates. To prove these estimates, we establish and utilize multi-linear bounds for the joint spectral projector associated to the Laplace--Beltrami operators on the individual sphere factors of the product manifold. To treat the particular case of the cubic NLS on a product of two spheres at critical regularity, we prove a sharp $L^\infty_xL^p_t$ estimate of the solution to the linear Schr\"odinger equation on the two-torus.

math.AP

The GeV $\gamma$-ray emission from the composite SNR CTB 87

We report the GeV $\gamma$-ray emission around the composite supernova remnant (SNR) CTB 87 with more than 16 yrs PASS 8 data recorded by the Fermi Large Area Telescope. Two separate point sources with the different GeV spectra are identified in this region: one has a soft $\gamma$-ray spectrum, likely due to interactions between the SNR shock and molecular clouds (MCs); and another source with a hard GeV $\gamma$-ray spectrum aligns with the TeV spectrum of VER J2016+371, suggesting it as the GeV counterpart. Considering the observations of CTB 87 in the radio and X-ray bands, VER J2016+371 is proposed to originate from the pulsar wind nebula (PWN) associated with PSR J2016+3711. A leptonic model with a broken power-law electron distribution could explain the multi-wavelength data of VER J2016+371, with fitted parameters matching typical $\gamma$-ray PWNe. Deeper searching for the SNR shock of CTB 87 in other bands and the future TeV observations by LHAASO and CTA are crucial to reveal the nature of CTB 87.

astro-ph.HE

Global well-posedness and equicontinuity for mKdV in modulation spaces

We establish global well-posedness for both the defocusing and focusing complex-valued modified Korteweg--de Vries equations on the real line in modulation spaces $M_p^{s,2}(\mathbb{R})$, for all $1\leq p<\infty$ and $0\leq s<3/2-1/p$. We will also show that such solutions admit global-in-time bounds in these spaces and that equicontinuous sets of initial data lead to equicontinuous ensembles of orbits. Indeed, such information forms a crucial part of our well-posedness argument.

math.AP

Bounds of restriction of characters to submanifolds

A fruitful approach to studying the concentration of Laplace--Beltrami eigenfunctions on a compact manifold, as the eigenvalue tends to infinity, is to bound their restriction to submanifolds. In this paper, we adopt this approach in the setting of compact Lie groups and provide sharp restriction bounds for general Laplace--Beltrami eigenfunctions, as well as for important special cases such as sums of matrix coefficients and, in particular, characters of irreducible representations. We prove sharp asymptotic $L^p$ bounds for the restriction of general Laplace--Beltrami eigenfunctions to maximal flats and all of their submanifolds, for all $p \geq 2$. Furthermore, we establish sharp asymptotic $L^p$ bounds for the restriction of characters to maximal tori and all of their submanifolds for all $p>0$, and to torus-generated conjugation-invariant submanifolds for all $p \geq 2$. We also obtain sharp $L^p$ bounds for the restriction of general sums of matrix coefficients to maximal flats and all of their submanifolds, for all $p \geq 2$.

math.RT

Simulation-to-reality UAV Fault Diagnosis in windy environments

Monitoring propeller failures is vital to maintain the safe and reliable operation of quadrotor UAVs. The simulation-to-reality UAV fault diagnosis technique offer a secure and economical approach to identify faults in propellers. However, classifiers trained with simulated data perform poorly in real flights due to the wind disturbance in outdoor scenarios. In this work, we propose an uncertainty-based fault classifier (UFC) to address the challenge of sim-to-real UAV fault diagnosis in windy scenarios. It uses the ensemble of difference-based deep convolutional neural networks (EDDCNN) to reduce model variance and bias. Moreover, it employs an uncertainty-based decision framework to filter out uncertain predictions. Experimental results demonstrate that the UFC can achieve 100% fault-diagnosis accuracy with a data usage rate of 33.6% in the windy outdoor scenario.

cs.RO

DDCNN: A Promising Tool for Simulation-To-Reality UAV Fault Diagnosis

Identifying the fault in propellers is important to keep quadrotors operating safely and efficiently. The simulation-to-reality (sim-to-real) UAV fault diagnosis methods provide a cost-effective and safe approach to detecting propeller faults. However, due to the gap between simulation and reality, classifiers trained with simulated data usually underperform in real flights. In this work, a novel difference-based deep convolutional neural network (DDCNN) model is presented to address the above issue. It uses the difference features extracted by deep convolutional neural networks to reduce the sim-to-real gap. Moreover, a new domain adaptation (DA) method is presented to further bring the distribution of the real-flight data closer to that of the simulation data. The experimental results demonstrate that the DDCNN+DA model can increase the accuracy from 52.9% to 99.1% in real-world UAV fault detection.

cs.RO

Simulation-to-reality UAV Fault Diagnosis with Deep Learning

Accurate diagnosis of propeller faults is crucial for ensuring the safe and efficient operation of quadrotors. Training a fault classifier using simulated data and deploying it on a real quadrotor is a cost-effective and safe approach. However, the simulation-to-reality gap often leads to poor performance of the classifier when applied in real flight. In this work, we propose a deep learning model that addresses this issue by utilizing newly identified features (NIF) as input and utilizing domain adaptation techniques to reduce the simulation-to-reality gap. In addition, we introduce an adjusted simulation model that generates training data that more accurately reflects the behavior of real quadrotors. The experimental results demonstrate that our proposed approach achieves an accuracy of 96\% in detecting propeller faults. To the best of our knowledge, this is the first reliable and efficient method for simulation-to-reality fault diagnosis of quadrotor propellers.

cs.RO

Strichartz estimates for the Schr\"odinger equation on products of odd-dimensional spheres

We prove Strichartz estimates for the Schr\"odinger equation which are scale-invariant up to an $\varepsilon$-loss on products of odd-dimensional spheres. Namely, for any product of odd-dimensional spheres $M=\mathbb{S}^{d_1}\times\cdots\times\mathbb{S}^{d_r}$ (so that $M$ is of dimension $d=d_1+\cdots+d_r$ and rank $r$) equipped with rational metrics, the following Strichartz estimate \begin{equation*} \|e^{it\Delta}f\|_{L^p(I\times M)}\leq C_\varepsilon\|f\|_{H^{\frac{d}{2}-\frac{d+2}{p}+\varepsilon}(M)} \end{equation*} holds for any $p\geq 2+\frac{8(s-1)}{sr}$, where $$s=\max\left\{\frac{2d_i}{d_i-1}, i=1,\ldots,r\right\}.$$

math.AP

Algebraic and analytic properties of invariant differential operators on a homogeneous space of complexity $1$

Denote by $SL_3(\mathbb R)$ the special linear group of degree 3 over the real numbers, $A$ the subgroup consisting of the diagonal matrices with positive entries. In this paper, we study the algebraic and analytic properties of the invariant differential operators on the homogeneous space $SL_3(\mathbb R)/A$. Firstly, we specify the noncommutative algebra of invariant differential operators in terms of generators and their relations. Secondly, we describe the center of this algebra and prove that all of its symmetric elements are essentially self-adjoint. Thirdly, for the first time on homogeneous spaces, we identify several essentially self-adjoint invariant differential operators which do not lie in the center of the algebra of invariant differential operators.

math.RT

Exponential canonical correlation analysis with orthogonal variation

Canonical correlation analysis (CCA) is a standard tool for studying associations between two data sources; however, it is not designed for data with count or proportion measurement types. In addition, while CCA uncovers common signals, it does not elucidate which signals are unique to each data source. To address these challenges, we propose a new framework for CCA based on exponential families with explicit modeling of both common and source-specific signals. Unlike previous methods based on exponential families, the common signals from our model coincide with canonical variables in Gaussian CCA, and the unique signals are exactly orthogonal. These modeling differences lead to a non-trivial estimation via optimization with orthogonality constraints, for which we develop an iterative algorithm based on a splitting method. Simulations show on par or superior performance of the proposed method compared to the available alternatives. We apply the method to analyze associations between gene expressions and lipids concentrations in nutrigenomic study, and to analyze associations between two distinct cell-type deconvolution methods in prostate cancer tumor heterogeneity study.

stat.CO

Careful Seeding for k-Medois Clustering with Incremental k-Means++ Initialization

K-medoids clustering is a popular variant of k-means clustering and widely used in pattern recognition and machine learning. A main drawback of k-medoids clustering is that an improper initialization can cause it to get trapped in local optima. An improved k-medoids clustering algorithm, called INCKM algorithm, which is the first to apply incremental initialization to k-medoids clustering, was recently proposed to overcome this drawback. The INCKM algorithm requires the construction of a subset of candidate medoids determined by one hyperparameter for initialization, and meanwhile, it always fails when dealing with imbalanced datasets with an incorrect hyperparameter selection. In this paper, we propose a novel k-medoids clustering algorithm, called incremental k-means++ (INCKPP) algorithm, which initializes with a novel incremental manner, attempting to optimally add one new cluster center at each stage through a nonparametric and stochastic k-means++ initialization. The INCKPP algorithm overcomes the difficulty of hyperparameter selection in the INCKM algorithm, improves the clustering performance, and can deal with imbalanced datasets well. However, the INCKPP algorithm is not computationally efficient enough. To deal with this, we further propose an improved INCKPP algorithm, called INCKPPsample algorithm, which improves the clustering efficiency while maintaining the clustering performance of the INCKPP algorithm. Extensive results from experiments on both synthetic and real-world datasets, including imbalanced datasets, illustrate that the proposed algorithms outperforms than the other compared algorithms.

cs.LG

Connecting Algorithmic Research and Usage Contexts: A Perspective of Contextualized Evaluation for Explainable AI

Recent years have seen a surge of interest in the field of explainable AI (XAI), with a plethora of algorithms proposed in the literature. However, a lack of consensus on how to evaluate XAI hinders the advancement of the field. We highlight that XAI is not a monolithic set of technologies -- researchers and practitioners have begun to leverage XAI algorithms to build XAI systems that serve different usage contexts, such as model debugging and decision-support. Algorithmic research of XAI, however, often does not account for these diverse downstream usage contexts, resulting in limited effectiveness or even unintended consequences for actual users, as well as difficulties for practitioners to make technical choices. We argue that one way to close the gap is to develop evaluation methods that account for different user requirements in these usage contexts. Towards this goal, we introduce a perspective of contextualized XAI evaluation by considering the relative importance of XAI evaluation criteria for prototypical usage contexts of XAI. To explore the context dependency of XAI evaluation criteria, we conduct two survey studies, one with XAI topical experts and another with crowd workers. Our results urge for responsible AI research with usage-informed evaluation practices, and provide a nuanced understanding of user requirements for XAI in different usage contexts.

cs.AI