arXiv ScienceSearch

arXiv subjects

Hui Xiao

Publications and source records attributed to Hui Xiao.

At least 19 recordsLinked to original sources

Strong non-arithmeticity for Zariski dense subsemigroups

We prove a non-arithmeticity property for the complex eigenvalues of a Zariski dense subsemigroups of real reductive algebraic groups. This property will be used in [7] in order to establish a local limit theorem for the operator norm of products of random matrices.

math.GR

Conditioned random walks on linear groups I: construction of the target harmonic measure

Our objective is to explore random walks on the general linear group, constrained to a specific domain, with a primary focus on establishing the conditioned local limit theorem. This paper represents the first step toward achieving this goal, specifically entailing the construction of a novel entity -- the target harmonic measure. This measure, together with the harmonic function, serves as a pivotal component in establishing the conditioned local limit theorem. Using a reversal identity, we introduce a reversed sequence characterized as a dual random walk with a perturbation depending on future observations. The investigation of such walks, which rely on future information, lies at the heart of this paper. To carry out this study, we develop an approach grounded in the finite-size approximation of perturbations, enabling us to simplify the investigation to an array of Markov chains with increasing dimensions.

math.PR

Multi-Scale Distillation for RGB-D Anomaly Detection on the PD-REAL Dataset

We present PD-REAL, a novel large-scale dataset for unsupervised anomaly detection (AD) in the 3D domain. It is motivated by the fact that 2D-only representations in the AD task may fail to capture the geometric structures of anomalies due to uncertainty in lighting conditions or shooting angles. PD-REAL consists entirely of Play-Doh models for 15 object categories and focuses on the analysis of potential benefits from 3D information in a controlled environment. Specifically, objects are first created with six types of anomalies, such as \textit{dent}, \textit{crack}, or \textit{perforation}, and then photographed under different lighting conditions to mimic real-world inspection scenarios. To demonstrate the usefulness of 3D information, we use a commercially available RealSense camera to capture RGB and depth images. Compared to the existing 3D dataset for AD tasks, the data acquisition of PD-REAL is significantly cheaper, easily scalable, and easier to control variables. Furthermore, we introduce a multi-scale teacher--student framework with hierarchical distillation for multimodal anomaly detection. This architecture overcomes the inherent limitation of single-scale distillation approaches, which often struggle to reconcile global context with local features. Leveraging multi-level guidance from the teacher network, the student network can effectively capture richer features for anomaly detection. Extensive evaluations with our method and state-of-the-art AD algorithms on our dataset qualitatively and quantitatively demonstrate the higher detection accuracy of our method. Our dataset can be downloaded from https://github.com/Andy-cs008/PD-REAL

cs.CV

A Multi-View Consistency Framework with Semi-Supervised Domain Adaptation

Semi-Supervised Domain Adaptation (SSDA) leverages knowledge from a fully labeled source domain to classify data in a partially labeled target domain. Due to the limited number of labeled samples in the target domain, there can be intrinsic similarity of classes in the feature space, which may result in biased predictions, even when the model is trained on a balanced dataset. To overcome this limitation, we introduce a multi-view consistency framework, which includes two views for training strongly augmented data. One is a debiasing strategy for correcting class-wise prediction probabilities according to the prediction performance of the model. The other involves leveraging pseudo-negative labels derived from the model predictions. Furthermore, we introduce a cross-domain affinity learning aimed at aligning features of the same class across different domains, thereby enhancing overall performance. Experimental results demonstrate that our method outperforms the competing methods on two standard domain adaptation datasets, DomainNet and Office-Home. Combining unsupervised domain adaptation and semi-supervised learning offers indispensable contributions to the industrial sector by enhancing model adaptability, reducing annotation costs, and improving performance.

cs.CV

Exploiting Minority Pseudo-Labels for Semi-Supervised Fine-grained Road Scene Understanding

In fine-grained road scene understanding, semantic segmentation plays a crucial role in enabling vehicles to perceive and comprehend their surroundings. By assigning a specific class label to each pixel in an image, it allows for precise identification and localization of detailed road features, which is vital for high-quality scene understanding and downstream perception tasks. A key challenge in this domain lies in improving the recognition performance of minority classes while mitigating the dominance of majority classes, which is essential for achieving balanced and robust overall performance. However, traditional semi-supervised learning methods often train models overlooking the imbalance between classes. To address this issue, firstly, we propose a general training module that learns from all the pseudo-labels without a conventional filtering strategy. Secondly, we propose a professional training module to learn specifically from reliable minority-class pseudo-labels identified by a novel mismatch score metric. The two modules are crossly supervised by each other so that it reduces model coupling which is essential for semi-supervised learning. During contrastive learning, to avoid the dominance of the majority classes in the feature space, we propose a strategy to assign evenly distributed anchors for different classes in the feature space. Experimental results on multiple public benchmarks show that our method surpasses traditional approaches in recognizing tail classes.

cs.CV

Local limit theorems for conditioned random walks by the heat kernel approximation

We study the random walk $(S_n)_{n\geq 1}$ with independent and identically distributed real-valued increments having zero mean and an absolute moment of order $2 + δ$ for some $δ> 0$. For any starting point $x \in \mathbb{R}$, let $τ_x = \inf\{k \geq 1 : x + S_k < 0\}$ denote the first exit time of the random walk $x + S_n$ from the half-line $[0, \infty)$. In the previous work [25], we established a Gaussian heat kernel approximation for both the persistence probability $\mathbb{P}(τ_x > n)$ and the joint distribution $\mathbb{P}(x + S_n \leq \cdot, τ_x > n)$, uniformly over $x \in \mathbb{R}$ as $n \to \infty$. In this paper, we leverage these results to establish a novel conditioned local limit theorem for the walk $(x + S_n)_{n \geq 1}$. For $\mathbb{Z}$-valued random walks, we prove that the joint probability $\mathbb{P}(x + S_n = y, τ_x > n)$ is uniformly approximated by a distribution governed by the Gaussian heat kernel over all $x, y \in \mathbb{Z}$ as $n \to \infty$. Our new asymptotic unifies into a single comprehensive formula the classical local limit theorem by Caravenna [6], as well as various results relying on specific assumptions on $x$ and $y$. As a corollary, we obtain a new uniform-in-$x$ asymptotic formula for the local probability $\mathbb{P}(τ_x = n)$. We also extend our analysis to non-lattice random walks.

math.PR

Type R $λ$-Permutation Approach to Velleman's Open Problem

Previously, mathematicians Steven Krantz and Jeffery McNeal studied a type of positive numbers permutation called $λ$-permutation. This type of permutation, when applied to the index of terms of a series, is defined to be both convergence-preserving and "fixing" at least one divergent series, that is, rearranging the terms of any convergent series will result in a convergent series, while rearranging the terms of some divergent series will result in a convergent series. In general, if a divergent series can be fixed to converge in some way (it does not need to be by $λ$-permutation), it is called a "conditionally divergent series". In 2006, another mathematician Daniel Velleman raised an open problem related to $λ$-permutation: for a conditionally divergent series $\sum_{n=0}^{\infty}a_n,n\in \mathbb{N},a_n\in \mathbb{R}$, let $S=\{L \in \mathbb{R} \colon L = \sum_{n=0}^{\infty}{a_{σ\left(n\right)}}$ $\text{for some } λ\text{-permutation } σ\}$, can $S$ ever be something between $\emptyset$ and $\mathbb{R}$? This paper is devoted to partially answering this open problem by considering a subset of $λ$-permutation constraint by how we can permute, named type R $λ$-permutation. Then we answer the analogous question about a subset of S with respect to type R $λ$-permutation, named $Z_{R}=\{L \in \mathbb{R} \colon L = \sum_{n=0}^{\infty}{a_{σ\left(n\right)}}$ $\text{for some type R } λ\text{-permutation } σ\}$. We show that $Z_R$ is either $\emptyset$, a singleton or $\mathbb{R}$. We also provide sufficient conditions on the conditionally divergent series $\sum_{n=0}^{\infty}a_n$ for $Z_R$ to be a singleton or $\mathbb{R}$, by introducing a "substantial property" on the series.

math.CO

Spinal decomposition, martingale convergence and the Seneta-Heyde scaling for matrix branching random walks

We consider a matrix branching random walk on the semi-group of nonnegative matrices, where we are able to derive, under general assumptions, an analogue of Biggins' martingale convergence theorem for the additive martingale $W_n$, a spinal decomposition theorem, convergence of the derivative martingale $D_n$, and finally, the Seneta-Heyde scaling stating that in the boundary case $c \sqrt{n} W_n \to D_\infty$ a.s., where $D_\infty$ is the limit of the derivative martingale and $c$ is a positive constant. As an important tool that is of interest in its own right, we provide explicit duality results for the renewal measure of centered Markov random walks, relating the renewal measure of the process, killed when the random walk component becomes negative, to the renewal measure of the ascending ladder process.

math.PR

Conditioned local limit theorems for products of positive random matrices

Let $(g_{n})_{n\geq 1}$ be a sequence of independent and identically distributed positive random $d\times d$ matrices, where $d\geq 2$ is an integer. For any starting point $x \in \mathbb{R}_+^d$ with $|x| = 1$ and $y \in \mathbb R$, we define the exit time $τ_{x, y} = \inf \{ k \geq 1: y + \log |g_k \cdots g_1 x| < 0 \}$. In this paper, we investigate the conditioned local probability $\mathbb{P} (y + \log |g_n \cdots g_1 x| \in z + [0, Δ], τ_{x, y} > n)$ under various assumptions on $y$, $z$ and $Δ$. For the case where $z = O(\sqrt{n})$, we establish an exact asymptotic result as $n \to \infty$, uniformly in $y$ and $Δ$, which extends the classical Caravenna conditioned local limit theorem to the case of products of positive random matrices. Our proof does not rely on the reversibility techniques. Furthermore, for arbitrary $z \in \mathbb R_+$, we deduce a uniform upper bound with rate $n^{-3/2}$.

math.PR

Multi-modal expressive personality recognition in data non-ideal audiovisual based on multi-scale feature enhancement and modal augment

Automatic personality recognition is a research hotspot in the intersection of computer science and psychology, and in human-computer interaction, personalised has a wide range of applications services and other scenarios. In this paper, an end-to-end multimodal performance personality is established for both visual and auditory modal datarecognition network , and the through feature-level fusion , which effectively of the two modalities is carried out the cross-attention mechanismfuses the features of the two modal data; and a is proposed multiscale feature enhancement modalitiesmodule , which enhances for visual and auditory boththe expression of the information of effective the features and suppresses the interference of the redundant information. In addition, during the training process, this paper proposes a modal enhancement training strategy to simulate non-ideal such as modal loss and noise interferencedata situations , which enhances the adaptability ofand the model to non-ideal data scenarios improves the robustness of the model. Experimental results show that the method proposed in this paper is able to achieve an average Big Five personality accuracy of , which outperforms existing 0.916 on the personality analysis dataset ChaLearn First Impressionother methods based on audiovisual and audio-visual both modalities. The ablation experiments also validate our proposed , respectivelythe contribution of module and modality enhancement strategy to the model performance. Finally, we simulate in the inference phase multi-scale feature enhancement six non-ideal data scenarios to verify the modal enhancement strategy's improvement in model robustness.

cs.SD

Berry-Esseen bound and precise moderate deviations for products of random matrices

Let $(g_{n})_{n\geq 1}$ be a sequence of independent and identically distributed (i.i.d.) $d\times d$ real random matrices. For $n\geq 1$ set $G_n = g_n \ldots g_1$. Given any starting point $x=\mathbb R v\in\mathbb{P}^{d-1}$, consider the Markov chain $X_n^x = \mathbb R G_n v $ on the projective space $\mathbb P^{d-1}$ and the norm cocycle $σ(G_n, x)= \log \frac{|G_n v|}{|v|}$, for an arbitrary norm $|\cdot|$ on $\mathbb R^{d}$. Under suitable conditions we prove a Berry-Esseen type theorem and an Edgeworth expansion for the couple $(X_n^x, σ(G_n, x))$. These results are established using a brand new smoothing inequality on complex plane, the saddle point method and additional spectral gap properties of the transfer operator related to the Markov chain $X_n^x$. Cramér type moderate deviation expansions as well as a local limit theorem with moderate deviations are proved for the couple $(X_n^x, σ(G_n, x))$ with a target function $φ$ on the Markov chain $X_n^x$.

math.PR

Edgeworth expansion and large deviations for the coefficients of products of positive random matrices

Consider the matrix products $G_n: = g_n \ldots g_1$, where $(g_{n})_{n\geq 1}$ is a sequence of independent and identically distributed positive random $d\times d$ matrices. Under the optimal third moment condition, we first establish a Berry-Esseen theorem and an Edgeworth expansion for the $(i,j)$-th entry $G_n^{i,j}$ of the matrix $G_n$, where $1 \leq i, j \leq d$. Using the Edgeworth expansion for $G_n^{i,j}$ under the changed probability measure, we then prove precise upper and lower large deviation asymptotics for the entries $G_n^{i,j}$ subject to an exponential moment assumption. As applications, we deduce local limit theorems with large deviations for $G_n^{i,j}$ and upper and lower large deviations bounds for the spectral radius $ρ(G_n)$ of $G_n$. A byproduct of our approach is the local limit theorem for $G_n^{i,j}$ under the optimal second moment condition. In the proofs we develop a spectral gap theory for the norm cocycle and for the coefficients, which is of independent interest.

math.PR

Gaussian heat kernel asymptotics for conditioned random walks

Consider a random walk $S_n=\sum_{i=1}^n X_i$ with independent and identically distributed real-valued increments with zero mean, finite variance and moment of order $2 + δ$ for some $δ>0$. For any starting point $x\in \mathbb R$, let $τ_x = \inf \left\{ k\geq 1: x+S_{k} < 0 \right\}$ denote the first time when the random walk $x+S_n$ exits the half-line $[0,\infty)$. We investigate the uniform asymptotic behavior over $x\in \mathbb R$ of the persistence probability $\mathbb P (τ_x >n)$ and the joint distribution $\mathbb{P} \left( x + S_n \leq u, τ_x > n \right)$, for $u\geq 0$, as $n \to \infty$. New limit theorems for these probabilities are established based on the heat kernel approximations. Additionally, we evaluate the rate of convergence by proving Berry-Esseen type bounds.

math.PR

The extremal position of a branching random walk on the general linear group

Consider a branching random walk $(G_u)_{u\in \mathbb T}$ on the general linear group $\textrm{GL}(V)$ of a finite dimensional space $V$, where $\mathbb T$ is the associated genealogical tree with nodes $u$. For any starting point $v \in V \setminus\{0\}$ with $\|v\|=1$ and $x = \mathbb R v \in \mathbb P(V)$, let $M^x_n=\max_{|u| = n} \log \| G_u v \|$ denote the maximal position of the walk $\log \| G_u v \|$ in the generation $n$. We first show that under suitable conditions, $\lim_{n \to \infty} \frac{M_n^x }{n} = γ$ almost surely, where $γ\in \mathbb R$ is a constant. Then, in the case when $γ= 0$, under appropriate {\textit boundary conditions}, we refine the last statement by determining the rate of convergence at which $M_n^x$ converges to $-\infty$. We prove in particular that $\lim_{n \to \infty} \frac{M_n^x}{\log n} = -\frac{3}{2α}$ in probability, where $α>0$ is a constant determined by the boundary conditions. Analogous properties are established for the minimal position. As a consequence we derive the asymptotic speed of the maximal and minimal positions for the coefficients, the operator norm and the spectral radius of $G_u$.

math.PR

Limit theorems for first passage times of multivariate perpetuity sequences

We study the first passage time $τ_u = \inf \{ n \geq 1: |V_n| > u \}$ for the multivariate perpetuity sequence $V_n = Q_1 + M_1 Q_2 + \cdots + (M_1 \ldots M_{n-1}) Q_n$, where $(M_n, Q_n)$ is a sequence of independent and identically distributed random variables with $M_1$ a $d \times d$ ($d \geq 1$) random matrix with nonnegative entries, and $Q_1$ a nonnegative random vector in $\mathbb R^d$. Here $|\cdot|$ denotes the vector norm. The exact asymptotic for the probability $\mathbb P (τ_u < \infty)$ as $u \to \infty$ has been found by Kesten (Acta Math. 1973). In this paper we prove a conditioned weak law of large numbers for $τ_u$: conditioned on the event $\{ τ_u < \infty \}$, $\frac{τ_u}{\log u}$ converges in probability to a certain constant $ρ> 0$ as $u \to \infty$. A conditioned central limit theorem for $τ_u$ is also obtained. We further establish precise large deviation asymptotics for the lower probability $\mathbb P (τ_u \leq (β- l) \log u)$ as $u \to \infty$, where $β\in (0, ρ)$ and $l \geq 0$ is a vanishing perturbation satisfying $l \to 0$ as $u \to \infty$. Our results extend those of Buraczewski et al. (Ann. Probab. 2016) from the univariate case ($d=1$) to the multivariate case ($d>1$). As consequences, we deduce exact asymptotics for the pointwise probability $\mathbb P (τ_u = [(β- l) \log u] )$ and the local probability $\mathbb P (τ_u - (β- l) \log u \in (a, a + m ] )$, where $a<0$ and $m \in \mathbb Z_+$. We also establish analogous results for the first passage time $τ_u^y = \inf \{ n \geq 1: \langle y, V_n \rangle > u \}$, where $y$ is a nonnegative vector in $\mathbb R^d$ with $|y| = 1$.

math.PR

Conditioned random walks on linear groups II: local limit theorems

We investigate random walks on the general linear group constrained within a specific domain, with a focus on their asymptotic behavior. In a previous work [38], we constructed the associated harmonic measure, a key element in formulating the local limit theorem for conditioned random walks on groups. The primary aim of this paper is to prove this theorem. The main challenge arises from studying the conditioned reverse walk, whose increments, in the context of random walks on groups, depend on the entire future. To achieve our goal, we combine a Caravenna-type conditioned local limit theorem with the conditioned version of the central limit theorem for the reversed walk. The resulting local limit theorem is then applied to derive the local behavior of the exit time.

math.PR

Prior-based Objective Inference Mining Potential Uncertainty for Facial Expression Recognition

Annotation ambiguity caused by the inherent subjectivity of visual judgment has always been a major challenge for Facial Expression Recognition (FER) tasks, particularly for largescale datasets from in-the-wild scenarios. A potential solution is the evaluation of relatively objective emotional distributions to help mitigate the ambiguity of subjective annotations. To this end, this paper proposes a novel Prior-based Objective Inference (POI) network. This network employs prior knowledge to derive a more objective and varied emotional distribution and tackles the issue of subjective annotation ambiguity through dynamic knowledge transfer. POI comprises two key networks: Firstly, the Prior Inference Network (PIN) utilizes the prior knowledge of AUs and emotions to capture intricate motion details. To reduce over-reliance on priors and facilitate objective emotional inference, PIN aggregates inferential knowledge from various key facial subregions, encouraging mutual learning. Secondly, the Target Recognition Network (TRN) integrates subjective emotion annotations and objective inference soft labels provided by the PIN, fostering an understanding of inherent facial expression diversity, thus resolving annotation ambiguity. Moreover, we introduce an uncertainty estimation module to quantify and balance facial expression confidence. This module enables a flexible approach to dealing with the uncertainties of subjective annotations. Extensive experiments show that POI exhibits competitive performance on both synthetic noisy datasets and multiple real-world datasets. All codes and training logs will be publicly available at https://github.com/liuhw01/POI.

cs.CV