arXiv Science⌕ Search

arXiv subjects

Xin Yu

Publications and source records attributed to Xin Yu.

At least 253 records · Page 14Linked to original sources

Learning Strict Identity Mappings in Deep Residual Networks

A family of super deep networks, referred to as residual networks or ResNet, achieved record-beating performance in various visual tasks such as image recognition, object detection, and semantic segmentation. The ability to train very deep networks naturally pushed the researchers to use enormous resources to achieve the best performance. Consequently, in many applications super deep residual networks were employed for just a marginal improvement in performance. In this paper, we propose epsilon-ResNet that allows us to automatically discard redundant layers, which produces responses that are smaller than a threshold epsilon, with a marginal or no loss in performance. The epsilon-ResNet architecture can be achieved using a few additional rectified linear units in the original ResNet. Our method does not use any additional variables nor numerous trials like other hyper-parameter optimization techniques. The layer selection is achieved using a single training process and the evaluation is performed on CIFAR-10, CIFAR-100, SVHN, and ImageNet datasets. In some instances, we achieve about 80% reduction in the number of parameters.

cs.CV↗

Can generalised relative pose estimation solve sparse 3D registration?

Popular 3D scan registration projects, such as Stanford digital Michelangelo or KinectFusion, exploit the high-resolution sensor data for scan alignment. It is particularly challenging to solve the registration of sparse 3D scans in the absence of RGB components. In this case, we can not establish point correspondences since the same 3D point cannot be captured in two successive scans. In contrast to correspondence based methods, we take a different viewpoint and formulate the sparse 3D registration problem based on the constraints from the intersection of line segments from adjacent scans. We obtain the line segments by modeling every horizontal and vertical scan-line as piece-wise linear segments. We propose a new alternating projection algorithm for solving the scan alignment problem using line intersection constraints. We develop two new minimal solvers for scan alignment in the presence of plane correspondences: 1) 3 line intersections and 1 plane correspondence, and 2) 1 line intersection and 2 plane correspondences. We outperform other competing methods on Kinect and LiDAR datasets.

cs.CV↗

Identity-preserving Face Recovery from Stylized Portraits

Given an artistic portrait, recovering the latent photorealistic face that preserves the subject's identity is challenging because the facial details are often distorted or fully lost in artistic portraits. We develop an Identity-preserving Face Recovery from Portraits (IFRP) method that utilizes a Style Removal network (SRN) and a Discriminative Network (DN). Our SRN, composed of an autoencoder with residual block-embedded skip connections, is designed to transfer feature maps of stylized images to the feature maps of the corresponding photorealistic faces. Owing to the Spatial Transformer Network (STN), SRN automatically compensates for misalignments of stylized portraits to output aligned realistic face images. To ensure the identity preservation, we promote the recovered and ground truth faces to share similar visual features via a distance measure which compares features of recovered and ground truth faces extracted from a pre-trained FaceNet network. DN has multiple convolutional and fully-connected layers, and its role is to enforce recovered faces to be similar to authentic faces. Thus, we can recover high-quality photorealistic faces from unaligned portraits while preserving the identity of the face in an image. By conducting extensive evaluations on a large-scale synthesized dataset and a hand-drawn sketch dataset, we demonstrate that our method achieves superior face recovery and attains state-of-the-art results. In addition, our method can recover photorealistic faces from unseen stylized portraits, artistic paintings, and hand-drawn sketches.

cs.CV↗

Recovering Faces from Portraits with Auxiliary Facial Attributes

Recovering a photorealistic face from an artistic portrait is a challenging task since crucial facial details are often distorted or completely lost in artistic compositions. To handle this loss, we propose an Attribute-guided Face Recovery from Portraits (AFRP) that utilizes a Face Recovery Network (FRN) and a Discriminative Network (DN). FRN consists of an autoencoder with residual block-embedded skip-connections and incorporates facial attribute vectors into the feature maps of input portraits at the bottleneck of the autoencoder. DN has multiple convolutional and fully-connected layers, and its role is to enforce FRN to generate authentic face images with corresponding facial attributes dictated by the input attribute vectors. %Leveraging on the spatial transformer networks, FRN automatically compensates for misalignments of portraits. % and generates aligned face images. For the preservation of identities, we impose the recovered and ground-truth faces to share similar visual features. Specifically, DN determines whether the recovered image looks like a real face and checks if the facial attributes extracted from the recovered image are consistent with given attributes. %Our method can recover high-quality photorealistic faces from unaligned portraits while preserving the identity of the face images as well as it can reconstruct a photorealistic face image with a desired set of attributes. Our method can recover photorealistic identity-preserving faces with desired attributes from unseen stylized portraits, artistic paintings, and hand-drawn sketches. On large-scale synthesized and sketch datasets, we demonstrate that our face recovery method achieves state-of-the-art results.

cs.CV↗

Periodic parabola solitons for the nonautonomous KP equation

Kadomtsev-Petviashvili (KP) equation, who can describe different models in fluids and plasmas, has drawn investigation for its solitonic solutions with various methods. In this paper, we focus on the periodic parabola solitons for the (2+1) dimensional nonautonomous KP equations where the necessary constraints of the parameters are figured out. With Painleve analysis and Hirota bilinear method, we find that the solution has six undetermined parameters as well as analyze the features of some typical cases of the solutions. Based on the constructed solutions, the conditions of their convergence are also discussed.

nlin.SI↗

Bringing a Blurry Frame Alive at High Frame-Rate with an Event Camera

Event-based cameras can measure intensity changes (called `{\it events}') with microsecond accuracy under high-speed motion and challenging lighting conditions. With the active pixel sensor (APS), the event camera allows simultaneous output of the intensity frames. However, the output images are captured at a relatively low frame-rate and often suffer from motion blur. A blurry image can be regarded as the integral of a sequence of latent images, while the events indicate the changes between the latent images. Therefore, we are able to model the blur-generation process by associating event data to a latent image. In this paper, we propose a simple and effective approach, the \textbf{Event-based Double Integral (EDI)} model, to reconstruct a high frame-rate, sharp video from a single blurry frame and its event data. The video generation is based on solving a simple non-convex optimization problem in a single scalar variable. Experimental results on both synthetic and real images demonstrate the superiority of our EDI model and optimization method in comparison to the state-of-the-art.

cs.CV↗

Generalized Lyapunov criteria on finite-time stability of stochastic nonlinear systems

This paper considers the problem of finite-time stability for stochastic nonlinear systems. A new Lyapunov theorem of stochastic finite-time stability is proposed, and an important corollary is obtained. Some comparisons with the existing results are given, and it shows that this new Lyapunov theorem not only is a generalization of classical stochastic finite-time theorem, but also reveals the important role of white-noise in finite-time stabilizing stochastic systems. In addition, multiple Lyapunov functions-based criteria on stochastic finite-time stability are presented, which further relax the constraint of the infinitesimal generator $\mathcal{L}V$. Some examples are constructed to show significant features of the proposed theorems. Finally, simulation results are presented to demonstrate the theoretical analysis.

math.PR↗

Estimating the Distribution of Random Parameters in a Diffusion Equation Forward Model for a Transdermal Alcohol Biosensor

We estimate the distribution of random parameters in a distributed parameter model with unbounded input and output for the transdermal transport of ethanol in humans. The model takes the form of a diffusion equation with the input being the blood alcohol concentration and the output being the transdermal alcohol concentration. Our approach is based on the idea of reformulating the underlying dynamical system in such a way that the random parameters are now treated as additional space variables. When the distribution to be estimated is assumed to be defined in terms of a joint density, estimating the distribution is equivalent to estimating the diffusivity in a multi-dimensional diffusion equation and thus well-established finite dimensional approximation schemes, functional analytic based convergence arguments, optimization techniques, and computational methods may all be employed. We use our technique to estimate a bivariate normal distribution based on data for multiple drinking episodes from a single subject.

math.OC↗

VLASE: Vehicle Localization by Aggregating Semantic Edges

In this paper, we propose VLASE, a framework to use semantic edge features from images to achieve on-road localization. Semantic edge features denote edge contours that separate pairs of distinct objects such as building-sky, road- sidewalk, and building-ground. While prior work has shown promising results by utilizing the boundary between prominent classes such as sky and building using skylines, we generalize this approach to consider semantic edge features that arise from 19 different classes. Our localization algorithm is simple, yet very powerful. We extract semantic edge features using a recently introduced CASENet architecture and utilize VLAD framework to perform image retrieval. Our experiments show that we achieve improvement over some of the state-of-the-art localization algorithms such as SIFT-VLAD and its deep variant NetVLAD. We use ablation study to study the importance of different semantic classes and show that our unified approach achieves better performance compared to individual prominent features such as skylines.

cs.CV↗

Identity-preserving Face Recovery from Portraits

Recovering the latent photorealistic faces from their artistic portraits aids human perception and facial analysis. However, a recovery process that can preserve identity is challenging because the fine details of real faces can be distorted or lost in stylized images. In this paper, we present a new Identity-preserving Face Recovery from Portraits (IFRP) to recover latent photorealistic faces from unaligned stylized portraits. Our IFRP method consists of two components: Style Removal Network (SRN) and Discriminative Network (DN). The SRN is designed to transfer feature maps of stylized images to the feature maps of the corresponding photorealistic faces. By embedding spatial transformer networks into the SRN, our method can compensate for misalignments of stylized faces automatically and output aligned realistic face images. The role of the DN is to enforce recovered faces to be similar to authentic faces. To ensure the identity preservation, we promote the recovered and ground-truth faces to share similar visual features via a distance measure which compares features of recovered and ground-truth faces extracted from a pre-trained VGG network. We evaluate our method on a large-scale synthesized dataset of real and stylized face pairs and attain state of the art results. In addition, our method can recover photorealistic faces from previously unseen stylized portraits, original paintings and human-drawn sketches.

cs.CV↗

Face Destylization

Numerous style transfer methods which produce artistic styles of portraits have been proposed to date. However, the inverse problem of converting the stylized portraits back into realistic faces is yet to be investigated thoroughly. Reverting an artistic portrait to its original photo-realistic face image has potential to facilitate human perception and identity analysis. In this paper, we propose a novel Face Destylization Neural Network (FDNN) to restore the latent photo-realistic faces from the stylized ones. We develop a Style Removal Network composed of convolutional, fully-connected and deconvolutional layers. The convolutional layers are designed to extract facial components from stylized face images. Consecutively, the fully-connected layer transfers the extracted feature maps of stylized images into the corresponding feature maps of real faces and the deconvolutional layers generate real faces from the transferred feature maps. To enforce the destylized faces to be similar to authentic face images, we employ a discriminative network, which consists of convolutional and fully connected layers. We demonstrate the effectiveness of our network by conducting experiments on an extensive set of synthetic images. Furthermore, we illustrate our network can recover faces from stylized portraits and real paintings for which the stylized data was unavailable during the training phase.

cs.CV↗

Solitons and breathers for nonisospectral mKdV equation with Darboux transformation

Under investigation in this paper is the nonisospectral and variable coefficients modified Kortweg-de Vries (vc-mKdV) equation, which manifests in diverse areas of physics such as fluid dynamics, ion acoustic solitons and plasma mechanics. With the degrees of restriction reduced, a simplified constraint is introduced, under which the vc-mKdV equation is an integrable system and the spectral flow is time-varying. The Darboux transformation for such equation is constructed, which gives rise to the generation of variable kinds of solutions including the double-breather coherent structure, periodical soliton-breather and localized solitons and breathers. In addition, the effect of variable coefficients and initial phases is discussed in terms of the soliton amplitude, polarity, velocity and width, which might provide feasible soliton management with certain conditions taken into account.

nlin.PS↗

Darboux Transformation for the Nonisospectral and Variable-coefficient KdV Equation

With the nonuniform media taken into account, the nonisospectral and variable-coefficient Korteweg-de Vries equation, which describes various physical situations such as fluid dynamics and plasma, is under investigation in this paper. With appropriate selection of wave functions, the Darboux transformation is constructed, by which the multi-soliton solutions are derived and graphs are presented. The spectral parameters, coefficients and initial phase are discussed analytically and numerically to demonstrate their respective effect on the soliton dynamics, which plays a role in achieving the feasible soliton management with explicit conditions taken into account.

nlin.PS↗

Spacial inhomogeneity and nonlinear tunneling for the forced KdV equation

A variable-coefficient forced Korteweg-de Vries equation with spacial inhomogeneity is investigated in this paper. Under constraints, this equation is transformed into its bilinear form, and multi-soliton solutions are derived. Effects of spacial inhomogeneity for soliton velocity, width and background are discussed. Nonlinear tunneling for this equation is presented, where the soliton amplitude can be amplified or compressed. Our results might be useful for the relevant problems in fluids and plasmas.

nlin.PS↗

Semileptonic decays $B_c^+\to D^{(*)}_{(s)}(l^+ν,l^+l^-,ν\barν)$ in the perturbative QCD approach

In this paper we study the semileptonic decays of $B_c^+\to D^{(*)}_{(s)}(l^+ν_l,l^+l^-,ν\barν)$ (here $l$ stands for $e$, $μ$, or $τ$). After evaluating the $B_c^+ \to (D_{(s)},D^*_{(s)})$ transition form factors $F_{0,+,T}(q^2)$ and $V(q^2), A_{0,1,2}(q^2), T_{1,2,3}(q^2)$ by employing the perturbative QCD factorization approach, we calculate the branching ratios for all these semileptonic decays. Our predictions for the values of the $B_c^+ \to D_{(s)}$ and $B_c^+ \to D^*_{(s)}$ transition form factors are consistent with those obtained by using other methods. The branching ratios of the decay modes with $\barνν$ are almost an order of magnitude larger than the corresponding decays with $l^+l^-$ after the summation over the three neutrino generations. The branching ratios for the decays with $b\to d$ transitions are much smaller than those decays with the $b\to s$ transitions, due to the Cabibbo-Kobayashi-Maskawa suppression. We define ratios $R_D$ and $R_{D^*}$ for the branching ratios with the $τ$ lepton versus $μ$, $e$ lepton final states to cancel the uncertainties of the form factors, which could possibly be tested in the near future.

hep-ph↗

The NLO twist-3 contributions to $B \to π$ form factors in $k_{T}$ factorization

In this paper, we calculate the next-to-leading-order (NLO) twist-3 contribution to the form factors of $B \to π$ transitions by employing the $k_{T}$ factorization theorem. All the infrared divergences regulated by the logarithms $\ln(k_{iT}^{2})$ cancel between those from the quark diagrams and from the effective diagrams for the initial $B$ meson wave function and the final pion meson wave function. An infrared finite NLO hard kernel is therefore obtained, which confirms the application of the $k_{T}$ factorization theorem to $B$ meson semileptonic decays at twist-3 level. From our analytical and numerical evaluations, we find that the NLO twist-3 contributions to the form factors $f^{+,0}(q^2)$ of $B \to π$ transition are similar in size, but have an opposite sign with the NLO twist-2 contribution, which leads to a large cancelation between these two NLO parts. For the case of $f^+(0)$, for example, the $24\%$ NLO twist-2 enhancement to the full LO prediction is largely canceled by the negative ( about $-17\%$ ) NLO twist-3 contribution, leaving a small and stable $7\%$ enhancement to the full LO prediction in the whole range of $0\leq q^2\leq 12$ GeV$^2$. At the full NLO level, the perturbative QCD prediction is $F^{B \to π}(0)=0.269^{+0.054}_{-0.050}$. We also studied the possible effects on the pQCD predictions when different sets of the B meson and pion distribution amplitudes are used in the numerical evaluation.

hep-ph↗

Global existence of null-form wave equations on small asymptotically Euclidean manifolds

We prove the global existence of the small solutions to the Cauchy problem for quasilinear wave equations satisfying the null condition on $(R^3, g)$, where the metric $g$ is a small perturbation of the flat metric and approaches the Euclidean metric like $(1+|x|)^{-a}$ with $a>1$. Global and almost global existence for systems without the null condition are also discussed for certain small time-dependent perturbations of the flat metric in the appendix.

math.AP↗

Perturbative QCD study of $B_s$ decays to a pseudoscalar meson and a tensor meson

We study two-body hadronic $B_s\to PT$ decays, with $P (T)$ being a light pseudoscalar (tensor) meson, in the perturbative QCD approach. The CP-averaged branching ratios and the direct CP asymmetries of the $ΔS=0$ modes are predicted, where $ΔS$ is the difference between the strange numbers of final and initial states. We also define and calculate experimental observables for the $ΔS=1$ modes under the $B_s^0-\bar{B}_s^0$ mixing, including CP averaged branching ratios, time-integrated CP asymmetries, and the CP observables $C_{f}$, $D_{f}$ and $S_{f}$. Results are compared to the $B_s\to PV$ ones in the literature, and to the $B\to PT$ ones, which indicate considerable U-spin symmetry breaking. Our work provides theoretical predictions for the $B_s\to PT$ decays for the first time, some of which will be potentially measurable at future experiments.

hep-ph↗