arXiv ScienceSearch

arXiv subjects

Qing Yu

Publications and source records attributed to Qing Yu.

At least 55 records · Page 3Linked to original sources

Can Pre-trained Networks Detect Familiar Out-of-Distribution Data?

Out-of-distribution (OOD) detection is critical for safety-sensitive machine learning applications and has been extensively studied, yielding a plethora of methods developed in the literature. However, most studies for OOD detection did not use pre-trained models and trained a backbone from scratch. In recent years, transferring knowledge from large pre-trained models to downstream tasks by lightweight tuning has become mainstream for training in-distribution (ID) classifiers. To bridge the gap between the practice of OOD detection and current classifiers, the unique and crucial problem is that the samples whose information networks know often come as OOD input. We consider that such data may significantly affect the performance of large pre-trained networks because the discriminability of these OOD data depends on the pre-training algorithm. Here, we define such OOD data as PT-OOD (Pre-Trained OOD) data. In this paper, we aim to reveal the effect of PT-OOD on the OOD detection performance of pre-trained networks from the perspective of pre-training algorithms. To achieve this, we explore the PT-OOD detection performance of supervised and self-supervised pre-training algorithms with linear-probing tuning, the most common efficient tuning method. Through our experiments and analysis, we find that the low linear separability of PT-OOD in the feature space heavily degrades the PT-OOD detection performance, and self-supervised models are more vulnerable to PT-OOD than supervised pre-trained models, even with state-of-the-art detection methods. To solve this vulnerability, we further propose a unique solution to large-scale pre-trained models: Leveraging powerful instance-by-instance discriminative representations of pre-trained models and detecting OOD in the feature space independent of the ID decision boundaries. The code will be available via https://github.com/AtsuMiyai/PT-OOD.

cs.CV

CPR-Coach: Recognizing Composite Error Actions based on Single-class Training

The fine-grained medical action analysis task has received considerable attention from pattern recognition communities recently, but it faces the problems of data and algorithm shortage. Cardiopulmonary Resuscitation (CPR) is an essential skill in emergency treatment. Currently, the assessment of CPR skills mainly depends on dummies and trainers, leading to high training costs and low efficiency. For the first time, this paper constructs a vision-based system to complete error action recognition and skill assessment in CPR. Specifically, we define 13 types of single-error actions and 74 types of composite error actions during external cardiac compression and then develop a video dataset named CPR-Coach. By taking the CPR-Coach as a benchmark, this paper thoroughly investigates and compares the performance of existing action recognition models based on different data modalities. To solve the unavoidable Single-class Training & Multi-class Testing problem, we propose a humancognition-inspired framework named ImagineNet to improve the model's multierror recognition performance under restricted supervision. Extensive experiments verify the effectiveness of the framework. We hope this work could advance research toward fine-grained medical action analysis and skill assessment. The CPR-Coach dataset and the code of ImagineNet are publicly available on Github.

cs.CV

Open-Set Domain Adaptation with Visual-Language Foundation Models

Unsupervised domain adaptation (UDA) has proven to be very effective in transferring knowledge obtained from a source domain with labeled data to a target domain with unlabeled data. Owing to the lack of labeled data in the target domain and the possible presence of unknown classes, open-set domain adaptation (ODA) has emerged as a potential solution to identify these classes during the training phase. Although existing ODA approaches aim to solve the distribution shifts between the source and target domains, most methods fine-tuned ImageNet pre-trained models on the source domain with the adaptation on the target domain. Recent visual-language foundation models (VLFM), such as Contrastive Language-Image Pre-Training (CLIP), are robust to many distribution shifts and, therefore, should substantially improve the performance of ODA. In this work, we explore generic ways to adopt CLIP, a popular VLFM, for ODA. We investigate the performance of zero-shot prediction using CLIP, and then propose an entropy optimization strategy to assist the ODA models with the outputs of CLIP. The proposed approach achieves state-of-the-art results on various benchmarks, demonstrating its effectiveness in addressing the ODA problem.

cs.CV

Noisy Universal Domain Adaptation via Divergence Optimization for Visual Recognition

To transfer the knowledge learned from a labeled source domain to an unlabeled target domain, many studies have worked on universal domain adaptation (UniDA), where there is no constraint on the label sets of the source domain and target domain. However, the existing UniDA methods rely on source samples with correct annotations. Due to the limited resources in the real world, it is difficult to obtain a large amount of perfectly clean labeled data in a source domain in some applications. As a result, we propose a novel realistic scenario named Noisy UniDA, in which classifiers are trained using noisy labeled data from the source domain as well as unlabeled domain data from the target domain that has an uncertain class distribution. A multi-head convolutional neural network framework is proposed in this paper to address all of the challenges faced in the Noisy UniDA at once. Our network comprises a single common feature generator and multiple classifiers with various decision bounds. We can detect noisy samples in the source domain, identify unknown classes in the target domain, and align the distribution of the source and target domains by optimizing the divergence between the outputs of the various classifiers. The proposed method outperformed the existing methods in most of the settings after a thorough analysis of the various domain adaption scenarios. The source code is available at \url{https://github.com/YU1ut/Divergence-Optimization}.

cs.CV

Qubit Energy Tuner Based on Single Flux Quantum Circuits

A device called qubit energy tuner (QET) based on single flux quantum (SFQ) circuits is proposed for Z control of superconducting qubits. Created from the improvement of flux digital-to-analog converters (flux DACs), a QET is able to set the energy levels or the frequencies of qubits, especially flux-tunable transmons, and perform gate operations requiring Z control. The circuit structure of QET is elucidated, which consists of an inductor loop and flux bias units for coarse tuning or fine tuning. The key feature of a QET is analyzed to understand how SFQ pulses change the inductor loop current, which provides external flux for qubits. To verify the functionality of the QET, three simulations are carried out. The first one verifies the responses of the inductor loop current to SFQ pulses. The results show that there is about 4.2% relative deviation between analytical solutions of the inductor loop current and the solutions from WRSpice time-domain simulation. The second and the third simulations with QuTip show how a Z gate and an iSWAP gate can be performed by this QET, respectively, with corresponding fidelities 99.99884% and 99.93906% for only once gate operation to specific initial states. These simulations indicate that the SFQ-based QET could act as an efficient component of SFQ-based quantum-classical interfaces for digital Z control of large-scale superconducting quantum computers.

quant-ph

New determination of $|V_{\rm cb}|$ using the three-loop QCD corrections for the $B\to D^{\ast}$ semi-leptonic decays

We present a new determination of the Cabibbo-Kobayashi-Maskawa matrix element $|V_{\rm cb}|$ by using the three-loop perturbative QCD corrections for the $B\to D^{\ast}$ semi-leptonic decay. The decay width of $B\to D^{\ast}$ semi-leptonic decay can be factorized as perturbatively calculable short-distance part and the non-perturbative but universal long-distance part. We adopt the principle of maximum conformality (PMC) single-scale setting approach to deal with the perturbative series so as to achieve a precise fixed-order prediction for the short-distance parameter $η_{A}$. By applying the PMC, an overall effective $α_s$ value is achieved by recursively using the renormalization group equation, which inversely results in a precise scale-invariant pQCD series. Such scale-invariant series also provides a reliable basis for predicting the contributions from uncalculated perturbative terms. We then obtain $η_{A}=0.9225^{+0.0117}_{-0.0168}$, where the error is the squared average of those from $Δα_{s}(M_Z)=\pm0.0010$ and the uncertainties caused by the uncalculated higher-order perturbative terms. By using the data of $B\to D^{\ast}\ell\barν_{\ell}$, we finally obtain $|V_{\rm cb}|_{\rm PMC} =(40.60^{+0.53}_{-0.57})\times10^{-3}$, which is consistent with the PDG value within errors.

hep-ph

Thermodynamics and quark condensates of three-flavor QCD at low temperature

We use three-flavor chiral perturbation theory ($χ$PT) to calculate the pressure, light and $s$-quark condensates of QCD in the confined phase at finite temperature to ${\cal O}(p^6)$ in the low-energy expansion. We also include electromagnetic effects to order $e^2$, where the electromagnetic coupling $e$ counts as order $p$. Our results for the pressure and the condensates suggest that $χ$PT converges very well for temperatures up to approximately 150 MeV. We combine $χ$PT and the Hadron Resonance Gas (HRG) model by adding heavier baryons and mesons. Our results are compared with lattice simulations an d the agreement is very good for temperatures below {170} MeV, in contrast to the results from $χ$PT which agree with the lattice only up to $T\approx120$ MeV. Our value for the chiral crossover temperature is 160.1 MeV, which compares favorably to the lattice result of $157.3$ MeV.

hep-ph

Rethinking Rotation in Self-Supervised Contrastive Learning: Adaptive Positive or Negative Data Augmentation

Rotation is frequently listed as a candidate for data augmentation in contrastive learning but seldom provides satisfactory improvements. We argue that this is because the rotated image is always treated as either positive or negative. The semantics of an image can be rotation-invariant or rotation-variant, so whether the rotated image is treated as positive or negative should be determined based on the content of the image. Therefore, we propose a novel augmentation strategy, adaptive Positive or Negative Data Augmentation (PNDA), in which an original and its rotated image are a positive pair if they are semantically close and a negative pair if they are semantically different. To achieve PNDA, we first determine whether rotation is positive or negative on an image-by-image basis in an unsupervised way. Then, we apply PNDA to contrastive learning frameworks. Our experiments showed that PNDA improves the performance of contrastive learning. The code is available at \url{ https://github.com/AtsuMiyai/rethinking_rotation}.

cs.CV

Reheating constraints on modified single-field Natural Inflation models

In this paper, we discuss three modified single-field natural inflation models in detail, including Special generalized Natural Inflation model(SNI), Extended Natural Inflation model(ENI) and Natural Inflation inspired model(NII). We derive the analytical expression of the tensor-to-scalar ratio $r$ and the spectral index $n_s$ for those models. Then the reheating temperature $T_{re}$ and reheating duration $N_{re}$ are analytically derived. Moreover, considering the CMB constraints, the feasible space of the SNI model in $(n_s, r)$ plane is almost covered by that of the NII, which means the NII is more general than the SNI. In addition, there is no overlapping space between the ENI and the other two models in $(n_s, r)$ plane, which indicates that the ENI and the other two models exclude each other, and more accurate experiments can verify them. Furthermore, the reheating brings tighter constraints to the inflation models, but they still work for a different reheating universe. Considering the constraints of $n_s$, $r$, $N_k$ and choosing $T_{re}$ near the electroweak energy scale, one can find that the decay constants of the three models have no overlapping area and the effective equations of state $ω_{re}$ should be within $\frac{1}{4}\lesssim ω_{re} \lesssim \frac{4}{5}$ for the three models.

hep-ph

Novel and self-consistency analysis of the QCD running coupling $α_s(Q)$ in both the perturbative and nonperturbative domains

The QCD coupling $α_s$ is the most important parameter for achieving precise QCD predictions. By using the well measured effective coupling $α^{g_1}_{s}(Q)$ defined from the Bjorken sum rules as a basis, we suggest a novel and self-consistency way to fix the $α_s$ at all scales: The QCD light-front holographic model is adopted for its infrared behavior, and the fixed-order pQCD prediction under the principle of maximum conformality (PMC) is used for its high-energy behavior. Using the PMC scheme-and-scale independent perturbative series, and by transforming it into the one under the physical $V$-scheme, we observe that a precise $α_s$ running behavior in both the perturbative and nonperturbative domains with a smooth transition from small to large scales can be achieved.

hep-ph

A Survey of Video-based Action Quality Assessment

Human action recognition and analysis have great demand and important application significance in video surveillance, video retrieval, and human-computer interaction. The task of human action quality evaluation requires the intelligent system to automatically and objectively evaluate the action completed by the human. The action quality assessment model can reduce the human and material resources spent in action evaluation and reduce subjectivity. In this paper, we provide a comprehensive survey of existing papers on video-based action quality assessment. Different from human action recognition, the application scenario of action quality assessment is relatively narrow. Most of the existing work focuses on sports and medical care. We first introduce the definition and challenges of human action quality assessment. Then we present the existing datasets and evaluation metrics. In addition, we summarized the methods of sports and medical care according to the model categories and publishing institutions according to the characteristics of the two fields. At the end, combined with recent work, the promising development direction in action quality assessment is discussed.

cs.CV

Noisy Annotation Refinement for Object Detection

Supervised training of object detectors requires well-annotated large-scale datasets, whose production is costly. Therefore, some efforts have been made to obtain annotations in economical ways, such as cloud sourcing. However, datasets obtained by these methods tend to contain noisy annotations such as inaccurate bounding boxes and incorrect class labels. In this study, we propose a new problem setting of training object detectors on datasets with entangled noises of annotations of class labels and bounding boxes. Our proposed method efficiently decouples the entangled noises, corrects the noisy annotations, and subsequently trains the detector using the corrected annotations. We verified the effectiveness of our proposed method and compared it with the baseline on noisy datasets with different noise levels. The experimental results show that our proposed method significantly outperforms the baseline.

cs.CV

A new analysis of the pQCD contributions to the electroweak parameter $ρ$ using the single-scale approach of principle of maximum conformality

It has been observed that conventional renormalization scheme and scale ambiguities for the pQCD predictions can be eliminated by using the principle of maximum conformality (PMC). However, being the intrinsic nature of any perturbative theory, there are still two types of residual scale dependences due to uncalculated higher-order terms. In the paper, as a step forward of our previous work [Phys.Rev.D {\bf 89},116001(2014)], we reanalyze the electroweak $ρ$ parameter by using the PMC single-scale approach. Using the PMC conformal series and the Pad$\acute{e}$ approximation approach, we observe that the residual scale dependence can be greatly suppressed and then a more precise pQCD prediction up to ${\rm N^4LO}$-level can be achieved, e.g. $Δρ|_{\rm PMC}\simeq(8.204\pm0.012)\times10^{-3}$, where the errors are squared averages of those from unknown higher-order terms and $Δα_s(M_Z)=\pm 0.0010$. We then predict the magnitudes of the shifts of the $W$-boson mass and the effective leptonic weak-mixing angle: $δM_{W}|_{\rm N^4LO} =-0.26$ MeV and $δ\sin^2θ_{\rm eff}|_{\rm N^4LO}=0.14\times10^{-5}$, which are well below the precision anticipated for the future electron-position colliders such as FCC, CEPC and ILC. Thus by measuring those parameters, it is possible to test SM with high precision.

hep-ph

Generalized Crewther relation and a novel demonstration of the scheme independence of commensurate scale relations up to all orders

In the paper, we make a detailed study on the generalized Crewther Relation (GCR) between the Adler function ($D$) and the Gross-Llewellyn Smith sum rules coefficient ($C^{\rm GLS}$) by using the newly suggested single-scale approach of the principle of maximum conformality (PMC). The resultant GCR is scheme-independent, whose residual scale dependence due to unknown higher-order terms are highly suppressed. Thus a precise test of QCD theory without renormalization scheme and scale ambiguities can be achieved by comparing with the data. Moreover, a demonstration of the scheme independence of commensurate scale relation up to all orders has been presented. And as the first time, the Pade approximation approach has been adopted for estimating the unknown $5_{\rm th}$-loop contributions from the known four-loop perturbative series.

hep-ph

A novel determination of non-perturbative contributions to Bjorken sum rule

In the present paper, we first give a detailed study on the pQCD corrections to the leading-twist part of BSR. Previous pQCD corrections to the leading-twist part derived under conventional scale-setting approach up to ${\cal O}(α_s^4)$-level still show strong renormalization scale dependence. The principle of maximum conformality (PMC) provides a systematic way to eliminate conventional renormalization scale-setting ambiguity by determining the accurate $α_s$-running behavior of the process with the help of renormalization group equation. Our calculation confirms the PMC prediction satisfies the standard renormalization group invariance, e.g. its fixed-order prediction does scheme-and-scale independent. In low $Q^2$-region, the effective momentum of the process is small and to have a reliable prediction, we adopt four low-energy $α_s$ models to do the analysis. Our predictions show that even though the high-twist terms are generally power suppressed in high $Q^2$-region, they shall have sizable contributions in low and intermediate $Q^2$ domain. By using the more accurate scheme-and-scale independent pQCD prediction, we present a novel fit of the non-perturbative high-twist contributions by comparing with the JLab data.

hep-ph

Gravitational Wave From Axion-like Particle Inflation

In this paper, we investigate the Axion-like Particle inflation by applying the multi-nature inflation model, where the end of inflation is achieved through the phase transition (PT). The events of PT should not be less than $200$, which results in the free parameter $n\geq404$. Under the latest CMB restrictions, we found that the inflation energy is fixed at $10^{15} \rm{GeV}$. Then, we deeply discussed the corresponding stochastic background of the primordial gravitational wave (GW) during inflation. We study the two kinds of $n$ cases, i.e., $n=404, 2000$. We observe that the magnitude of $n$ is negligible for the physical observations, such as $n_s$, $r$, $Λ$, and $Ω_{\rm{GW}}h^2$. In the low-frequency regions, the GW is dominated by the quantum fluctuations, and this GW can be detected by Decigo at $10^{-1}~\rm{Hz}$. However, GW generated by PT dominates the high-frequency regions, which is expected to be detected by future 3DSR detector.

hep-ph

The $P$-wave charmonium annihilation into two photons $χ_{c0, c2}\rightarrow γγ$ with high-order QCD corrections

In this paper, we present a new analysis on the $P$-wave charmonium annihilation into two photons up to next-to-next-to-leading order (NNLO) QCD corrections by using the principle of maximum conformality (PMC). The conventional perturbative QCD prediction shows strong scale dependence and deviates largely from the BESIII measurements. After applying the PMC, we obtain a more precise scale-invariant pQCD prediction, which also agrees with the BESIII measurements within errors, i.e. $R={Γ_{γγ}(χ_{c2})} /{Γ_{γγ}(χ_{c0})}=0.246\pm0.013$, where the error is for $Δα_s(M_τ)=\pm0.016$. By further considering the color-octet contributions, even the central value can be in agreement with the data. This shows the importance of a correct scale-setting approach. We also give a prediction for the ratio involving $χ_{b0, b2} \toγγ$, which could be tested in future Belle II experiment.

hep-ph

The Gross-Llewellyn Smith sum rule up to ${\cal O}(α_s^4)$-order QCD corrections

In the paper, we analyze the properties of Gross-Llewellyn Smith (GLS) sum rule by using the $\mathcal{O}(α_s^4)$-order QCD corrections with the help of principle of maximum conformality (PMC). By using the PMC single-scale approach, we obtain an accurate renormalization scale-and-scheme independent fixed-order pQCD contribution for GLS sum rule, e.g. $S^{\rm GLS}(Q_0^2=3{\rm GeV}^2)|_{\rm PMC}=2.559^{+0.023}_{-0.024}$, where the error is squared average of those from $Δα_s(M_Z)$, the predicted $\mathcal{O}(α_s^5)$-order terms predicted by using the Padé approximation approach. After applying the PMC, a more convergent pQCD series has been obtained, and the contributions from the unknown higher-order terms are highly suppressed. In combination with the nonperturbative high-twist contribution, our final prediction of GLS sum rule agrees well with the experimental data given by the CCFR collaboration.

hep-ph