arXiv ScienceSearch

arXiv subjects

Yuanwei Wu

Publications and source records attributed to Yuanwei Wu.

At least 19 recordsLinked to original sources

An Approximately 70-Year Core-Related Modulation of Earth Rotation and Its Implications for the Leap Second

Recent observations of Universal Time (UT1) indicate an acceleration in Earth's rotation. If sustained under the current leap-second framework, this behavior could eventually prompt consideration of a negative leap second. We examine whether the recent acceleration is consistent with an approximately 70-year, core-related modulation of length of day (LOD). After removal of modeled tidal, surface-fluid, and secular contributions, residual LOD contains a near-70-year component, and a similar component is present in core angular momentum (CAM)-derived equivalent LOD inferred from geomagnetic observations. All harmonic, spectral, and LOD-CAM analyses reported here use the common 1883--2022 interval. Harmonic regression over trial periods of 50--100 yr gives periods of 69.7 yr for residual LOD and 71.8 yr for CAM-derived equivalent LOD, with amplitudes of 2.87 and 1.94 ms, respectively. Lomb--Scargle spectra show peaks near 67.8 and 70.5 yr. The annual series have a zero-lag correlation of 0.918. Their lagged correlation has a broad maximum for a CAM lead of approximately 1-3 yr, with a numerical maximum of 0.932 at 2 yr. Because both records are strongly autocorrelated, these coefficients are used to characterize their correspondence rather than to assess predictive significance. The results are consistent with a core-related contribution to low-frequency rotational variability, but they do not uniquely separate the contributions of electromagnetic, topographic, gravitational, and viscous core--mantle coupling mechanisms. Within the fitted model, the multidecadal component alone does not indicate sustained near-term shortening of the day that would, by itself, require a negative leap second. This is a model-dependent geophysical assessment, not an operational prediction of future UTC adjustments.

astro-ph.EP

East Asian VLBI Network astrometry toward the star-forming region G040.96+02.48 in the Extreme Outer Galaxy

Accurate astrometric measurements for star-forming regions located on the far side of the Milky Way remain scarce. In this work, we present the astrometric results for a 22\,GHz water maser associated with star-forming region G040.96+02.48 located on the far side of the Milky Way, using the East Asian VLBI Network. The target water maser's proper motion was determined to be ($\mu_{\alpha}\cos\delta, \mu_{\delta}$) = ($-2.06_{-0.51}^{+0.53}$, $-2.95_{-0.44}^{+0.45}$)~mas~yr$^{-1}$. The derived three-dimensional kinematic distance to the star-forming region is 20.2$\pm$3.2\,kpc, placing it slightly outside the Outer Scutum$-$Centaurus Arm. The corresponding vertical height of 872$\pm$139\,pc indicates a significant warp of the outer Galactic disk, which is in good agreement with the latest precessing warp model. Moreover, the resulting peculiar motions reveal a complex kinematic pattern, characterized by a large outward radial velocity of $-32\pm$18\,km~s$^{-1}$. Our observations substantially expand the valuable sample of star-forming regions with accurate astrometric measurements in the Extreme Outer Galaxy.

astro-ph.GA

Restore-R1: Efficient Image Restoration Agents via Reinforcement Learning with Multimodal LLM Perceptual Feedback

Complex image restoration aims to recover high-quality images from inputs affected by multiple degradations such as blur, noise, rain, and compression artifacts. Recent restoration agents, powered by vision-language models and large language models, offer promising restoration capabilities but suffer from significant efficiency bottlenecks due to reflection, rollback, and iterative tool searching. Moreover, their performance heavily depends on degradation recognition models that require extensive annotations for training, limiting their applicability in label-free environments. To address these limitations, we propose a policy optimization-based restoration framework that learns an lightweight agent to determine tool-calling sequences. The agent operates in a sequential decision process, selecting the most appropriate restoration operation at each step to maximize final image quality. To enable training within label-free environments, we introduce a novel reward mechanism driven by multimodal large language models, which act as human-aligned evaluator and provide perceptual feedback for policy improvement. Once trained, our agent executes a deterministic restoration plans without redundant tool invocations, significantly accelerating inference while maintaining high restoration quality. Extensive experiments show that despite using no supervision, our method matches SOTA performance on full-reference metrics and surpasses existing approaches on no-reference metrics across diverse degradation scenarios.

cs.CV

Can Large Language Models Automatically Jailbreak GPT-4V?

GPT-4V has attracted considerable attention due to its extraordinary capacity for integrating and processing multimodal information. At the same time, its ability of face recognition raises new safety concerns of privacy leakage. Despite researchers' efforts in safety alignment through RLHF or preprocessing filters, vulnerabilities might still be exploited. In our study, we introduce AutoJailbreak, an innovative automatic jailbreak technique inspired by prompt optimization. We leverage Large Language Models (LLMs) for red-teaming to refine the jailbreak prompt and employ weak-to-strong in-context learning prompts to boost efficiency. Furthermore, we present an effective search method that incorporates early stopping to minimize optimization time and token expenditure. Our experiments demonstrate that AutoJailbreak significantly surpasses conventional methods, achieving an Attack Success Rate (ASR) exceeding 95.3\%. This research sheds light on strengthening GPT-4V security, underscoring the potential for LLMs to be exploited in compromising GPT-4V integrity.

cs.CL

Jailbreaking GPT-4V via Self-Adversarial Attacks with System Prompts

Existing work on jailbreak Multimodal Large Language Models (MLLMs) has focused primarily on adversarial examples in model inputs, with less attention to vulnerabilities, especially in model API. To fill the research gap, we carry out the following work: 1) We discover a system prompt leakage vulnerability in GPT-4V. Through carefully designed dialogue, we successfully extract the internal system prompts of GPT-4V. This finding indicates potential exploitable security risks in MLLMs; 2) Based on the acquired system prompts, we propose a novel MLLM jailbreaking attack method termed SASP (Self-Adversarial Attack via System Prompt). By employing GPT-4 as a red teaming tool against itself, we aim to search for potential jailbreak prompts leveraging stolen system prompts. Furthermore, in pursuit of better performance, we also add human modification based on GPT-4's analysis, which further improves the attack success rate to 98.7\%; 3) We evaluated the effect of modifying system prompts to defend against jailbreaking attacks. Results show that appropriately designed system prompts can significantly reduce jailbreak success rates. Overall, our work provides new insights into enhancing MLLM security, demonstrating the important role of system prompts in jailbreaking. This finding could be leveraged to greatly facilitate jailbreak success rates while also holding the potential for defending against jailbreaks.

cs.CR

Maser Investigation toward Off-Plane Stars (MIOPS): detection of SiO masers in the Galactic thick disk and halo

Studying stars that are located off the Galactic plane is important for understanding the formation history of the Milky Way. We searched for SiO masers toward off-plane O-rich asymptotic giant branch (AGB) stars from the catalog presented by Mauron et al. (2019) in order to shed light on the origin of these objects. A total of 102 stars were observed in the SiO $J$=1-0, $v=1$ and 2 transitions with the Effelsberg-100 m and Tianma-65 m telescopes. SiO masers were discovered in eight stars, all first detections. The measured maser velocities allow the first estimates of the host AGB stars' radial velocities. We find that the radial velocities of three stars (namely G068.881-24.615, G070.384-24.886, and G084.453-21.863) significantly deviate from the values expected from Galactic circular motion. The updated distances and 3D motions indicate that G068.881$-$24.615 is likely located in the Galactic halo, while G160.648-08.846 is probably located in the Galactic thin disk, and the other six stars are probably part of the Galactic thick disk.

astro-ph.GA

Water Maser Survey towards off-plane O-rich AGBs around the orbital plane of the Sagittarius Stellar Stream

A 22 GHz water maser survey was conducted towards 178 O-rich AGB stars with the aim of identifying maser emission associated with the Sagittarius stellar stream. In this survey, maser emissions were detected in 21 targets, of which 20 were new detections. We studied the Galactic distributions of H2O and SiO maser-traced AGBs towards the Sgr orbital plane, and found an elongated structure towards the (l, b)~(340, 40) direction. In order to verify its association with the Sagittarius tidal stream, we further studied the 3D motions of these sources, but found, kinematically, these maser-traced AGBs are still Galactic disc sources rather than Stream debris. In addition, we found a remarkable outward motion, ~50 km/s away from the Galactic center of these maser-traced AGBs, but with no systermatic lag of rotational speed which were reported in 2000 for solar neighborhood Miras.

astro-ph.GA

Light Deflection under the Gravitational Field of Jupiter -- Testing General Relativity

We measured the relative positions between two pairs of compact extragalactic sources (CESs), J1925-2219 \& J1923-2104 (C1--C2) and J1925-2219 \& J1928-2035 (C1--C3) on 2020 October 23--25 and 2021 February 5 (totaling four epochs), respectively, using the Very Long Baseline Array (VLBA) at 15 GHz. Accounting for the deflection angle dominated by Jupiter, as well as the contributions from the Sun, planets other than Earth, the Moon and Ganymede (the most massive of the solar system's moons), our theoretical calculations predict that the dynamical ranges of the relative positions across four epochs in R.A. of the C1--C2 pair and C1--C3 pair are 841.2 and 1127.9 $\mu$as, respectively. The formal accuracy in R.A. is about 20 $\mu$as, but the error in Decl. is poor. The measured standard deviations of the relative positions across the four epochs are 51.0 and 29.7 $\mu$as in R.A. for C1--C2 and C1--C3, respectively. These values indicate that the accuracy of the post-Newtonian relativistic parameter, $\gamma$, is $\sim 0.061$ for C1--C2 and $\sim 0.026$ for C1--C3. Combining the two CES pairs, the measured value of $\gamma$ is $0.984 \pm 0.037$, which is comparable to the latest published results for Jupiter as a gravitational lens reported by Fomalont \& Kopeikin, i.e., $1.01 \pm 0.03$.

gr-qc

Training Deep Neural Networks via Branch-and-Bound

In this paper, we propose BPGrad, a novel approximate algorithm for deep nueral network training, based on adaptive estimates of feasible region via branch-and-bound. The method is based on the assumption of Lipschitz continuity in objective function, and as a result, it can adaptively determine the step size for the current gradient given the history of previous updates. We prove that, by repeating such a branch-and-pruning procedure, it can achieve the optimal solution within finite iterations. A computationally efficient solver based on BPGrad has been proposed to train the deep neural networks. Empirical results demonstrate that BPGrad solver works well in practice and compares favorably to other stochastic optimization methods in the tasks of object recognition, detection, and segmentation. The code is available at \url{https://github.com/RyanCV/BPGrad}.

cs.CV

Self-Orthogonality Module: A Network Architecture Plug-in for Learning Orthogonal Filters

In this paper, we investigate the empirical impact of orthogonality regularization (OR) in deep learning, either solo or collaboratively. Recent works on OR showed some promising results on the accuracy. In our ablation study, however, we do not observe such significant improvement from existing OR techniques compared with the conventional training based on weight decay, dropout, and batch normalization. To identify the real gain from OR, inspired by the locality sensitive hashing (LSH) in angle estimation, we propose to introduce an implicit self-regularization into OR to push the mean and variance of filter angles in a network towards 90 and 0 simultaneously to achieve (near) orthogonality among the filters, without using any other explicit regularization. Our regularization can be implemented as an architectural plug-in and integrated with an arbitrary network. We reveal that OR helps stabilize the training process and leads to faster convergence and better generalization.

cs.CV

MDFN: Multi-Scale Deep Feature Learning Network for Object Detection

This paper proposes an innovative object detector by leveraging deep features learned in high-level layers. Compared with features produced in earlier layers, the deep features are better at expressing semantic and contextual information. The proposed deep feature learning scheme shifts the focus from concrete features with details to abstract ones with semantic information. It considers not only individual objects and local contexts but also their relationships by building a multi-scale deep feature learning network (MDFN). MDFN efficiently detects the objects by introducing information square and cubic inception modules into the high-level layers, which employs parameter-sharing to enhance the computational efficiency. MDFN provides a multi-scale object detector by integrating multi-box, multi-scale and multi-level technologies. Although MDFN employs a simple framework with a relatively small base network (VGG-16), it achieves better or competitive detection results than those with a macro hierarchical structure that is either very deep or very wide for stronger ability of feature extraction. The proposed technique is evaluated extensively on KITTI, PASCAL VOC, and COCO datasets, which achieves the best results on KITTI and leading performance on PASCAL VOC and COCO. This study reveals that deep features provide prominent semantic information and a variety of contextual contents, which contribute to its superior performance in detecting small or occluded objects. In addition, the MDFN model is computationally efficient, making a good trade-off between the accuracy and speed.

cs.CV

Object Detection with Convolutional Neural Networks

In this chapter, we present a brief overview of the recent development in object detection using convolutional neural networks (CNN). Several classical CNN-based detectors are presented. Some developments are based on the detector architectures, while others are focused on solving certain problems, like model degradation and small-scale object detection. The chapter also presents some performance comparison results of different models on several benchmark datasets. Through the discussion of these models, we hope to give readers a general idea about the developments of CNN-based object detection.

cs.CV

Adaptively Denoising Proposal Collection for Weakly Supervised Object Localization

In this paper, we address the problem of weakly supervised object localization (WSL), which trains a detection network on the dataset with only image-level annotations. The proposed approach is built on the observation that the proposal set from the training dataset is a collection of background, object parts, and objects. Several strategies are taken to adaptively eliminate the noisy proposals and generate pseudo object-level annotations for the weakly labeled dataset. A multiple instance learning (MIL) algorithm enhanced by mask-out strategy is adopted to collect the class-specific object proposals, which are then utilized to adapt a pre-trained classification network to a detection network. In addition, the detection results from the detection network are re-weighted by jointly considering the detection scores and the overlap ratio of proposals in a proposal subset optimization framework. The optimal proposals work as object-level labels that enable a pseudo-strongly supervised dataset for training the detection network. Consequently, we establish a fully adaptive detection network. Extensive evaluations on the PASCAL VOC 2007 and 2012 datasets demonstrate a significant improvement compared with the state-of-the-art methods.

cs.CV

Unsupervised Deep Feature Transfer for Low Resolution Image Classification

In this paper, we propose a simple while effective unsupervised deep feature transfer algorithm for low resolution image classification. No fine-tuning on convenet filters is required in our method. We use pre-trained convenet to extract features for both high- and low-resolution images, and then feed them into a two-layer feature transfer network for knowledge transfer. A SVM classifier is learned directly using these transferred low resolution features. Our network can be embedded into the state-of-the-art deep neural networks as a plug-in feature enhancement module. It preserves data structures in feature space for high resolution images, and transfers the distinguishing features from a well-structured source domain (high resolution features space) to a not well-organized target domain (low resolution features space). Extensive experiments on VOC2007 test set show that the proposed method achieves significant improvements over the baseline of using feature extraction.

cs.CV

Parallaxes for star forming regions in the inner Perseus spiral arm

We report trigonometric parallax and proper motion measurements of 6.7-GHz CH3OH and 22-GHz H2O masers in eight high-mass star-forming regions (HMSFRs) based on VLBA observations as part of the BeSSeL Survey. The distances of these HMSFRs combined with their Galactic coordinates, radial velocities, and proper motions, allow us to assign them to a segment of the Perseus arm with ~< 70 deg. These HMSFRs are clustered in Galactic longitude from ~30 deg to ~50, neighboring a dirth of such sources between longitudes ~50 deg to ~90 deg.

astro-ph.GA

The spiral structure of the Milky Way

The morphology and kinematics of the spiral structure of the Milky Way is a long-standing problem in astrophysics. In this review we firstly summarize various methods with different tracers used to solve this puzzle. The astrometry of Galactic sources is gradually alleviating this difficult situation caused mainly by large distance uncertainties, as we can currently obtain accurate parallaxes (a few $\mu$as) and proper motions ($\approx$ 1 km~s$^{-1}$) by using Very Long Baseline Interferometry (VLBI). On the other hand, Gaia mission is providing the largest, uniform sample of parallaxes for O-type stars in the entire Milky Way. Based upon the VLBI maser and Gaia O-star parallax measurements, nearby spiral structures: the Perseus, Local, Sagittarius and Scutum arms are determined in unprecedented detail. Meanwhile, we estimate fundamental Galactic parameters, the distance to the Galactic center, $R_0$, to be 8.35 $\pm$ 0.18 kpc, a circular rotation speed at the Sun, $\Theta_0$, to be 240 $\pm$ 10 km~s$^{-1}$. We found kinematic differences between O stars and interstellar masers: the O stars, on average, rotate faster, $>$~8~ km s$^{-1}$ than maser-traced high-mass star forming regions.

astro-ph.GA

MDCN: Multi-Scale, Deep Inception Convolutional Neural Networks for Efficient Object Detection

Object detection in challenging situations such as scale variation, occlusion, and truncation depends not only on feature details but also on contextual information. Most previous networks emphasize too much on detailed feature extraction through deeper and wider networks, which may enhance the accuracy of object detection to certain extent. However, the feature details are easily being changed or washed out after passing through complicated filtering structures. To better handle these challenges, the paper proposes a novel framework, multi-scale, deep inception convolutional neural network (MDCN), which focuses on wider and broader object regions by activating feature maps produced in the deep part of the network. Instead of incepting inner layers in the shallow part of the network, multi-scale inceptions are introduced in the deep layers. The proposed framework integrates the contextual information into the learning process through a single-shot network structure. It is computational efficient and avoids the hard training problem of previous macro feature extraction network designed for shallow layers. Extensive experiments demonstrate the effectiveness and superior performance of MDCN over the state-of-the-art models.

cs.CV

BPGrad: Towards Global Optimality in Deep Learning via Branch and Pruning

Understanding the global optimality in deep learning (DL) has been attracting more and more attention recently. Conventional DL solvers, however, have not been developed intentionally to seek for such global optimality. In this paper we propose a novel approximation algorithm, BPGrad, towards optimizing deep models globally via branch and pruning. Our BPGrad algorithm is based on the assumption of Lipschitz continuity in DL, and as a result it can adaptively determine the step size for current gradient given the history of previous updates, wherein theoretically no smaller steps can achieve the global optimality. We prove that, by repeating such branch-and-pruning procedure, we can locate the global optimality within finite iterations. Empirically an efficient solver based on BPGrad for DL is proposed as well, and it outperforms conventional DL solvers such as Adagrad, Adadelta, RMSProp, and Adam in the tasks of object recognition, detection, and segmentation.

stat.ML