arXiv ScienceSearch

arXiv subjects

Biao Wang

Publications and source records attributed to Biao Wang.

At least 19 recordsLinked to original sources

Dynamical generalizations of Chowla's conjecture on short averages

In 1965, Chowla conjectured that the signs of the Liouville function become asymptotically uncorrelated at any fixed collection of distinct shifts. In 2015, Matomäki, Radziwiłł and Tao proved an averaged form of Chowla's conjecture. In 2022, Lichtman proved a variant of this conjecture over primes on average. In the same year, Lichtman and Teräväinen proved the Hardy--Littlewood--Chowla conjecture on average. In this article, motivated by the recent work of Bergelson and Richter on the dynamical generalizations of the prime number theorem, we will establish the dynamical generalizations of these three results related to Chowla's conjecture. In the proofs, we will show a uniform pretentious-distance estimate on the prime Omega function. Then exponential sum estimates are used to establish the averaged shift invariance of the distributions related to the shifted sum of the prime Omega function.

math.NT

One- and two-dimensional cluster states for topological phase simulation and measurement-based quantum computation

Quantum entanglement is a fundamental resource for quantum information processing and serves as a critical benchmark for quantum hardware performance. Cluster states are a special class of entangled states that serve as universal resources for measurement-based quantum computation and possess an intrinsic symmetry-protected topological order, which confers robustness against symmetry-respecting noise. Here we report the scalable preparation and verification of genuine multipartite cluster states on the 105-qubit Zuchongzhi 3.1 superconducting processor. We achieve one-dimensional cluster states of up to 95 qubits and two-dimensional cluster states of up to 72 qubits. The symmetry-protected topological cluster states exhibit input-state-dependent robustness under symmetry-breaking perturbations due to an operational parity structure that enhances the performance of measurement-based quantum computation. Furthermore, we use our two-dimensional cluster states to implement the Deutsch-Jozsa algorithm within the measurement-based quantum computation framework, achieving higher output-state fidelity compared with traditional circuit-based models and a query efficiency advantage over classical approaches. Our work establishes a scalable platform that combines large-scale entanglement generation, symmetry-protected topological order and practical quantum algorithms to enable robust, fault-tolerant measurement-based quantum computation.

quant-ph

Simple critical zeros and distinct zeros of the Riemann zeta-function in short intervals

Recently, on the non-trivial zeros of the Riemann zeta function, it is discovered by Claude and verified by Alpöge and Furman that more than 67.25% of the zeros are simple and on the critical line, and more than 83.62% are distinct. Later, Lamzouri gave a different and more direct proof. In this article, we will use the method of Lamzouri to give lower bounds on the number of the non-trivial zeros of the Riemann zeta function in short intervals. To prove the main result, we establish Montgomery's theorem on the pair correlation of zeros of the zeta function in short intervals by following the approach of Baluyot, Goldston, Suriajaya and Turnage-Butterbaugh, and then use Lamzouri's inequality on any finite multiset of complex numbers which is invariant under complex conjugation.

math.NT

Two averaged dynamical generalizations of Chowla's conjecture

Let $k\ge1$ be an integer and let $λ$ be the Liouville function. In 1965, Chowla gave a conjecture that the values of $λ(n+h_1),\dots, λ(n+h_k)$ are asymptotically unrelated for any distinct natural numbers $h_1, \dots, h_k$. In this article, motivated by the recent work of Bergelson and Richter on the dynamical generalizations of the prime number theorem, we will show a dynamical generalization of Chowla's conjecture on average. In the proof, we follow an approach of Qi and Zheng who established a variant of Bergelson and Richter's theorem over irreducible binary cubic forms. Moreover, we will use this approach to show an analogue of the dynamical Chowla's conjecture along the primes on average. In 2016, Tao proved that the two-point logarithmic Chowla's conjecture holds. Recently, Charamaras and Richter generalized Tao's theorem to bounded arithmetic functions and proposed a conjecture that generalizes Chowla's conjecture to bounded multi-variable arithmetic functions. Building on their work, we prove a dynamical generalization of Tao's theorem and a variant for the composition of the sum-of-digits function with the prime Omega function.

math.NT

Image-Conditional Diffusion Transformer for Underwater Image Enhancement

Underwater image enhancement (UIE) has attracted much attention owing to its importance for underwater operation and marine engineering. Motivated by the recent advance in generative models, we propose a novel UIE method based on image-conditional diffusion transformer (ICDT). Our method takes the degraded underwater image as the conditional input and converts it into latent space where ICDT is applied. ICDT replaces the conventional U-Net backbone in a denoising diffusion probabilistic model (DDPM) with a transformer, and thus inherits favorable properties such as scalability from transformers. Furthermore, we train ICDT with a hybrid loss function involving variances to achieve better log-likelihoods, which meanwhile significantly accelerates the sampling process. We experimentally assess the scalability of ICDTs and compare with prior works in UIE on the Underwater ImageNet dataset. Besides good scaling properties, our largest model, ICDT-XL/2, outperforms all comparison methods, achieving state-of-the-art (SOTA) quality of image enhancement.

cs.CV

A New Type of Adversarial Examples

Most machine learning models are vulnerable to adversarial examples, which poses security concerns on these models. Adversarial examples are crafted by applying subtle but intentionally worst-case modifications to examples from the dataset, leading the model to output a different answer from the original example. In this paper, adversarial examples are formed in an exactly opposite manner, which are significantly different from the original examples but result in the same answer. We propose a novel set of algorithms to produce such adversarial examples, including the negative iterative fast gradient sign method (NI-FGSM) and the negative iterative fast gradient method (NI-FGM), along with their momentum variants: the negative momentum iterative fast gradient sign method (NMI-FGSM) and the negative momentum iterative fast gradient method (NMI-FGM). Adversarial examples constructed by these methods could be used to perform an attack on machine learning systems in certain occasions. Moreover, our results show that the adversarial examples are not merely distributed in the neighbourhood of the examples from the dataset; instead, they are distributed extensively in the sample space.

cs.LG

On the minimum modulus problem in number fields

The minimum modulus problem on covering systems was posed in 1950 by Erdős, who asked whether the minimum modulus of a covering system with distinct moduli is bounded. In 2007, Filaseta, Ford, Konyagin, Pomerance and Yu affirmed it if the reciprocal sum of the moduli of a covering system is bounded. Later in 2015, Hough resolved this problem by showing that the minimum modulus in any covering system with distinct moduli is at most $10^{16}$. In 2022, Balister, Bollobás, Morris, Sahasrabudhe and Tiba reduced this bound to $616,000$ by developing a versatile method called the distortion method. Recently, Klein, Koukoulopoulos and Lemieux generalized Hough's result by using a suitable modification of the distortion method. In this paper, we develop the distortion method further by introducing the theory of probability measures associated to an inverse system. Following Klein et al.'s work, we provide a solution to Erdős' minimum modulus problem in number fields. As an application, we prove that the $j$-th smallest norm in a minimal covering system of a number field with distinct moduli is bounded.

math.NT

Asymptotic uncorrelations between functions with squarefull kernel and functions of invariant average

In 1986, Ivić and Tenenbaum introduced arithmetic functions with squarefull kernel, which are also called $s$-functions. Later, Erdős and Ivić gave an asymptotic estimate on the shifted convolution sums of $s$-functions. Recently, Bergelson and Richter studied the orbits along the prime Omega function in a uniquely ergodic topological dynamical system and established a new dynamical generalization of the prime number theorem (PNT). These orbits can be viewed as functions of invariant average under multiplications. In this paper, we show that both $s$-functions and their shifted convolutions are asymptotically uncorrelated to the orbits along the prime Omega function in a uniquely ergodic system. As a consequence, we obtain a refinement of the PNT via the local distribution of $s$-functions. Furthermore, several variants of these results are established as well.

math.NT

Surface code logical operations on a superconducting quantum processor

Fault-tolerant quantum computation requires logical operations that manipulate encoded information while preserving quantum error-correction protection. In planar surface-code architectures, code deformation and lattice surgery provide a local, measurement-based route to such operations. Here we experimentally realize key elements of patch-based surface-code logical processing on a 107-qubit superconducting quantum processor. We first implement a reusable primitive layer comprising merge and split, patch expansion and shrinkage, and deformations mediated by domain walls and twist defects. We then compose these primitives to realize logical state routing, the logical controlled-NOT gate, and the single-qubit Hadamard and phase gates, which together form a Clifford-generating set. All operations are implemented on distance-three rotated surface-code patches with multi-round syndrome extraction and neural-network decoding, without post-selection. Our results advance superconducting surface-code experiments from protected logical memory to active, patch-based fault-tolerant logical operations.

quant-ph

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: AI Flash Portrait (Track 3)

In this paper, we present a comprehensive overview of the NTIRE 2026 3rd Restore Any Image Model (RAIM) challenge, with a specific focus on Track 3: AI Flash Portrait. Despite significant advancements in deep learning for image restoration, existing models still encounter substantial challenges in real-world low-light portrait scenarios. Specifically, they struggle to achieve an optimal balance among noise suppression, detail preservation, and faithful illumination and color reproduction. To bridge this gap, this challenge aims to establish a novel benchmark for real-world low-light portrait restoration. We comprehensively evaluate the proposed algorithms utilizing a hybrid evaluation system that integrates objective quantitative metrics with rigorous subjective assessment protocols. For this competition, we provide a dataset containing 800 groups of real-captured low-light portrait data. Each group consists of a 1K-resolution low-light input image, a 1K ground truth (GT), and a 1K person mask. This challenge has garnered widespread attention from both academia and industry, attracting over 100 participating teams and receiving more than 3,000 valid submissions. This report details the motivation behind the challenge, the dataset construction process, the evaluation metrics, and the various phases of the competition. The released dataset and baseline code for this track are publicly available from the same \href{https://github.com/zsn1434/AI_Flash-BaseLine/tree/main}{GitHub repository}, and the official challenge webpage is hosted on \href{https://www.codabench.org/competitions/12885/}{CodaBench}.

cs.CV

Mamyshev oscillator based on gain-managed nonlinearity and chirped pulse amplification

We experimentally demonstrate a Mamyshev oscillator based on gain-managed nonlinearity and chirped pulse amplification. Different from other Mamyshev oscillators, the gain-managed nonlinear regime serves as a seed provider instead of a power amplifier in one arm of this laser. The output pulse energy over 300 nJ and pulse width of 739 fs has been achieved from the chirped pulsed amplification in another arm. This configuration provided a new approach to design a high-energy ultrafast laser.

physics.optics

Temperature-activated dislocation avalanches signaling brittle-to-ductile transition in BCC micropillars

We carry out strain-controlled in-situ compression experiments of micron-sized tungsten (W) micropillars in the temperature range 300-900 K, together with simulations of three-dimensional discrete dislocation dynamics (DDD) at the same scale. Two distinct regimes are observed. At low temperatures, plastic deformation appears smooth, both temporally and spatially. Stress fluctuations are consistent with a Wiener stochastic process resulting from uncorrelated dislocation activity within the pillars. However, high-temperature stress fluctuations are highly correlated and exhibit features of self-organized criticality (SOC), with deformation located within well-defined slip bands. The high-temperature stress relaxation statistics are consistent with a thermally activated nucleation process from the surface. The nature of the transition between the two regimes is a manifestation of the brittle to ductile transition in BCC metals.

cond-mat.mtrl-sci

OmniVideo-R1: Reinforcing Audio-visual Reasoning with Query Intention and Modality Attention

While humans perceive the world through diverse modalities that operate synergistically to support a holistic understanding of their surroundings, existing omnivideo models still face substantial challenges on audio-visual understanding tasks. In this paper, we propose OmniVideo-R1, a novel reinforced framework that improves mixed-modality reasoning. OmniVideo-R1 empowers models to "think with omnimodal cues" by two key strategies: (1) query-intensive grounding based on self-supervised learning paradigms; and (2) modality-attentive fusion built upon contrastive learning paradigms. Extensive experiments on multiple benchmarks demonstrate that OmniVideo-R1 consistently outperforms strong baselines, highlighting its effectiveness and robust generalization capabilities.

cs.AI

The prime number theorem over integers of power-free polynomial values

Let $f(x)\in \mathbb{Z}[x]$ be an irreducible polynomial of degree $d\ge 1$. Let $k\ge2$ be an integer. The number of integers $n$ such that $f(n)$ is $k$-free is widely studied in the literature. In principle, one expects that $f(n)$ is $k$-free infinitely often, if $f$ has no fixed $k$-th power divisor. In 2022, Bergelson and Richter established a new dynamical generalization of the prime number theorem (PNT). Inspired by their work, one may expect that this generalization of the PNT also holds over integers of power-free polynomial values. In this note, we establish such variants of Bergelson and Richter's theorem for several polynomials studied by Estermann, Hooley, Heath-Brown, Booker and Browning.

math.NT

On averages of completely multiplicative functions over co-prime integer pairs

Recently, Donoso, Le, Moreira and Sun studied the asymptotic behavior of the averages of completely multiplicative functions over the Gaussian integers. They derived Wirsing's theorem for Gaussian integers, answered a question of Frantzikinakis and Host for sum of two squares, and obtained a variant of a theorem of Bergelson and Richter on ergodic averages along the number of prime factors of integers. In this paper, we will show the analogue of these results for co-prime integer pairs. Moreover, building on Frantzikinakis and Host's results, we obtain some convergences on the multilinear averages of multiplicative functions over primitive lattice points.

math.NT

RISE-T2V: Rephrasing and Injecting Semantics with LLM for Expansive Text-to-Video Generation

Most text-to-video(T2V) diffusion models depend on pre-trained text encoders for semantic alignment, yet they often fail to maintain video quality when provided with concise prompts rather than well-designed ones. The primary issue lies in their limited textual semantics understanding. Moreover, these text encoders cannot rephrase prompts online to better align with user intentions, which limits both the scalability and usability of the models, To address these challenges, we introduce RISE-T2V, which uniquely integrates the processes of prompt rephrasing and semantic feature extraction into a single and seamless step instead of two separate steps. RISE-T2V is universal and can be applied to various pre-trained LLMs and video diffusion models(VDMs), significantly enhancing their capabilities for T2V tasks. We propose an innovative module called the Rephrasing Adapter, enabling diffusion models to utilize text hidden states during the next token prediction of the LLM as a condition for video generation. By employing a Rephrasing Adapter, the video generation model can implicitly rephrase basic prompts into more comprehensive representations that better match the user's intent. Furthermore, we leverage the powerful capabilities of LLMs to enable video generation models to accomplish a broader range of T2V tasks. Extensive experiments demonstrate that RISE-T2V is a versatile framework applicable to different video diffusion model architectures, significantly enhancing the ability of T2V models to generate high-quality videos that align with user intent. Visual results are available on the webpage at https://rise-t2v.github.io.

cs.CV

VC4VG: Optimizing Video Captions for Text-to-Video Generation

Recent advances in text-to-video (T2V) generation highlight the critical role of high-quality video-text pairs in training models capable of producing coherent and instruction-aligned videos. However, strategies for optimizing video captions specifically for T2V training remain underexplored. In this paper, we introduce VC4VG (Video Captioning for Video Generation), a comprehensive caption optimization framework tailored to the needs of T2V models. We begin by analyzing caption content from a T2V perspective, decomposing the essential elements required for video reconstruction into multiple dimensions, and proposing a principled caption design methodology. To support evaluation, we construct VC4VG-Bench, a new benchmark featuring fine-grained, multi-dimensional, and necessity-graded metrics aligned with T2V-specific requirements. Extensive T2V fine-tuning experiments demonstrate a strong correlation between improved caption quality and video generation performance, validating the effectiveness of our approach. We release all benchmark tools and code at https://github.com/alimama-creative/VC4VG to support further research.

cs.CV

Some ergodic theorems over $k$-full numbers

In 2022, Bergelson and Richter established a new dynamical generalization of the prime number theorem. Later, Loyd showed a disjoint form with the Erdős-Kac theorem. Recently, the author and his coauthors proved some ergodic theorems over squarefree numbers related to these results. In this paper, building on the previous work, we will derive the analogues of Bergelson-Richter's theorem, Erdős-Kac theorem and Loyd's theorem over $k$-full numbers for any integer $k\geq2$.

math.NT