arXiv ScienceSearch

arXiv subjects

Fan Xu

Publications and source records attributed to Fan Xu.

At least 19 recordsLinked to original sources

Deformation families of open Calabi-Yau manifolds and Steinness

This paper is to construct deformation families of open Calabi-Yau manifolds together with some discussions on Steinness. Firstly, we construct a nontrivial compactifiable deformation of open Calabi-Yau manifolds. Secondly, we construct a compactifiable deformation family admitting a holomorphic involution and claim the non-Steinness for all the compactifiable deformation families constructed. Finally, we give some discussions on the non-Steinness of the quasi-projective variety from the blow-up of CP2 at nine points corresponding to a conjecture by Macro Brunella and present deformation families whose fibers are Stein.

math.AG

A Missing Tool for Calculating Auto/Cross-correlation Function under Nonuniform Sampling Observations

Nonuniform sampling presents a long-standing challenge in astrophysical time-domain analysis, invalidating the standard autocorrelation and cross-correlation functions and forcing researchers to adopt ad-hoc methods like interpolation or binning, which introduce unquantified biases and lack rigorous error estimation. Here we introduce a new method for calculating the nonuniform autocorrelation function (NUACF) and nonuniform cross-correlation function (NUCCF) for irregularly sampled time series. Instead of relying on interpolation, it naturally evaluates the correlation function by incorporating time-interval weights and misalignment penalties. Monte Carlo simulations provide confidence bands for significance assessment and a complete error budget for the time delays that accounts for both flux uncertainties and sampling irregularity (essential but generally absent from existing methods). Through extensive simulations, we demonstrate our method outperforms traditional methods across various conditions, from strictly periodic to complex repeating variability patterns (e.g., intermittent but aperiodic). Its effectiveness is demonstrated via various real astrophysical data sets, revealing repetitive variability in stellar light curves, measuring time delays for multi-band disc reverberation in the AGN Fairall 9, and providing model-independent validation of time delays for the gravitationally lensed quasar HE 0435-1223. The method provides a rigorous and general solution to the ubiquitous problem of nonuniform sampling, positioning it as a useful tool for large-scale time-domain survey data analysis. The framework is also directly applicable to emerging time-domain phenomena such as fast radio bursts (FRBs), enabling, e.g., the study of correlations between persistent radio source luminosity and repeating FRB activity, or among the multi-parameter variability curves of FRB emission itself.

astro-ph.IM

Constraints on the Low-frequency Radio Emission of the Galactic FRB Source SGR 1935+2154

We present a search for radio pulses from the Galactic magnetar SGR 1935+2154, a well-known source of fast radio bursts (FRBs), at $\sim$110 MHz using the Large Phased Array (LPA) of the Pushchino Radio Astronomy Observatory. Data from two active periods in 2020 (March -- May and September -- November, with $\sim 3.5$ minutes of daily coverage) were analyzed with new methods tailored to both FRB-like single pulses and pulsar-like periodic signals. No significant FRB-like pulses were found. Using Monte Carlo simulations, $3\sigma$ upper limits were derived for the burst rate: for a log-normal energy distribution the limit is $\sim$${10}^{1.5}~{\rm{d}}^{-1}$ for a mean of average monochromatic isotropic luminosity $L_{\nu{\rm ,mean}}\sim1.3\times{10}^{29}~{\rm{erg~s^{-1}~ {Hz}^{-1}}}$ and a natural log-space scatter of $\sigma\sim0.85$; while for a power-law distribution it is $\sim$${10}^{1.8}~{\rm{d}}^{-1}$ for an index $\beta\lesssim3.0$ and a minimum average monochromatic isotropic luminosity $L_{\nu{\rm{,min}}}\lesssim0.7\times{10}^{25}~{\rm{erg~s^{-1}~{Hz}^{-1}}}$. When folded at the known 3.24781628 s period of SGR 1935+2154, a weak pulse was noted (S/N $<$ 3.16), but the significance is insufficient for a secure detection of the pulsar-like emission signal. A conservative upper limit on the average monochromatic isotropic luminosity of any possible periodic emission is $2.08\times{10}^{19}~{\rm{erg~s^{-1}~{Hz}^{-1}}}$. Our results offer meaningful low-frequency upper limits on the burst rate of SGR 1935+2154, and hint for very faint pulsar-like radiation at meter wavelengths.

astro-ph.HE

Training-Free Pseudo-Fusion for Composed Image Retrieval with Diffusion Models and Multimodal Large Language Models

Composed Image Retrieval (CIR) is an emerging paradigm in content-based image retrieval that enables users to formulate compositional queries by combining a reference image with an auxiliary modality, usually text-based. This approach supports fine-grained search where the target image shares structural elements with the user-provided image while incorporating the modifications specified by the auxiliary text. Conventional CIR methods rely on multimodal fusion to combine visual and textual features into a joint query embedding, which requires training modules that align composed queries with the targets. In this work, we propose PeFuse (for pseudo-fusion), a training-free framework that leverages pretrained Diffusion Models and Multimodal Large Language Models to bridge modalities via generative conversion. We introduce two novel strategies: uni-directional and bi-directional conversion, which convert CIR into four single-modality retrieval problems. These methods reformulate CIR as either intra-modal or cross-modal single-query retrieval tasks, bypassing the need for dedicated task-specific training. Extensive experiments on standard benchmarks demonstrate that converting CIR into text-to-image retrieval tasks is more effective than alternative conversion strategies, achieving competitive or superior performance compared with state-of-the-art methods, while maintaining high flexibility thanks to replaceable components of the conversion pipeline. These results highlight the effectiveness of the pseudo-fusion paradigm for zero-shot CIR. Our code is publicly available at: https://github.com/StevenXuf/PeFuse4CIR.

cs.CV

Congruence Decomposition with Neural Block Solvers for Large-Scale PCI Assignment

Physical Cell Identity (PCI) assignment is essential for interference management in dense 5G networks. As cellular networks scale, PCI reuse becomes unavoidable, which may cause collisions, confusions, and multiple forms of modular interference. Jointly mitigating these effects gives rise to a large-scale, multi-objective combinatorial optimization problem that is difficult to solve efficiently at practical network scales. In this work, we propose a congruence decomposition framework with neural block solvers for large-scale PCI assignment. The proposed decomposition exploits the arithmetic structure of PCI values to decouple multiple modular interference objectives into a collection of blockwise Min-$k$-Partition subproblems, followed by a graph coloring procedure to resolve PCI conflicts. For the resulting NP-hard Min-$k$-Partition subproblems, we develop neural block solvers by parameterizing their relaxed quadratic formulations with graph neural networks, enabling efficient optimization at large scales. Discrete assignments are recovered through conditional expectation rounding with theoretical guarantees. Experiments on synthetic cellular graphs and real-world 5G networks show that the proposed method consistently outperforms existing modular-interference-aware baselines in modular interference reduction, conflict elimination, and computational efficiency.

cs.LG

Dispersion Control of Chiral Exciton-Polariton Transport with Dielectric Metasurfaces

Exciton-polaritons provide a powerful platform for manipulating hybrid light-matter states with low effective masses and strong nonlinearities. Introducing chirality into these quasiparticles enables selective control over their spin and propagation, opening new opportunities for chiral transport and spin-selective polaritonic devices. We exploit the strong chiral light-matter coupling in silicon metasurfaces composed of tilted nanorod dimers to demonstrate selective transport of organic chiral exciton-polaritons. The metasurface supports surface lattice resonances and quasi-bound states in the continuum that simultaneously provide high photonic confinement and extrinsic chirality, giving rise to chiral exciton-polaritons in the achiral molecules. These exciton-polaritons exhibit a large magnitude of the dissymmetry factor, reaching a value of 0.93. Using photoluminescence Fourier microscopy and real-space imaging, we show that chiral exciton-polaritons propagate over distances exceeding 50 um without significant degradation of their dissymmetry, with characteristic propagation lengths of approximately 6-13 um. These propagation lengths correspond to an enhancement of 3 orders of magnitude compared to bare excitons. This work constitutes the first demonstration of enhanced and selective chiral transport of organic exciton-polaritons, driven by strong light-matter coupling, in achiral metasurfaces, paving the way for spin-selective polaritonic technologies using simple metasurfaces.

physics.optics

ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

EDA scripting with tool-specific, often undocumented APIs remains a long-tail bottleneck that existing LLMs fail to address. This paper presents ZhuLong, an execution-grounded LLM coding agent for PyAether and SKILL that combines API retrieval, documentation inspection, and sandbox execution via unified MCP tools, augmented by an offline API self-exploration mechanism that infers undocumented API behaviors through counterfactual experimentation. We evaluate ZhuLong on EDA-Eval-PyAether, a benchmark of 158 real-world tasks with assertion-based execution, where the complete system achieves 78.5% Pass@1 in the commercial Empyrean Aether environment, substantially outperforming a pure LLM baseline (23.6%). Ablation studies identify sandbox execution as the dominant performance driver (41.2 pp drop when removed), with the self-exploration mechanism contributing an additional 3.2 pp accuracy gain and a 22.1% reduction in per-task tool calls. On 20 interactive tasks involving unsaved layouts and schematics, ZhuLong achieves 60.0% Pass@1 for PyAether and 50.0% for SKILL.

cs.AI

DAVET: Denoising-Aware Visual Evidence Trajectory Allocation for Diffusion Vision-Language Models

Diffusion vision-language models (dVLMs) iteratively denoise masked responses while conditioning each denoising step on visual evidence, making visual conditioning a substantial recurring inference cost. Unlike autoregressive decoding, diffusion generation repeatedly revisits the entire response as uncertainty evolves. Our analysis reveals that visual evidence demand is strongly step-dependent, motivating adaptive allocation across denoising steps. Existing inference acceleration methods operate through decoding-side strategies or visual token compression via pruning and merging, but do not explicitly treat visual evidence as a resource whose demand evolves across the diffusion process. Therefore, we present Denoising-Aware Visual Evidence Trajectory Allocation (DAVET), a training-free framework that allocates visual evidence according to the evolving generation state. Starting from a phase-conditioned evidence trajectory, the proposed allocation policy uses operation demand to set an evidence reserve whose allocation at each denoising step is modulated by trajectory risk. DAVET realizes the resulting budgets through a hierarchy of evidence views constructed from a single visual encoding, separating when and how much evidence is needed from how the evidence views are constructed. Evaluated on two representative dVLMs, LLaDA-V and LaViDa, across multiple visual-understanding benchmarks, DAVET achieves an average speedup of 1.55$\times$ with an average relative performance drop of 1.86\%, showing that denoising-aware visual evidence allocation can reduce visual conditioning cost while largely preserving generation quality.

cs.CV

Beyond the Mean: Multi-Moment Policy Optimization for LLM Reasoning

Reinforcement learning has become a central paradigm for improving the reasoning capabilities of large language models. Existing methods generally aim to reduce the failure probabilities induced across problems. In this paper, we introduce a moment-based perspective on policy optimization for LLM reasoning by treating the failure probability of a randomly sampled problem as a random variable and characterizing optimization objectives through its moments. Under this perspective, many existing methods optimize only a single moment of the failure-probability distribution, leaving its broader distributional structure largely uncharacterized. We propose \textbf{M}ulti-\textbf{M}oment \textbf{P}olicy \textbf{O}ptimization (MMPO), a novel policy optimization framework that jointly minimizes multiple moments of the failure-probability distribution. MMPO admits a direct operational interpretation as minimizing the expected truncated time required to obtain the first successful response. Beyond MMPO, we further develop a general moment-transformation framework that systematically induces different moment profiles and provides a unified view of a broader family of policy optimization objectives. Experiments across five mathematical reasoning benchmarks and models of different scales demonstrate that MMPO consistently outperforms strong baselines. We hope this moment-based perspective offers new insights into the design of policy optimization objectives for LLM reasoning.

cs.AI

Diverse Morphologies of GRB X-Ray Plateaus within a Common Magnetar Framework

The origin of the X-ray plateau phase in gamma-ray bursts (GRBs) remains an open problem. In particular, it is unclear whether GRBs with different temporal morphologies (i.e., with a rising, flat, or decaying plateau) arise from a common underlying mechanism. Although magnetar energy injection is a leading explanation, previous studies have primarily inferred magnetar properties on a burst-by-burst basis and have not tested the model at the population level. Here we perform the first hierarchical population inference of magnetar parameters for a uniform sample of 185 long GRBs with X-ray plateaus within a conditional Poisson point-process framework. It is found that the observed plateau population is well reproduced by physically plausible magnetar populations. The inferred parameter distributions show no strong statistical separation among subclasses with different plateau morphologies. Nevertheless, all subclasses show a substantial intrinsic luminosity scatter, $\sigma_{L,\rm int}\sim0.5$--1.0 dex, whereas the intrinsic duration scatter remains considerably smaller. The results provide a population-level test of the magnetar interpretation of GRB X-ray plateaus, showing that the observed diversity of plateau morphologies does not require distinct magnetar populations.

astro-ph.HE

UPAIR: Diagnosing Reasoning States via Uncertainty-Progress Alignment for Selective Intervention

While test-time scaling improves the problem-solving ability of large reasoning models (LRMs) through additional inference-time computation, it can also exacerbate overthinking and underthinking, which we formulate as reasoning state--action mismatch. Resolving this mismatch requires reliable reasoning state diagnosis, yet single-signal monitors provide ambiguous evidence, while steering-based controllers often rely on outcome-labeled supervision or model-specific calibration. We introduce the Uncertainty--Progress Alignment Hypothesis, which posits that the relative transition timing of proxy answer uncertainty and latent reasoning progress distinguishes healthy, stagnant, and ready states that warrant different subsequent actions. Building on this insight, we propose UPAIR, a training-free framework that couples lightweight uncertainty monitoring with event-triggered joint diagnosis and maps the resulting state to native continuation, selective strategy switching, or verification-guided stopping. Across three LRMs and five cross-domain benchmarks, the stagnation diagnosis detects 64.3% of natural errors while flagging only 5.4% of correct samples, revealing a dynamic reasoning regularity shared across models and tasks. End to end, UPAIR improves accuracy by up to 16.67 percentage points and reduces generated tokens by up to 29.64%, demonstrating the effectiveness of its integrated diagnosis and intervention, while online diagnosis costs less than 1% of natural-generation time.

cs.AI

Information Gain-based Rollout Policy Optimization: An Adaptive Tree-Structured Rollout Approach for Multi-Turn LLM Agents

Reinforcement learning has become a promising paradigm for improving large language model (LLM) agents on long-horizon search tasks, where the agent must make a sequence of intermediate decisions before receiving a final outcome. However, existing methods still face a key limitation: the rollout budget is often allocated without explicitly assessing the utility of intermediate states. As a result, substantial computation may be spent on low-value states, even though different branches can vary drastically in their informativeness. In this paper, we propose Information Gain-based Rollout Policy Optimization (IGRPO), a policy optimization framework that treats intermediate-state informativeness as the organizing principle of rollout collection. Specifically, IGRPO performs budget-aware tree-structured rollouts by allocating expansion budget according to node-level informativeness, so that more informative branches are expanded more frequently while unpromising branches are progressively suppressed. We further demonstrate that the information gain-based rollout induces an explicit limiting teacher distribution over trajectories, which naturally yields a clear policy optimization target, thereby unifying adaptive tree-structured exploration with principled policy learning under a single framework. Experiments on seven challenging search-augmented QA benchmarks demonstrate that IGRPO consistently outperforms strong baselines under the same rollout budget constraints, validating the effectiveness of leveraging the induced teacher distribution to guide policy optimization for long-horizon search agents.

cs.AI

Speech-Driven End-to-End Language Discrimination towards Chinese Dialects

Language discrimination among similar languages, varieties, and dialects is a challenging natural language processing task. The traditional text-driven focus leads to poor results. In this paper, we explore the effectiveness of speech-driven features towards language discrimination among Chinese dialects. First, we systematically explore the appropriateness of speech-driven MFCC features towards CNN-based language discrimination. Then, we design an end-to-end speech recognition model based on HMM-DNN to predict Chinese dialect words. We adopt attention to extract the discriminative words related to different Chinese dialects. Finally, through a CNN, we combine the word-level embedding and the MFCC-based features. Evaluation of two benchmark Chinese dialect corpora shows the appropriateness and effectiveness of the proposed speech-driven approach to fine-grained Chinese dialect discrimination compared to the state-of-the-art methods.

cs.CL

Low-resource Language Discrimination Towards Chinese Dialects with Transfer learning and Data Augmentation

Chinese dialects discrimination is a challenging natural language processing task due to scarce annotation resource. In this article, we develop a novel Chinese dialects discrimination framework with transfer learning and data augmentation (CDDTLDA) in order to overcome the shortage of resources. To be more specific, we first use a relatively larger Chinese dialects corpus to train a source-side automatic speech recognition (ASR) model. Then, we adopt a simple but effective data augmentation method (i.e., speed, pitch, and noise disturbance) to augment the target-side low-resource Chinese dialects, and fine-tune another target ASR model based on the previous source-side ASR model. Meanwhile, the potential common semantic features between source-side and target-side ASR models can be captured by using self-attention mechanism. Finally, we extract the hidden semantic representation in the target ASR model to conduct Chinese dialects discrimination. Our extensive experimental results demonstrate that our model significantly outperforms state-of-the-art methods on two benchmark Chinese dialects corpora.

cs.CL

Quantum Chip Paradigm Framework

Quantum Electronic Design Automation (Q-EDA) is emerging as quantum chips move from laboratory prototypes to scalable engineering systems. This paper argues that superconducting quantum chip design is approaching a "SPICE moment" similar to early classical EDA, where growing qubit scale, control complexity, frequency planning, packaging, process variation, and cryogenic measurement feedback require a shift from experience-based design to model-driven engineering. We propose a Quantum Chip Paradigm Framework that treats Q-EDA not only as software, but as part of the quantum chip development paradigm. Unlike classical HDL-first design, quantum chip design must begin with physical structures such as Josephson junctions, resonators, couplers, readout elements, control lines, and packaging environments. The framework emphasizes PCell-based modeling, SPICE-Q simulation, Quantum PDKs, and design-technology-measurement co-optimization. We further outline a hierarchical Q-EDA system spanning physical structures, qubit PCells, logical qubits, quantum arithmetic, functional quantum IP, and Quantum SoC systems. The key goal is to turn physical models, layout rules, simulation results, fabrication data, and measurement feedback into reusable and auditable engineering objects for large-scale quantum processors and fault-tolerant quantum computing.

quant-ph

Bounded Deep Unfolding for Joint Beamforming and Scheduling in Multi-Cell MIMO Networks

This paper investigates the joint resource block group (RBG) scheduling and beamforming optimization problem for weighted sum-rate (WSR) maximization in multi-cell multiuser multiple-input multiple-output (MU-MIMO) downlink networks. While the Fast Fractional Programming (FastFP) framework provides a reliable model-driven solution, it suffers from conservative continuous beamforming updates and prohibitive computational overhead during the discrete RBG matching phase. To address these bottlenecks, we propose a joint deep unfolding framework comprising two core modules: P-Net and K-Net. Specifically, P-Net learns an adaptive relaxation factor along the FastFP direction, strictly bounded within an ascent-preserving interval to accelerate convergence while retaining stationary-point guarantees. Meanwhile, K-Net learns a long-horizon priority policy to guide a low-complexity greedy assignment, maintaining high assignment quality while bypassing the computationally expensive Hungarian matching. Both networks leverage analytical algorithmic priors and utilize recurrent parameter sharing, enabling flexible inference beyond the training horizon. Extensive simulations demonstrate that the proposed joint framework achieves higher WSR and faster execution times than conventional model-driven baselines, while generalizing robustly across unseen network scales and channel conditions without retraining.

cs.IT

Baseband-Efficient WMMSE Precoding: From a Signal Weighting Cost Perspective

For downlink transmission in massive multi-user multiple-input multiple-output (MU-MIMO) systems, conventional precoding research heavily focuses on reducing the computational complexity of precoding matrix design, while largely overlooking another critical bottleneck: the substantial signal weighting cost incurred by repeatedly applying the precoder to high-speed data streams. To address both challenges simultaneously, this paper proposes a novel sparse precoding framework tailored for fully-digital architectures. Within this framework, from the sum-rate maximization perspective, we design two sparse precoding architectures: a common-support row-sparse architecture and a user-specific row-sparse architecture, so as to reduce the number of multiplication operations required in baseband signal weighting without sacrificing system capacity. For the formulated mixed-integer non-linear programming (MINLP) problem, we rigorously prove, for the first time, that the optimal precoder under both sparse architectures strictly resides in a specific low-dimensional subspace determined by the channel matrices, thereby reducing the dimensionality of the optimization variables. Based on this insight, an alternating optimization algorithm is developed within the weighted minimum mean square error (WMMSE) framework to jointly optimize sparse beam selection and low-dimensional precoding coefficients. The combinatorial beam selection problem is handled using an efficient penalty-based majorize-minimization (MM) method, yielding a low-complexity closed-form solution. Simulation results demonstrate that the proposed scheme achieves near-optimal sum-rate performance while substantially reducing both the precoding computation complexity and the overall signal weighting cost.

eess.SP

PnP-Corrector: A Universal Correction Framework for Coupled Spatiotemporal Forecasting

Coupled spatiotemporal forecasting is important for predicting the future evolution of multiple interacting dynamical systems, such as in climate models. However, existing methods are severely constrained by the persistent bottleneck of compounding errors. In coupled systems, errors from each subsystem simulator propagate and amplify one another, a phenomenon we term Reciprocal Error Amplification, leading to a rapid collapse of long-range predictions. To address this challenge, we propose a universal framework called PnP-Corrector (Plug-and-Play Corrector). The core idea of our framework is to decouple the physical simulation from the error correction process: it freezes pre-trained physics simulation engines and exclusively trains a correction agent to proactively counteract the systematic biases emerging from the coupled system. Furthermore, we design an efficient predictive model architecture, DSLCast, to serve as the backbone of this framework. Extensive experiments demonstrate that our method significantly enhances the long-term stability and accuracy of coupled forecasting systems. For instance, in the challenging task of a 300-day global ocean-atmosphere coupled forecast, our PnP-Corrector framework reduces the prediction error of the baseline model by 28% and surpasses state-of-the-art models on several key metrics.

cs.AI