arXiv ScienceSearch

arXiv subjects

Huan Yang

Publications and source records attributed to Huan Yang.

At least 19 recordsLinked to original sources

From the Test-Mass Limit to Binary Black-Hole Waveforms in Higher-Derivative Gravity

Many higher-derivative theories predict stronger deviations from General Relativity for lower-mass black holes, while their nonlinear field equations often prevent reliable simulations of the full binary evolution. Here we develop a route from controlled black hole perturbation theory based on the modified Teukolsky formalism to comparable-mass waveforms, using parity-even cubic gravity as a representative example. We find that the tidal response of the secondary black hole enters at the same perturbative order as the direct higher-curvature correction and is therefore essential for a consistent leading-order waveform. The resulting strong-field fluxes and conservative dynamics produce an accumulated inspiral dephasing that grows toward merger. Embedding this test-mass information into an effective-one-body model, we construct inspiral-merger-ringdown waveforms for comparable-mass binaries and find coupling-dependent dephasing and waveform-peak shifts. Our results demonstrate how strong-field test-mass calculations can anchor waveform models for higher-derivative gravity when theory-specific numerical-relativity simulations are unavailable.

gr-qc

Imaging Stars at the Quantum Compatibility Limit

Imaging astrophysical sources with a multi-station interferometer is intrinsically a multiparameter quantum-estimation problem. {Using tools from multiparameter quantum metrology,} we show that time-resolved repetitive or adaptive measurements in an \(N\)-station array suffer a fundamental array-level incompatibility among visibility estimators. Collective measurements, {which coherently process the received starlight across multiple time bins in a single joint readout}, remove the array-size penalty up to an order-unity factor, yielding an asymptotic \(O(\sqrt{N})\) enhancement for the {directional-averaged} SNR of visibility measurement. We then propose a memory-assisted interferometric architecture designed to implement collective readout through coherent storage and joint quantum processing. Imaging simulations and Fisher-information analyses demonstrate that collective measurements improve image reconstruction in near-term arrays and enhance the resolving power of future long-baseline architectures, with pronounced benefits for representative AGN targets such as NGC~4151 and 3C~273. These results highlight collective measurement as a promising building block for future quantum-assisted interferometric arrays for stellar imaging.

quant-ph

Gravitational Waves from Green's Function Decomposition for a Kerr black hole: I. Equatorial ISCO Plunge

We present a decomposition of the Kerr Green's function in the time domain, motivated by the frequency-domain split previously studied in the Schwarzschild limit. We show that the identification of a quasinormal-mode contribution, a direct part, and a late-time tail is still available, where the split times are determined by the black hole spin and positions of the emitter and receiver. We have checked this Green's function with time-domain Teukolsky numerical simulations and find excellent agreement. We also apply this decomposed Green's function in the time domain to a model problem with a test particle plunging into a Kerr black hole. The dynamically excited direct wave and quasinormal modes are obtained by convoluting the Green's function with the particle's source term, which may be viewed as the first order in mass ratio of a spinning black hole ringdown.

gr-qc

SpatialDiff: 3D-Aware Object Movement via Implicit Spatial Modeling

Recent advances in image editing allow impressive manipulation of objects, existing methods still struggle to handle spatial movement in complex scenes, such as objects span different depth layers or are partially occluded. Most image editing methods focus solely on prior information from 2D datasets, emphasizing planar features while lacking support for spatial structures. Even approaches that incorporate explicit positional information fail to capture true 3D spatial relationships, thus limiting accurate object movement in complex scenes. In this paper, we present SpatialDiff, a method that effectively captures 3D spatial structures, enabling precise and consistent object movements in complex scenes. Our core innovations are twofold: (1) Implicit 3D Spatial Modeling, which introduces 3D prior knowledge and enables the model to internally build a comprehensive understanding of the three-dimensional spatial structure; and (2) Global Spatial Supervision, which constrains the latent spatial features to enable the model to perceive changes in object spatial positions caused by editing operations. Experimental results demonstrate that our method significantly improves the accuracy and fidelity of spatial movement in complex scenes.

cs.CV

Probing Intrinsic Ellipticity in Compact Star Binaries

We present a novel resonance mechanism that can occur in {compact-star} binaries: a spin-orbit resonance. This resonance locks the binary into a unique state where {the spin of one component} evolves alongside the orbit. The resonance requires this component to possess a finite ellipticity $ε$, and we find that the locking probability is proportional to $\sqrtε$. We show that resonance locking and its subsequent breaking produce a characteristic phase signature in the gravitational waveform, opening a new observational channel for probing intrinsic ellipticity in compact-star binaries, including exotic compact objects. In addition, as an illustrative astrophysical scenario, we discuss magnetars, whose strong internal fields can source the required ellipticity and may place the signal in the ground-based low-frequency band, although their abundance at merger remains uncertain. We have also conducted a search in all neutron star binaries up to the O4a gravitational-wave catalog, with no positive event found so far.

astro-ph.HE

Black hole spectroscopy: from theory to experiment

The "ringdown" radiation emitted by oscillating black holes has great scientific potential. By carefully predicting the frequencies and amplitudes of black hole quasinormal modes and comparing them with gravitational-wave data from compact binary mergers we can advance our understanding of the two-body problem in general relativity, verify the predictions of the theory in the regime of strong and dynamical gravitational fields, and search for physics beyond the Standard Model or new gravitational degrees of freedom. We summarize the state of the art in our understanding of black hole quasinormal modes in general relativity and modified gravity, their excitation, and the modeling of ringdown waveforms. We also review the status of LIGO-Virgo-KAGRA ringdown observations, data analysis techniques, and the bright prospects of the field in the era of LISA and next-generation ground-based gravitational-wave detectors.

gr-qc

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning

Reinforcement learning (RL) post-training improves the reward alignment of flow-based generators, but often degrades perceptual quality in ways that are not captured by the reward proxy. We identify a simple structural signature of this drift: across three post-training methods (NFT, AWM, DPO), RL fine-tuning inflates the per-step velocity norm $\|v_θ\|$ by $5\%$ to $15\%$ relative to the reference. A form of norm inflation has been studied in classifier-free guidance (CFG), where rescaling the velocity back to a reference norm at inference time can mitigate the resulting artifacts. However, this inference-time correction does not transfer cleanly to RL: rescaling $v_θ$ to match $\|v_{\text{ref}}\|$ at inference time neither improves reward nor fixes the quality degradation, because the inflation is co-adapted into the model weights. Furthermore, an adjoint sensitivity analysis shows that velocity magnitude rescaling carries no coherent first-order reward signal at the batch level, indicating that suppressing norm inflation is unlikely to remove a consistently reward-carrying component. Since inference-time renormalization fails while norm suppression carries no reward cost, training-time intervention is the appropriate strategy. Together, these findings motivate NormGuard, a hinge penalty that activates only when $\|v_θ\|$ exceeds $\|v_{\text{ref}}\|$ and composes additively with any velocity-local base loss. Across two base models, three post-training methods, and two reward proxies, NormGuard consistently improves MLLM-judged image quality and forensic realism while preserving reward, with gains that amplify under few-step inference and are not explained by early stopping.

cs.LG

Identifying Kilonovae in the Presence of Optical Afterglow for the Wide-Field Survey Telescope

Identifying kilonovae associated with binary neutron star mergers is often complicated by the presence of a dominant synchrotron afterglow. In this work, we evaluate the performance of the Wide-Field Survey Telescope (WFST) in identifying kilonova signals in composite afterglow-kilonova transients. Using a numerical framework based on the Fisher information matrix, we simulate $10,000$ realizations for each of two scenarios: an AT2017gfo-based template model and a physically sampled population that accounts for kilonova diversity. Our results indicate that kilonova identification is primarily limited by source distance. In both scenarios, the identification efficiency is largely insensitive to variations in afterglow microphysical parameters and exceeds $80\%$ at distances within approximately $600~\rm Mpc$ for AT2017gfo-like events. Under our adopted assumptions, we estimate that WFST could identify approximately $1$--$16$ kilonovae per year. Furthermore, we find that the discriminating power of color-based filters rapidly saturates, reaching a stable plateau by the second night after the merger. We therefore propose a staged observing strategy that prioritizes high-cadence $g$ and $r$-band monitoring during the first night and incorporates the $z$ band from the second night onward. This strategy improves the identification precision by exploiting the increasingly prominent red excess produced by the kilonova. Our results provide a physical basis for optimizing WFST observing resources to efficiently detect and characterize kilonovae in the multimessenger era.

astro-ph.HE

Torsional-X Seismometer for Lunar Decihertz Gravitational-Wave Detection

The lunar gravitational-wave antenna concept uses the Moon as a resonant detector instrumented with precision seismometers, targeting the decihertz band between ground- and space-based observatories. We propose a compact monolithic fused-silica torsional-X seismometer that re-engineers garden-gate acceleration-to-rotation transduction for this regime through a high-tension dual-fiber suspension. Its designed millihertz-scale resonance and ultra-low mechanical dissipation enable a nearly order-of-magnitude improvement around $0.1\,\mathrm{Hz}$ compared with existing lunar seismometer concepts. Achieving this performance requires room-temperature operation, where fused-silica exhibits low mechanical loss, together with subdominant actuation noise. We demonstrate a room-temperature vacuum prototype validating the operating principle and core mechanical design, and derive requirements for a future lunar implementation capable of approaching the target sensitivity.

astro-ph.IM

Making Image Editing Easier via Adaptive Task Reformulation with Agentic Executions

Instruction guided image editing has advanced substantially with recent generative models, yet it still fails to produce reliable results across many seemingly simple cases. We observe that a large portion of these failures stem not from insufficient model capacity, but from poorly formulated editing tasks, such as those involving small targets, implicit spatial relations, or under-specified instructions. In this work, we frame image editing failures as a task formulation problem and propose an adaptive task reformulation framework that improves editing performance without modifying the underlying model. Our key idea is to transform the original image-instruction pair into a sequence of operations that are dynamically determined and executed by a MLLM agent through analysis, routing, reformulation, and feedback-driven refinement. Experiments on multiple benchmarks, including ImgEdit, PICA, and RePlan, across diverse editing backbones such as Qwen Image Edit and Nano Banana, show consistent improvements, with especially large gains on challenging cases. These results suggest that task reformulation is a critical but underexplored factor, and that substantial gains can be achieved by better matching editing tasks to the effective operating regime of existing models.

cs.CV

A UAV-Based Multi-Modal Vision System for Automated Sideslope Deformation Monitoring and Hazard Detection

Slope hazards constitute a major safety threat to expressway infrastructure, and their evolution is typically manifested as slow surface deformation. Conventional manual inspection suffers from low efficiency and inadequate operational safety, especially on severely deteriorated slopes. Accordingly, there is an urgent need for an automated, high-precision solution capable of large-area slope observation and analysis. This study aims to develop a highly automated workflow for slope hazard detection using Unmanned Aerial Vehicle (UAV)-borne Light Detection and Ranging (LiDAR). The proposed workflow consists of a shared data-acquisition and ground-surface extraction stage, a single-observation hazard-screening branch based on RandLA-Net, and a multi-epoch deformation-monitoring branch based on grid-wise elevation differencing. To validate the effectiveness of the proposed system, we conducted multiple UAV-borne LiDAR data-acquisition flights in real expressway slope environments. The results show that the workflow can extract usable ground-surface point clouds under vegetation cover, identify potential hazard zones from single-observation point clouds, and quantify centimeter-level elevation changes using multi-epoch grid differencing. This study establishes an end-to-end UAV-borne LiDAR-based workflow for slope inspection and demonstrates its feasibility through controlled experiments, field tests, and simulation-based validation, thereby providing an implementable solution for automated slope-hazard monitoring and intelligent early warning.

cs.CV

Towards Efficient and Secure Cloud-Assisted Autonomous Systems: A Review of Architectures, Algorithms, Security, and Deployment Challenges

Networked Control Systems (NCSs) have been instrumental in realizing fully connected and responsive intelligent environments within the context of real-time virtual control and management. However, traditional NCSs face considerable challenges in handling the vast amounts of data generated by large-scale control applications, particularly in terms of data acquisition, storage, and computational processing. To address these challenges, the emergence of cloud computing and advancements in control theory have empowered the new paradigm known as Cloud Control Systems (CCSs). Recently, CCSs have received substantial attention from industries for their potential properties, such as large-scale data management, complex computations, and data-centric optimized decisions. This study presents an extensive review of recent progress in CCSs spanning over multiple studies published between 2012 and 2025. Specifically, the focus is on providing a taxonomy of the current findings in CCS research, encompassing various perspectives, such as its efficient implementations in industrial automation, security and privacy considerations, and cloud-based control techniques. Each category is examined in depth through selected state-of-the-art analyses of different approaches and contrasting methodologies. Furthermore, we discuss future directions aimed at designing more efficient and practical CCSs. The insights gained from this study can help researchers, practitioners, and decision-makers in their domain for effective CCS design and deployment.

eess.SY

Modified Teukolsky Formalism for Extreme Mass-Ratio Inspirals in Higher-Derivative Gravity

In this work, we study a model problem involving a point particle spiraling into a non-rotating black hole in higher-derivative theories of gravity. In such theories, both the background spacetime and the generation and propagation of gravitational waves differ from those in General Relativity. We develop a modified Teukolsky formalism to describe gravitational waves sourced by the point particle and, as an illustrative example, compute the resulting fluxes to the black hole horizon and null infinity for a cubic gravity theory. The formalism is constructed in a way that can be naturally extended to rotating black holes. These results represent essential steps to build extreme mass-ratio-inspiral waveforms in modified gravity theories, which may also be rescaled to approximate waveforms from comparable-mass binary black hole systems, analogous to existing approaches in General Relativity.

gr-qc

MaskAlign: Token-Subset Representation Alignment for Efficient Diffusion Training

Representation alignment with pretrained vision models has recently shown strong potential for accelerating diffusion transformer training. By aligning intermediate diffusion features with clean-image representations from self-supervised vision encoders, existing methods improve convergence and generation quality. However, such alignment also introduces a non-trivial constraint: diffusion models operate on noisy inputs whose usable information varies across timesteps, while the reference features are extracted from clean images. In this paper, we revisit this mismatch from a token-level perspective. We find that, under full-token representation alignment, tokens with large alignment-gradient norms exhibit a stable spatial preference, suggesting that the alignment objective does not affect all tokens uniformly and may encourage the model to rely on the complete set of clean-image tokens. To address this issue, we propose MaskAlign, a token-subset representation alignment method that applies alignment to randomly sampled token subsets during training. By exposing the model to different token subsets across iterations, MaskAlign reduces the dependence of representation alignment on the complete token set and encourages alignment behavior that is more stable under token-subset perturbations. To mitigate the information loss caused by directly dropping tokens, we further introduce a lightweight pre-mask token mixing block that shares information across tokens before masking.

cs.CV

Spin Precession Signatures as an Indicator of Microlensing in Strongly Lensed Gravitational Waves

Microlensing by the stellar field in a strong-lensing galaxy can introduce wave-optics distortions into the waveforms of strongly lensed gravitational waves (SLGWs). If these signals are analyzed with waveform templates that do not include microlensing, the lensing-induced modulation may be misinterpreted as intrinsic source physics. In particular, microlensing can mimic spin precession, since both effects can produce beat-pattern-like features in the waveform. In this work, we study the degeneracy between stellar-field microlensing and spin precession, and ask to what extent microlensed SLGWs may show false evidence of precession. We analyze simulated SLGW events for two detector sensitivities, O5 and a lower-noise configuration with a power spectral density reduced by a factor of 4 (named O5 Plus), assuming binary black holes with parallel spins. We find that microlensing can indeed produce apparent evidence for precession, and that this effect becomes more visible at higher signal-to-noise ratios. Under O5 sensitivity, 4.88% of microlensed events lie above the one-sided Gaussian-equivalent 3$σ$ background threshold, corresponding to the 99.9th percentile of the unlensed-background distribution, while under O5 Plus sensitivity this fraction increases to 14.91%. We also find that the evidence for precession is positively correlated with the strength of microlensing. This correlation is weak under O5 sensitivity, but becomes clear under O5 Plus sensitivity. In addition, Type II (saddle-point) images show a stronger correlation than Type I (minimum-point) images. These results show that evidence for precession in GW data should be interpreted with care, as it may also arise from microlensing wave effects in SLGWs.

astro-ph.CO

Artificial Intelligence Driven Channel Coding and Resource Optimization for Wireless Networks: A Systematic Survey

The ongoing evolution of 5G and its enhanced version, 5G+, has significantly transformed the telecommunications landscape, driving an unprecedented demand for ultra-high-speed data transmission, ultra-low latency, and resilient connectivity. These capabilities are essential for enabling mission-critical applications such as the Internet of Things, autonomous vehicles, and smart city infrastructures. This survey investigates the important role of Artificial Intelligence (AI) in addressing the key challenges faced by 5G/5G+ networks, including interference mitigation, dynamic resource allocation, and maintaining seamless network operation. The study particularly focuses on AI-driven innovations in coding theory, which offer advanced solutions to the limitations of conventional error correction and modulation techniques. By employing deep learning, reinforcement learning, and neural network-based approaches, including convolutional neural networks, recurrent neural networks, and Transformer-based models, this research demonstrates significant advancements in error correction performance, decoding efficiency, and adaptive transmission strategies. Additionally, the integration of AI with emerging technologies, such as massive multiple-input and multiple-output, intelligent reflecting surfaces, and privacy-enhancing mechanisms, is discussed, highlighting their potential to propel the next generation of wireless networks. This survey provides an insightful overview of the transformative impact of AI on modern wireless communication, establishing a foundation for scalable, adaptive, and more efficient network architectures.

eess.SP

Fundamental Physics and Cosmology with TianQin

The exploration of the surrounding world and the universe is an important theme in the legacy of humankind. The detection of gravitational waves is adding a new dimension to this grand effort. What are the fundamental physical laws governing the dynamics of the universe? What is the fundamental composition of the universe? How has the universe evolved in the past and how will it evolve in the future? These are the basic questions that press for answers. The space-based gravitational wave detector TianQin will tune in to gravitational waves in the millihertz frequency range ($10^{-4} \sim 1$ Hz, to be specific), opening a new gravitational wave spectrum window to explore many of the previously hidden sectors of the universe. TianQin will discover many astrophysical systems, populating the universe at different redshifts: some will be of new types that have never been detected before, some will have very high signal-to-noise ratios, and some will have very high parameter estimation precision. The plethora of information collected will bring us to new fronts on which to search for the breaking points of general relativity, the possible violation of established physical laws, the signature of possible new gravitational physics and new fundamental fields, and to improve our knowledge on the expansion history of the universe. In this white paper, we highlight the advances that TianQin can bring to fundamental physics and cosmology.

gr-qc

UniPPTBench: A Unified Benchmark for Presentation Generation Across Diverse Input Settings

Existing works typically focus on presentation generation under isolated input settings, whereas real-world use cases span diverse scenarios, including vague user prompts, long documents, multimodal materials, and multiple heterogeneous sources. Moreover, current evaluations are often insufficiently scenario-specific. They mainly rely on generic presentation-quality criteria, such as visual appeal, layout quality, and overall coherence, but fail to assess the core capabilities required by different input settings, including grounded compression, visual-text alignment, and cross-source synthesis. Consequently, the field lacks a unified benchmark and a scenario-aware evaluation framework for faithfully diagnosing presentation-generation systems across diverse real-world settings. We present UniPPTBench, a unified benchmark for presentation generation across four representative input settings: vague-prompt, long-document, multimodal-document, and multi-source generation. We further introduce UniPPTEval, a scenario-aware evaluation protocol that combines shared metrics for cross-setting comparison with scenario-specific metrics tailored to the core requirements of each setting. We also provide transparent reference baselines to support reproducible comparison. Experiments on UniPPTBench reveal substantial performance variation across settings and recurring failure modes in content grounding, multimodal integration, and cross-source synthesis. In particular, strong performance on generic presentation-quality metrics does not necessarily imply strong task fulfillment in grounded scenarios. Together, UniPPTBench and UniPPTEval provide a faithful and diagnostic foundation for evaluating presentation generation across diverse real-world scenarios. Code and data will be publicly available.

cs.CV