arXiv ScienceSearch

arXiv subjects

Heng Guo

Publications and source records attributed to Heng Guo.

At least 19 recordsLinked to original sources

Complex scalar field thick branes: stability of linear perturbation and evolution of scalar Kaluza-Klein modes coupled with gravity

We investigate the stability and localization properties of Minkowski, de Sitter (dS), and anti-de Sitter (AdS) thick branes generated by a complex scalar field $Φ$. The scalar self-interaction potential contains a temperature parameter $a$. As $a$ approaches its critical value, the imaginary part of $Φ$ becomes double-kinked and the energy density splits into two peaks, indicating the formation of two sub-branes. The analysis of linear scalar, vector, and tensor perturbations shows no signs of instability in the scalar and vector sectors, while the factorization of the tensor perturbation equation excludes tachyonic tensor modes. The graviton zero mode is localized on the Minkowski and dS branes but is not normalizable on the AdS brane. We further study a real test scalar field $Ψ$ with a coupling function $F(R)$ between its kinetic term and the spacetime curvature. The scalar zero mode can always be localized on the Minkowski and dS branes, with a continuous gapless spectrum of scalar Kaluza--Klein (KK) modes. Massive scalar resonances emerge near the critical value of $a$ for suitable values of $β$. Increasing $β$ produces more resonances and prolongs the lifetimes of the lower resonances, while their time evolution exhibits the characteristic decay of metastable modes. For the AdS brane, the scalar zero mode is also localized, whereas the divergent effective potential at the boundaries traps the massive scalar KK modes as bound states, yielding a discrete mass spectrum.

hep-th

Optimal Simulated Annealing for Partition Function Estimation

In this note, we give a simple analysis of a non-adaptive simulated annealing algorithm for estimating the partition function of Gibbs distributions. This yields the most efficient reduction of this kind so far. We also establish lower bounds for both general and non-adaptive algorithms, showing that our algorithm is optimal over a broad range of parameters.

cs.DS

Fast FPRAS for the Permanent

We give an FPRAS for the permanent of an $n\times n$ $0/1$ matrix with running time $\widetilde{O}(n^{3.5}\varepsilon^{-2})$. Our algorithm extends to a strongly polynomial FPRAS for arbitrary nonnegative matrices, as in previous works. Jerrum, Sinclair, and Vigoda (2004) gave the first FPRAS for the permanent of a nonnegative matrix. The running time was subsequently improved to $\widetilde{O}(n^7)$ by Bezáková, Štefankovič, Vazirani, and Vigoda (2008), and recently to $\widetilde{O}(n^6)$ by Chen, Vigoda, and Yang (2026). We introduce a multicommodity-flow bound inspired by electrical flows, replacing the usual path-length factor by routing energy. For a boosted version of the classical JSV chain, we prove a relaxation-time bound of $O(n^3\log n)$ and show that stationary trajectories of this length estimate all stationary hole-pattern probabilities, yielding an $\widetilde O(n^5)$-time FPRAS algorithm. Our new hole-weighted slide (HWS) chain improves both bounds to $O(n^2\log n)$, yielding an $\widetilde O(n^4)$-time algorithm. Finally, we obtain the claimed $\widetilde O(n^{3.5})$ running time by using a subset of $\widetilde{O}(\sqrt{n})$ checkpoint temperatures in an iterated sequence of warm-starts to obtain initializations at every temperature.

cs.DS

Reduction of the six-dimensional $q$-form fields to the four-dimensional fields by coupling with gravity

In this paper, we investigate the localization of various $q$-form fields on a codimension-two brane. In particular, the $0$-form scalar field, the $1$-form $U(1)$ gauge vector field, and the $2$-form Kalb-Ramond field are considered with gravitational coupling, where a coupling function $F(R)$ is introduced into the six-dimensional actions of these fields. The function $F(R)$ depends on the scalar curvature of the bulk. Within this framework, we find that the massless modes of different $q$-form fields can be localized on the thick brane for positive values of the coupling parameters $t_1$ and $t_2$. For the massive modes, the different $q$-form fields exhibit similar localization properties determined by the coupling parameter $t_2$. When $0 v^2/24$, the effective potentials associated with the Kaluza-Klein modes of these $q$-form fields form infinitely deep potential wells, so that all massive modes can be localized on the brane. Moreover, the tachyonic massive modes can always be excluded for different $q$-form fields.

hep-th

FlashNormal: Detailed Surface Normal Estimation from Flash and No-Flash Images

High-quality surface normal estimation is preferred for detailed surface shape recovery and image editing. Existing single image-based methods, though being a practical setup, often struggle to recover fine surface details and are sensitive to inherent shape-reflectance ambiguity. While photometric stereo achieves high-fidelity surface normal estimation from images under varying lights, its applicability is strictly limited by requiring a multi-illumination capture setup. To this end, we propose FlashNormal, a diffusion-based surface normal estimator from flash/no-flash image pairs. While retaining high practicability on modern smartphones, our proposal takes advantage of flash-induced shading variations, and leverages curvature-guided detail enhancement strategy, improving surface detail recovery and mitigating shape-reflectance ambiguity effectively. To evaluate our proposed method, we further present EvalFlash, the first real-world flash/no-flash evaluation dataset containing 20 objects aligned with ground-truth surface normals for quantitative benchmarking. Extensive experiments demonstrate the effectiveness of FlashNormal over state-of-the-art single image-based methods and show a significant out-performance over flash/no-flash-based normal estimation method on EvalFlash.

cs.CV

Approximate counting of vertices of 0/1 polytopes: a stronger hardness result

We show that approximately counting the vertices of a bounded 0/1 polytope, presented as a system of rational linear inequalities, is, informally speaking, NP-hard. In particular, there is no FPRAS for this problem unless RP=NP. The proof is by a reduction from approximately counting homomorphisms from a given graph to a particular four-vertex graph. The main proof ideas were found using GPT-5.6 Sol Ultra.

cs.CC

Approximating spin systems on planar graphs

We show that the hard-core partition function admits a fully polynomial-time randomised approximation scheme (FPRAS) on planar graphs when the activity is a sufficiently small constant. In contrast, we show that for any constant $q\ge 4$, approximately counting $q$-colourings in planar graphs is NP-hard. We also give a complete characterisation of when an FPRAS exists for a sufficiently small external field for 2-spin systems on planar graphs. The main ideas of all proofs were found using GPT-5.6 Sol Ultra.

cs.DS

Approximating two-terminal network reliability

We present a fully polynomial-time randomised approximation scheme (FPRAS) for the two-terminal reliability problem on general graphs, both directed and undirected. We also show that the complementary unreliability question is \BIS-hard. The key idea of the algorithm was discovered by GPT-5.6 Sol Ultra.

cs.DS

OF$^3$GS: On-the-Fly Feed-Forward 3D Gaussian Splatting from Unposed Images

Feed-forward 3D Gaussian Splatting (3DGS) enables efficient and high-fidelity novel view synthesis (NVS) from offline image sequences. However, achieving on-the-fly NVS from unposed images remains challenging: the system must reconstruct renderable 3D Gaussians as images arrive, without access to future observations. Although online feed-forward geometry methods have been developed for causal depth and point-cloud recovery, directly adapting them to NVS often leads to severe rendering artifacts because Gaussian-based rendering demands stricter multi-view consistency in primitive scale and pose-geometry alignment. Even minor deviations can accumulate under causal inference and visibly degrade rendering quality. To this end, we propose OF$^3$GS, a feed-forward framework for efficient and high-quality on-the-fly NVS from sparse-view unposed images under causal constraints. We introduce two mechanisms for causal geometric stability: a Decoupled Intrinsic Recovery Head that mitigates cumulative camera-intrinsic bias and scene-scale jitter, and Dynamic Point Refinement Offsets that relax rigid unprojection to compensate for coupled pose-depth drift. Extensive experiments show that OF$^3$GS outperforms online baselines and approaches offline feed-forward 3DGS methods under comparable sparse-input settings. It also remains memory-feasible with denser inputs. Homepage: https://richardchen225.github.io/of3gs/

cs.CV

Fast counting and sampling for ferromagnetic two-spin systems

We introduce two new models equivalent to ferromagnetic two-spin systems: a weighted subgraph model and a random cluster type model. Using these new connections, we obtain an efficient sampling algorithm and a new randomised algorithm that efficiently approximates the partition function of ferromagnetic two-spin systems in certain parameter regimes. No efficient sampling algorithms are known before in this regime, and our new estimation algorithm runs in near-quadratic time for bounded degree graphs and in polynomial time for general graphs, improving upon the previous algorithm of Guo, Liu, and Lu (2020).

cs.DS

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment

Video diffusion models lack explicit geometric supervision during training, leading to inconsistency artifacts such as object deformation, spatial drift, and depth violations in generated videos. To address this limitation, we propose a geometry-based reward model that leverages pretrained geometric foundation models to evaluate multi-view consistency through cross-frame reprojection error. Unlike previous geometric metrics that measure inconsistency in pixel space, where pixel intensity may introduce additional noise, our approach conducts error computation in a pointwise fashion, yielding a more physically grounded and robust error metric. Furthermore, we introduce a geometry-aware sampling strategy that filters out low-texture and non-semantic regions, focusing evaluation on geometrically meaningful areas with reliable correspondences to improve robustness. We apply this reward model to align video diffusion models through two complementary pathways: post-training of a bidirectional model via SFT or Reinforcement Learning and inference-time optimization of a Causal Video Model (e.g., Streaming video generator) via test-time scaling with our reward as a path verifier. Experimental results validate the effectiveness of our design, demonstrating that our geometry-based reward provides superior robustness compared to other variants. By enabling efficient inference-time scaling, our method offers a practical solution for enhancing open-source video models without requiring extensive computational resources for retraining.

cs.CV

Two-scalar-field $f(R)$ Thick Branes, Gravitational Resonances and Quasinormal Modes

In this paper, we investigate thick brane worlds in $f(R)$ gravity supported by two-scalar-field. The two-scalar sector provides an analytical warped background with tunable energy-density splitting, allowing us to test whether a Bloch-type internal structure can generate long-lived tensor perturbations resonances in the physically admissible region. We impose the positivity of \(f_R\equiv df/dR\), the derivative of the gravitational Lagrangian with respect to the Ricci scalar, which plays the role of an effective gravitational coupling in \(f(R)\) gravity. This separates the smooth ghost-free branch from a singular branch where this effective coupling vanishes. In the ghost-free branch, neither the relative-probability spectrum nor the phase-shift transmission spectrum shows narrow real-axis resonant peaks. These real-axis diagnostics indicate that the internal brane structure alone does not produce long-lived tensor resonances in the ghost-free region. Sharp quasi-localization peaks appear only in the singular branch, where the vanishing effective coupling induces divergent structures in the tensor potential; these peaks should therefore be interpreted as singular-boundary signals rather than ghost-free resonances of the smooth brane background. We then characterize the ghost-free massive Kaluza-Klein modes in the complex-frequency plane. Using the Asymptotic Iteration Method where applicable and time-domain evolutions with a supersymmetric partner potential as a zero-mode filtering tool, we extract the fundamental quasinormal frequencies. The modes have negative imaginary parts and quality factors \(Q\simeq0.9-1.9\), showing that the ghost-free massive tensor excitations are broad, short-lived dissipative modes. Thus the QNM spectrum provides the appropriate complex-frequency description of the Kaluza-Klein dynamics when no narrow real-axis resonances are resolved.

hep-th

Simulating Gaussian boson sampling on graphs in polynomial time

We show that a distribution related to Gaussian Boson Sampling (GBS) on graphs can be sampled classically in polynomial time. Graphical applications of GBS typically sample from this distribution, and thus quantum algorithms do not provide exponential speedup for these applications. We also show that another distribution related to Boson sampling can be sampled classically in polynomial time.

quant-ph

Beyond Text Following: Repairable Arbitration Reversals in Audio-Language Models

Audio-language models (ALMs) often follow text that conflicts with audio, even when the audio evidence is clear. This raises a basic question: is the audio-supported answer unavailable, or is it represented but overridden by the conflicting text? We examine this question using a same-audio counterfactual that keeps the audio fixed, removes only the conflicting text, and measures the resulting shift in model preference. Across five ALMs and four conflict tasks, 64.1% of conflict samples show a sign flip: the same-audio branch prefers the audio-supported answer, whereas the joint branch prefers the text-supported answer. This pattern suggests that the relevant audio evidence is encoded but loses in arbitration. Activation patching further localizes the reversal to answer-position computation, and patching effects closely track output candidate-score differences (Spearman rho=0.93). Using this diagnostic, we propose Gated Audio Counterfactual Logit Correction (GACL), a training-free decoding rule that interpolates between joint and same-audio scores. Under a strict 5 pp faithfulness-drop budget, GACL improves nAUC by 17.8 points over the best contrastive baseline and transfers without retuning to vision-text arbitration (up to +40.5 pp).

cs.SD

Rapid mixing in positively weighted restricted Boltzmann machines

We show polylogarithmic mixing time bounds for the alternating-scan sampler for positively weighted restricted Boltzmann machines. This is done via analysing the same chain and the Glauber dynamics for ferromagnetic two-spin systems, where we obtain new mixing time bounds up to the critical thresholds.

cs.DS

SignRoundV2: Toward Closing the Performance Gap in Extremely Low-Bit Post-Training Quantization for LLMs

Extremely low-bit quantization is critical for efficiently deploying Large Language Models (LLMs), yet it often leads to severe performance degradation at 2 bits and even at 4 bits (e.g., MXFP4). We present SignRoundV2, a post-training quantization framework designed to maintain high performance even under aggressive compression. SignRoundV2 introduces (1) a simple yet efficient adaptive mixed-precision strategy that leverages gradient information and quantization-induced reconstruction errors to guide layer-wise bit allocation, and (2) a set of lightweight stabilization techniques, including loss filtering and a pre-tuning scale search, to improve tuning effectiveness in extremely low-bit regimes. Our approach takes a significant step toward closing the performance gap between quantized and full-precision models. Experimental results across diverse LLMs demonstrate that SignRoundV2 achieves near-lossless performance in mixed MXFP settings, narrowing the gap to $\sim$1\% at an average of 4.5 bits, while substantially improving accuracy in challenging 2-bit weight-only quantization. The source code is available at \url{https://github.com/intel/auto-round}.

cs.CL

NTIRE 2026 The Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images: Methods and Results

This paper presents an overview of the NTIRE 2026 Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images. Building upon the success of the first edition, this challenge attracted a wide range of impressive solutions, all developed and evaluated on our real-world Raindrop Clarity dataset~\cite{jin2024raindrop}. For this edition, we adjust the dataset with 14,139 images for training, 407 images for validation, and 593 images for testing. The primary goal of this challenge is to establish a strong and practical benchmark for the removal of raindrops under various illumination and focus conditions. In total, 168 teams have registered for the competition, and 17 teams submitted valid final solutions and fact sheets for the testing phase. The submitted methods achieved strong performance on the Raindrop Clarity dataset, demonstrating the growing progress in this challenging task.

cs.CV