arXiv ScienceSearch

arXiv subjects

Yuzhou Wang

Publications and source records attributed to Yuzhou Wang.

12 recordsLinked to original sources

Computational Thresholds for Balanced and Fixed-Slice Independent Sets in Bipartite Graphs

Motivated by recent work of Kocurek, Oveis Gharan, and Tjowasi, which gives an efficient sampling algorithm for the hard-core model on random regular bipartite graphs by decomposing into fixed-size slices, we study the worst-case tractability of approximate counting and sampling of fixed-size slices for bipartite independent set problems. Let $G=(L\sqcup R,E)$ be a bipartite graph with $|L|=|R|=n$ and maximum degree $\Delta$. The fixed-slice problem asks to sample uniformly from independent sets satisfying $|I\cap L|=\alpha_L n$ and $|I\cap R|=\alpha_R n$. We show that if the overall density $\alpha$ lies in the interval $(\frac{1}{\Delta}, \tfrac{1}{2})$, and the densities on the two sides are more balanced than the typical phase densities of a random $\Delta$-regular bipartite graph, then there is no FPRAS or efficient sampling scheme unless $\mathbf{NP}=\mathbf{RP}$. We then study a related fugacity model in which the densities are not fixed, but the independent set is required to be balanced between the two sides of the bipartition. For $\lambda>0$, the balanced hard-core model is the ordinary hard-core model with fugacity $\lambda$, conditioned on the event $|I\cap L|=|I\cap R|$. We prove that this model has the same computational threshold as the hard-core model on general bounded-degree graphs. That is, for every fixed $\Delta\ge 3$, if $\lambda<\lambda_c(\Delta)$, then the balanced partition function admits an FPTAS and the balanced hard-core distribution admits an efficient sampling scheme. Conversely, if $\lambda>\lambda_c(\Delta)$, then no FPRAS or efficient sampler exists on this graph class unless $\mathbf{NP}=\mathbf{RP}$.

cs.DS

A Gaia-linked High-purity QSO Candidate Catalog in Selected Fields with Extinction-binned Calibration and Spectrum-informed Training

We present an extinction-calibrated, Gaia-source-level QSO candidate catalog for selected fields, designed as a high-purity input catalog for fiber-spectroscopic follow-up rather than as an all-sky QSO census. The deployed selector uses Gaia astrometry and photometry, optical/infrared catalog features, and E(B-V)-binned threshold calibration; spectra are used only during training via a source-grouped spectrum-teacher model. The sample definition is layered: a four-field core domain ladder provides the main validation baseline, four application/stress-test fields probe portability, and COSMOS is treated separately as an Extreme Deep boundary case. At the recommended conservative operating point, calibrated to a validation-set purity of 0.98, the P3 spectrum-informed catalog selector achieves a measured test-set purity of 0.9809 and a spectroscopic-label completeness of 0.8869 within the frozen Gaia-linked benchmark, whereas the Gaia official QSO probability yields a spectroscopic-label completeness of 0.4493 under the same threshold protocol. The evaluation protocol excludes downstream validation/test Gaia source IDs from teacher fitting and checkpoint selection, and uses teacher probabilities only for downstream training rows. Relative to the earlier P2 teacher, P3 yields a modest mean completeness gain across five seeds, with a small decrease in purity and a small increase in false positives; the gain is most evident in higher-extinction and faint-source diagnostics. The released product is a catalog and empirical selection-function data product with source identifiers, field-layer assignments, input-coverage flags, calibrated scores, threshold flags, validation metadata, and provenance/QC fields. In COSMOS, the Gaia-linked parent set is much shallower than COSMOS2020; the robust 39-object subset is interpreted as a purity-oriented priority list rather than a completeness measurement.

astro-ph.IM

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth

Video depth estimation extends monocular prediction into the temporal domain to ensure coherence. However, existing methods often suffer from spatial blurring in fine-detail regions and temporal inconsistencies. We argue that current approaches, which primarily rely on temporal smoothing via Transformers, struggle to maintain strict 3D geometric consistency-particularly under rotations or drastic view changes. To address this, we propose GemDepth, a framework built on the insight that an explicit awareness of camera motion and global 3D structure is a prerequisite for 3D consistency. Distinctively, GemDepth introduces a Geometry-Embedding Module (GEM) that predicts inter-frame camera poses to generate implicit geometric embeddings. This injection of motion priors equips the network with intrinsic 3D perception and alignment capabilities. Guided by these geometric cues, our Alternating Spatio-Temporal Transformer (ASTT) captures latent point-level correspondences to simultaneously enhance spatial precision for sharp details and enforce rigorous temporal consistency. Furthermore, GemDepth employs a data-efficient training strategy, effectively bridging the gap between high efficiency and robust geometric consistency. As shown in Fig.2, comprehensive evaluations demonstrate that GemDepth achieves state-of-the-art performance across multiple datasets, particularly in complex dynamic scenarios. The code is publicly available at: https://github.com/Yuecheng919/GemDepth.

cs.CV

Constant-Stepsize Stochastic Approximation: Finite-Time Convergence, Gaussian Approximation, and Tail Bounds

Constant-stepsize stochastic approximation (SA) is widely used in learning for computational efficiency, yet the distribution of the iterates is typically intractable. Classical asymptotics results give $X_k^{(\alpha)} \approx X^{(\alpha)} \approx x^\star+\sqrt{\alpha}Y$, where $X^{(\alpha)}$ is the steady state and $Y$ is an appropriate Gaussian limit, by progressively taking the time $k\uparrow\infty$ and stepsize $\alpha\downarrow0$. Such limit results, however, do not quantify finite-time, finite-stepsize errors. We develop an explicit pre-limit characterization for SA with i.i.d.\ and Markovian noise. We establish existence and uniqueness of the stationary law, a geometric Wasserstein convergence to stationarity, and almost-sure and $L^3$ convergence of the steady state to the root $x^\star$, identifying the scale $\sqrt{\alpha}$ as first-order fluctuation. At this scale, we derive a higher-order quantitative Gaussian approximation with a Wasserstein error, using Stein's method and Poisson equation techniques. We further obtain non-uniform Berry--Esseen-type tail bounds, incorporating both steady-state approximation and finite-time convergence errors. We instantiate the theory for strongly convex smooth SGD, linear SA, and nonlinear contractive SA. Beyond strong convexity, for general convex SGD, we identify a Gibbs limiting law and prove a pre-limit Wasserstein approximation error under stability and Stein-equation hypothesis, which are validated numerically.

cs.LG

On the chromatic number of random triangle-free graphs

We study the chromatic number of typical triangle-free graphs with $\Theta \left( n^{3/2} (\log n)^{1/2} \right)$ edges and establish the width of the scaling window for the transitions from $\chi = 3$ to $\chi = 4$ and from $\chi = 4$ to $\chi = 5$. The transition from $3$- to $4$-colorability has scaling window of width $\Theta(n^{4/3} (\log n)^{-1/3})$. To prove this, we show a high probability equivalence of the $3$-colorability of a random triangle-free graph at this density and the satisfiability of an instance of bipartite random $2$-SAT, for which we establish the width of the scaling window following the techniques of Bollob{\'a}s, Borgs, Chayes, Kim, and Wilson. The transition from $4$- to $5$-colorability has scaling window of width $\Theta(n^{3/2} (\log n)^{-1/2})$. To prove this, we show a high probability equivalence of the $4$-colorability of a random triangle-free graph at this density and the simultaneous $2$-colorability of two independent Erd\H{o}s--R\'enyi random graphs. For this transition, we also establish the limiting probability of $4$-colorability inside the scaling window.

math.CO

Balanced colorings of Erd\H{o}s-R\'enyi hypergraphs

An $r$-uniform hypergraph $H = (V, E)$ is $r$-partite if there exists a partition of the vertex set into $r$ parts such that each edge contains exactly one vertex from each part. We say an independent set in such a hypergraph is balanced if it contains an equal number of vertices from each partition. The balanced chromatic number of $H$ is the minimum value $q$ such that $H$ admits a proper $q$-coloring where each color class is a balanced independent set. In this note, we determine the asymptotic behavior of the balanced chromatic number for sparse $r$-uniform $r$-partite Erd\H{o}s--R\'enyi hypergraphs. A key step in our proof is to show that any balanced colorable hypergraph of average degree $d$ admits a proper balanced coloring with $r(r-1)d + 1$ colors. This extends a result of Feige and Kogan on bipartite graphs to this more general setting.

math.CO

Unifying 2D and 3D Vision-Language Understanding

Progress in 3D vision-language learning has been hindered by the scarcity of large-scale 3D datasets. We introduce UniVLG, a unified architecture for 2D and 3D vision-language understanding that bridges the gap between existing 2D-centric models and the rich 3D sensory data available in embodied systems. Our approach initializes most model weights from pre-trained 2D models and trains on both 2D and 3D vision-language data. We propose a novel language-conditioned mask decoder shared across 2D and 3D modalities to ground objects effectively in both RGB and RGB-D images, outperforming box-based approaches. To further reduce the domain gap between 2D and 3D, we incorporate 2D-to-3D lifting strategies, enabling UniVLG to utilize 2D data to enhance 3D performance. With these innovations, our model achieves state-of-the-art performance across multiple 3D vision-language grounding tasks, demonstrating the potential of transferring advances from 2D vision-language learning to the data-constrained 3D domain. Furthermore, co-training on both 2D and 3D data enhances performance across modalities without sacrificing 2D capabilities. By removing the reliance on 3D mesh reconstruction and ground-truth object proposals, UniVLG sets a new standard for realistic, embodied-aligned evaluation. Code and additional visualizations are available at https://univlg.github.io .

cs.CV

MonSter++: Unified Stereo Matching, Multi-view Stereo, and Real-time Stereo with Monodepth Priors

We introduce MonSter++, a geometric foundation model for multi-view depth estimation, unifying rectified stereo matching and unrectified multi-view stereo. Both tasks fundamentally recover metric depth from correspondence search and consequently face the same dilemma: struggling to handle ill-posed regions with limited matching cues. To address this, we propose MonSter++, a novel method that integrates monocular depth priors into multi-view depth estimation, effectively combining the complementary strengths of single-view and multi-view cues. MonSter++ fuses monocular depth and multi-view depth into a dual-branched architecture. Confidence-based guidance adaptively selects reliable multi-view cues to correct scale ambiguity in monocular depth. The refined monocular predictions, in turn, effectively guide multi-view estimation in ill-posed regions. This iterative mutual enhancement enables MonSter++ to evolve coarse object-level monocular priors into fine-grained, pixel-level geometry, fully unlocking the potential of multi-view depth estimation. MonSter++ achieves new state-of-the-art on both stereo matching and multi-view stereo. By effectively incorporating monocular priors through our cascaded search and multi-scale depth fusion strategy, our real-time variant RT-MonSter++ also outperforms previous real-time methods by a large margin. As shown in Fig.1, MonSter++ achieves significant improvements over previous methods across eight benchmarks from three tasks -- stereo matching, real-time stereo matching, and multi-view stereo, demonstrating the strong generality of our framework. Besides high accuracy, MonSter++ also demonstrates superior zero-shot generalization capability. We will release both the large and the real-time models to facilitate their use by the open-source community.

cs.CV

The Low-Degree Hardness of Finding Large Independent Sets in Sparse Random Hypergraphs

We study the algorithmic task of finding large independent sets in Erdos-Renyi $r$-uniform hypergraphs on $n$ vertices having average degree $d$. Krivelevich and Sudakov showed that the maximum independent set has density $\left(\frac{r\log d}{(r-1)d}\right)^{1/(r-1)}$. We show that the class of low-degree polynomial algorithms can find independent sets of density $\left(\frac{\log d}{(r-1)d}\right)^{1/(r-1)}$ but no larger. This extends and generalizes earlier results of Gamarnik and Sudan, Rahman and Virag, and Wein on graphs, and answers a question of Bal and Bennett. We conjecture that this statistical-computational gap holds for this problem. Additionally, we explore the universality of this gap by examining $r$-partite hypergraphs. A hypergraph $H=(V,E)$ is $r$-partite if there is a partition $V=V_1\cup\cdots\cup V_r$ such that each edge contains exactly one vertex from each set $V_i$. We consider the problem of finding large balanced independent sets (independent sets containing the same number of vertices in each partition) in random $r$-partite hypergraphs with $n$ vertices in each partition and average degree $d$. We prove that the maximum balanced independent set has density $\left(\frac{r\log d}{(r-1)d}\right)^{1/(r-1)}$ asymptotically. Furthermore, we prove an analogous low-degree computational threshold of $\left(\frac{\log d}{(r-1)d}\right)^{1/(r-1)}$. Our results recover and generalize recent work of Perkins and the second author on bipartite graphs. While the graph case has been extensively studied, this work is the first to consider statistical-computational gaps of optimization problems on random hypergraphs. Our results suggest that these gaps persist for larger uniformities as well as across many models. A somewhat surprising aspect of the gap for balanced independent sets is that the algorithm achieving the lower bound is a simple degree-1 polynomial.

cs.CC

On the hardness of finding balanced independent sets in random bipartite graphs

We consider the algorithmic problem of finding large \textit{balanced} independent sets in sparse random bipartite graphs, and more generally the problem of finding independent sets with specified proportions of vertices on each side of the bipartition. In a bipartite graph it is trivial to find an independent set of density at least half (take one of the partition classes). In contrast, in a random bipartite graph of average degree $d$, the largest balanced independent sets (containing equal number of vertices from each class) are typically of density $(2+o_d(1)) \frac{\log d}{d}$. Can we find such large balanced independent sets in these graphs efficiently? By utilizing the overlap gap property and the low-degree algorithmic framework, we prove that local and low-degree algorithms (even those that know the bipartition) cannot find balanced independent sets of density greater than $(1+\epsilon) \frac{\log d}{d}$ for any $\epsilon>0$ fixed and $d$ large but constant. This factor $2$ statistical--computational gap between what exists and what local algorithms can achieve is analogous to the gap for finding large independent sets in (non-bipartite) random graphs. Our results therefor suggest that this gap is pervasive in many models, and that hard computational problems can lurk inside otherwise tractable ones. A particularly striking aspect of the gap in bipartite graphs is that the algorithm achieving the lower bound is extremely simple and can be implemented as a $1$-local algorithm and a degree-$1$ polynomial (a linear function).

cs.DS

Temperature-dependent elastic constants of thorium dioxide probed using time-domain Brillouin scattering

We report the adiabatic elastic constants of single-crystal thorium dioxide over a temperature range of 77 - 350 K. Time-domain Brillouin scattering (TDBS), an all-optical, non-contact picosecond ultrasonic technique, is used to generate and detect coherent acoustic phonons that propagate in the bulk perpendicular to the surface of the crystal. These coherent acoustic lattice vibrations have been monitored in two hydrothermally grown single-crystal thorium dioxide samples along the (100) and (311) crystallographic directions. The three independent elastic constants of the cubic crystal (C11, C12 and C44) are determined from the measured bulk acoustic velocities. The longitudinal wave along the (100) orientation provided a direct measurement of C11. Measurement of C44 and C12 was achieved by enhancing the intensity of quasi-shear mode in a (311) oriented crystal by adjusting the polarization angle relative to the crystal axes. We find the magnitude of softening of the three elastic constants to be ~2.5% over the measured temperature range. Good agreement is found between the measured elastic constants with previously reported values at room temperature, and between the measured temperature-dependent bulk modulus with calculated values. We find that semi-empirical models capturing lattice anharmonicity adequately reproduce the observed trend. We also determine the acoustic Gruneisen anharmonicity parameter from the experimentally derived temperature-dependent bulk modulus and previously reported temperature-dependent values of volume thermal expansion coefficient and heat capacity. This work presents measurements of the temperature-dependent elasticity in single-crystal thorium dioxide at cryogenic temperature and provides a basis for testing ab initio theoretical models and evaluating the impact of anharmonicity on thermophysical properties.

cond-mat.mtrl-sci

Maximum determinant and permanent of sparse 0-1 matrices

We prove that the maximum determinant of an $n \times n $ matrix, with entries in $\{0,1\}$ and at most $n+k$ non-zero entries, is at most $2^{k/3}$, which is best possible when $k$ is a multiple of 3. This result solves a conjecture of Bruhn and Rautenbach. We also obtain an upper bound on the number of perfect matchings in $C_4$-free bipartite graphs based on the number of edges, which, in the sparse case, improves on the classical Bregman's inequality for permanents. This bound is tight, as equality is achieved by the graph formed by vertex disjoint union of 6-vertex cycles.

math.CO