arXiv ScienceSearch

arXiv subjects

Xiaohu Yang

Publications and source records attributed to Xiaohu Yang.

At least 19 recordsLinked to original sources

ELUCID. IX. Recovering the Substructures and History of the Coma Cluster

We employ constrained hydrodynamic simulations of the Coma galaxy cluster from the ELUCID project to study its substructures and assembly history. Our simulations accurately reproduce the global properties of Coma, including its position, virial mass, radius, and surrounding large-scale filaments. Using a combined HBT+SKID method, we obtain a total intracluster light (ICL) fraction of $12.2\%-23.5\%$, consistent with recent observations. The mass-weighted ICL profiles of velocity dispersion, stellar age, and iron abundance generally agree with MaNGA measurements, and its clumpy east-west elongation is closely linked to the merger history of the two brightest cluster galaxies (BCGs). The simulation also successfully reproduces the north and west intracluster filaments (ICFs) detected by weak lensing, and the surrounding galaxy groups are distributed in alignment with the directions toward the nearby A2199 and A1367 clusters. The simulation further predict a complex cluster assembly history, with two major mergers at $z=0.74$ and $0.45$, and a pericentric passage of the two BCGs occurring $\sim 0.64 \ {\rm Gyr}$ before its current state. Our results demonstrate that constrained simulations are a powerful tool for connecting observed structures to the unobservable assembly histories of individual galaxy clusters.

astro-ph.GA

Growth from KiDS: Stellar mass assembly of galaxies since z=2 in the light of Kilo-Degree Survey

We measure the galaxy stellar mass function at $z = 0-2$ using 28M galaxies from the KiDS DR4 which spans an effective survey area of 676.9 deg$^2$. Combining nine-band photometry ($u$--$K_s$) and redshifts derived using deep learning photometry, we use five SED fitting configurations, CIGALE/LePhare and CB19/BC03/M05 Stellar Population Synthesis (SPS). We present the impact of the systematics related to the code, SPS modeling and Star formation History (SFH) , which combined can introduce systematics of $\sim 0.6$ dex. The GSMFs outlined in this work agree well with previous studies (within the uncertainties introduced by systematics) that typically are able to cover measurements for the GSMF only up to $\log(M_*/M_\odot)\sim 11.75$, due to their limited volume. The large KiDS area and galaxy counts, allows us to gain some insights on the ``extreme" high-mass end, and our GSMFs imply that possibly there are more massive galaxies than we would expect, from a simple exponential cut-off, at $\log(M_*/M_\odot)>11.75$. However, we heed caution that despite advances done both in redshift estimation techniques and SPS, future studies are required to confirm this behavior. We consider the effects of Eddington-bias (EB) to our observations via two different methods and compare our results with the predictions of semi-analytic models and hydrodynamical simulations. We use the Mean Absolute Difference (MAD) of residuals as a diagnostic metric. Due to inconsistencies ($> 0.5$ dex) between observations and Eddington biased simulations found, especially at the high mass end, we suggest that some theoretical models would achieve better performance once parameter tuning takes into account the effects of EB prior to the tuning of the uncertain parameters involved related to stellar physics and feedback.

astro-ph.GA

A graph-based Neural Network surrogate model for accelerating semi-analytical model of galaxy formation and evolution

Understanding how galaxy populations emerge and evolve from the growth of dark matter structure is a central challenge in galaxy formation theory. Semi-analytic models (SAMs) provide an efficient framework to address this problem, but exploring large ensembles of merger trees across broad parameter spaces remains computationally demanding. We develop a conditional graph neural network surrogate model that combines merger tree information with SAM parameters to predict galaxy properties across cosmic time. Using merger trees of dark matter halos from the Uchuu simulation and the Galacticus SAM, the model predicts stellar mass, luminosity, angular momentum, gas metal mass, and specific star formation rate across the wide redshift range of 0 <= z <= 5. For instance, the model can predict stellar mass at 0 <= z <= 3 with a scatter of 0.19-0.28 dex and coefficient of determination R^2 of 0.946-0.973 (R^2 close to 1 indicates prediction closely matching the truth). The results show that a single graph based model can reproduce these galaxy properties with good accuracy over multiple SAM realizations, merger trees and redshifts. This catalog-level model provides a practical route for accelerating SAM based studies of galaxy formation to enable a more detailed investigation of the model parameter space. The inference code, trained models, and example data products are publicly available at https://github.com/MutongCat/sam2galaxy-gnn.

astro-ph.GA

Quantifying Environmental Effects on Galaxy Properties using Non-spherical Voids Identified from SDSS DR7

Cosmic voids provide a distinct low-density region for studying the environmental effects of galaxy properties. Using the SDSS DR7 catalog, we identify non-spherical voids via Voronoi tessellation and the watershed algorithm, and classify void galaxies based on their local volume. We compare and find that void galaxies classified by this method are systematically less massive, fainter, bluer, and have higher specific star formation rate (sSFR) than non-void galaxies and all galaxy samples. We then divide void and non-void galaxies into stellar mass bins to focus on the environmental dependence of $g-r$ color and sSFR. By further classifying galaxies into blue/red and star-forming/quiescent populations, we calculate the ratio of blue to red and star-forming to quiescent for void and non-void galaxies separately. Comparing the ratio of the void value to the non-void value for both metrics presents an overall decreasing trend with stellar mass $M_*$ over the $9.4-10.4$ range in $\log[M_*/\mathrm{M}_\odot]$, indicating a stronger environmental effect in lower-mass systems. These results show that our classification of void galaxies in non-spherical voids based on local volume offers a robust approach for quantifying the influence of underdense environments on galaxy evolution.

astro-ph.GA

Massive Galaxy Halos Contain Less Inner Dark Matter Than Predicted

The mass profiles of galaxy halos encode how baryons reshape dark matter distribution, yet direct observational constraints across the full radial range remain scarce. Here we combine stellar kinematics from MaNGA, H I dynamical measurements from ALFALFA, and independently calibrated halo masses of SDSS groups to statistically reconstruct the mass distribution of central galaxies over nearly two orders of magnitude in radius. We demonstrate that H I data alone do not provide reliable total halo mass estimates, necessitating an independent group-based halo-mass scale. Compared to the IllustrisTNG and EAGLE simulations, the observational profiles of low-mass halos are broadly consistent; in contrast, massive observed halos exhibit systematically lower dynamical masses at the H I radius, lower inner dark-matter masses, and lower central dark-matter fractions (about 4$σ$ difference in units of population scatter) at fixed total halo mass. After subtracting baryonic contributions, the inferred dark-matter profiles remain broadly consistent with an NFW form, but with lower effective concentrations than predicted for massive halos. These results suggest that the inner dark-matter content of massive halos has been reduced more significantly than predicted by current hydrodynamical simulations, plausibly due to long-term baryonic halo heating in massive systems.

astro-ph.GA

Reconstructing the Projected Dark Matter Field across 0.1-100 Mpc Scales from the SDSS Survey

Dark matter sets the gravitational environment in which galaxies form and evolve, but cannot be observed directly. We present a conditional diffusion model that reconstructs the projected dark matter density field from the galaxy stellar-mass density field for direct application to galaxy surveys. The model is trained on CAMELS and validated on the independent IllustrisTNG300-1 simulation. Halo masses inferred from the reconstructed projected-aperture measurements agree well with the corresponding true values, with a scatter below 0.2 dex. On 100 kpc scales, reconstructed surface densities show a typical scatter of ~0.3 dex in the regime most relevant for observations. We apply the model to SDSS galaxies with $M_\star\ge10^9\,M_\odot$ in a contiguous low-redshift region. Averaging over 100 stochastic realizations, we reconstruct and publicly release a projected dark matter field covering $90\times90\,(h^{-1}\mathrm{Mpc})^2$ with a pixel size of $0.097\,h^{-1}\mathrm{Mpc}$. This pixel area corresponds to the characteristic projected area of halos with masses of ~$10^{10.6}\,h^{-1}\,M_\odot$. The map reveals the multiscale projected cosmic web, including cluster-scale overdensities, filaments and voids. Projected-aperture masses are statistically consistent with SDSS group-catalog masses, while the derived halo mass function broadly matches mock-catalog expectations. The reconstructed projected potential places Coma in one of the deepest wells and near a convergence region of the inferred projected acceleration field, suggesting that the reconstruction retains both local overdensities and coherent large-scale projected gravitational structure. This work shows that diffusion-based dark matter reconstruction can be applied to real galaxy surveys, enabling halo-mass- and spatially resolved dark-matter-environment-based studies of galaxy evolution in SDSS and future wide-area surveys.

astro-ph.GA

Black Hole-Galaxy Correlations in Cluster Zoomed-in Simulations: GIZMO-SIMBA and TNG-Cluster

We investigate the co-evolution of supermassive black holes (SMBHs) and central galaxies in massive clusters using the GIZMO-SIMBA and TNG-Cluster zoom-in simulations at $z=0-5$. We find that the distinct subgrid physics of these two models suggest fundamentally different evolutionary pathways. On the one hand, GIZMO-SIMBA, employs torque-limited accretion and predicts a supply-driven scenario where the SMBHs rapidly assemble synchronized with dark matter halo ($M_{200c}$) growth (i.e. the halo mass-BH mass relation is set by $z=3.0$ and similar to the present-day relationship). On the other hand, TNG-Cluster, exhibits a feedback-regulated growth phase delayed by an early thermal suppression. While both models successfully reproduce some local black hole-galaxy scaling relations, they imply significantly different evolution for these relations. Analysis of the BH mass-gas mass ratio relations suggests that TNG-Cluster's isotropic kinetic winds efficiently deplete cold gas, resulting in a "hard quench" of star formation. In the black hole accretion rate (BHAR)-star formation rate (SFR) relation we find that both simulations successfully reproduce the decoupling of BHAR and star formation observed in recent massive cluster ellipticals. The divergent evolutionary trends emphasize the importance of the multiphase intracluster medium; while these subgrid models do not have the necessary resolution and employ distinct formalisms, the sustained BHAR in quenched systems resemble outcomes broadly consistent with modern multiphase feeding paradigms such as chaotic cold accretion in turbulent cluster cores. Furthermore, we demonstrate that for both models, black hole mass is a primary regulator of atomic and molecular gas depletion in galaxy clusters.

astro-ph.GA

The Sinking Statistics of Dark Matter Subhalos Across Hierarchical Levels

We investigate the mergers among subhalos in a $Λ$CDM simulation, focusing on two fundamental aspects overlooked by previous studies: 1) how to identify mergers robustly; 2) the statistics of mergers across the subhalo hierarchy. To this end, we make use of the HBT+ subhalo finder that tracks subhalo evolution across hierarchy levels, identifying the coalescence of subhalo cores in phase space as a "sinking" event. This coalescence marks a distinct stalled phase in orbital decay, providing a physically motivated and natural definition of a resolved merger. Moreover, the phase transition occurs over a very short timescale, making the statistics robust to numerical resolutions. Our main findings are as follows. (1) More than 90% of sinking events occur between adjacent subhalo levels, while cross-level pathways arise from tidal stripping, group accretion, and numerical constraints. (2) Resolved sinking events are predominantly major mergers (mass ratios > 1/10), whereas the contribution from minor mergers decreases with the dynamical age of the host halo. (3) Although deep-level subhalos typically have low mass ratios relative to the host halo, their mass ratios relative to their direct parents are substantially larger, significantly enhancing their sinking probability. Consequently, the satellite-satellite sinking rate can rival or even exceed the central-satellite sinking rate at lower mass thresholds. (4) Satellite-satellite sinking events are spatially biased toward the outer regions of the host halo, suggesting that the central tidal field suppresses orbital decay between satellite systems.

astro-ph.GA

The Web4 Agent Economy: A Large-Scale Empirical Study of the Landscape, Challenges, and Opportunities

The Internet is transitioning from Web3 toward Web4, where autonomous agents serve as independent economic actors. These agents can now hold crypto wallets, execute on-chain trades, and pay for external API calls. This transition calls for a new infrastructure stack capable of supporting key agent operations, including agent-to-tool interaction, agent-to-agent payments, and verifiable agent identity, represented by emerging protocols such as the Model Context Protocol, x402, and EIP-8004. Despite growing industrial interest in these protocols, the real-world Web4 agent ecosystem remains largely underexplored. To bridge this gap, we conduct the first large-scale empirical study of the Web4 ecosystem. Specifically, our study targets three interconnected questions: how Web4 agents are deployed and used in practice; what engineering challenges developers face when building Web4 agents; how current project communities respond to these challenges. To answer these questions, we analyze 99,448 multi-chain identity registrations, 317,596,323 transaction logs, the source code of 341 MCP projects, and 349 filtered GitHub issues. Our findings reveal that autonomous agents have established a highly active machine-to-machine payment economy, processing millions of daily transactions. However, this growth is built on immature infrastructure, including identity/authorization practice, cross-environment operation, and payment interoperability. Our follow-up analysis shows that community responses are visible but unevenly distributed across repositories, and payment interoperability remains the most persistent unresolved bottleneck. Overall, this study reveals a critical gap between the rapid growth of the Web4 agent economy and its fragile underlying infrastructure, highlighting future directions for building a more secure Web4 agent ecosystem.

cs.SE

ELUCID-DESI II. Revealing dark matter mass, tidal, and velocity (MTV) fields using galaxy group phase information

We introduce a novel method for reconstructing the cosmic mass, tidal, and velocity (MTV) fields over the redshift range $0 < z < 0.6$ using the phase information of galaxy groups. This approach replaces the explicit theoretical bias correction typically needed to relate galaxy groups to the underlying dark matter density field with a simulation-calibrated statistical mapping, reducing a major source of systematic uncertainty and making the method directly applicable to spectroscopic redshift surveys such as the DESI Bright Galaxy Survey (BGS). We evaluate the performance of our MTV reconstruction pipeline with mock redshift surveys that include a comprehensive set of observational selection effects. The galaxy groups used as tracers are identified with an extended halo-based group finder applied to the DESI mock galaxy catalogue with an apparent magnitude limit of $m_z < 19.65$, yielding a galaxy number comparable to that of the DESI BGS faint sample ($m_r < 20.175$). Our tests show that the reconstructed velocities are accurate and unbiased, with a residual dispersion of $\sim 120\ \mathrm{km\,s^{-1}}$ across the redshift bins. The recovered velocity field allows us to shift galaxy groups to their real-space positions, thereby correcting for the Kaiser effect. By iteratively applying this Kaiser correction to the galaxy groups, we further reconstruct the tidal field and the mass-density distribution. The reconstruction is stable with respect to the grid resolution. Overall, our results demonstrate that this group-based phase-space reconstruction provides a robust pathway to recovering the dark matter MTV fields, with strong prospects for application to DESI BGS data.

astro-ph.CO

Diffploit: Facilitating Cross-Version Exploit Migration for Open Source Library Vulnerabilities

Exploits are commonly used to demonstrate the presence of library vulnerabilities and validate their impact across different versions. However, their direct application to alternative versions often fails due to breaking changes introduced during evolution. These failures stem from both changes in triggering conditions (e.g., API refactorings) and broken dynamic environments (e.g., build or runtime errors), which are challenging to interpret and adapt manually. Existing techniques primarily focus on code-level trace alignment through fuzzing, which is both time-consuming and insufficient for handling environment-level failures. Moreover, they often fall short when dealing with complicated triggering condition changes across versions. To overcome this, we propose Diffploit, an iterative, diff-driven exploit migration method structured around two key modules: the Context Module and the Migration Module. The Context Module dynamically constructs contexts derived from analyzing behavioral discrepancies between the target and reference versions, which capture the failure symptom and its related diff hunks. Leveraging these contexts, the Migration Module guides an LLM-based adaptation through an iterative feedback loop, balancing exploration of diff candidates and gradual refinement to resolve reproduction failures effectively. We evaluate Diffploit on a large-scale dataset containing 102 Java CVEs and 689 version-migration tasks across 79 libraries. Diffploit successfully migrates 84.2% exploits, outperforming the change-aware test repair tool TARGET by 52.0% and the rule-based tool in IDEA by 61.6%. Beyond technical effectiveness, Diffploit identifies 5 CVE reports with incorrect affected version ranges, three of which have been confirmed. It also discovers 111 unreported vulnerable versions in the GitHub Advisory Database.

cs.SE

RACE-Bench: A Reasoning-Augmented Benchmark for Repository-Level Code Agents on Feature Addition

Repository-level code agents have shown strong promise in real-world feature addition tasks, making reliable evaluation of their capabilities increasingly important. However, existing benchmarks primarily evaluate these agents as black boxes based on final test correctness, providing limited insight into how they reason and where failures arise. To address this limitation, we introduce RACE-bench, a reasoning-augmented benchmark for evaluating code agents on repository-level feature addition tasks. RACE-bench contains 528 real-world feature addition instances from 12 open-source repositories. Each instance is paired with executable patch verification and structured intermediate reference reasoning covering issue understanding, file localization, implementation tasks, and step decomposition. Based on this design, we introduce a dual-track evaluation framework that jointly measures patch correctness and intermediate reasoning alignment with developer-accepted reference trajectories. We evaluate three representative repository-level code agents on RACE-bench. On the full benchmark, Resolved Rate ranges from 29% to 70% across different agents. Our reasoning-level analysis further shows that while current agents perform well at understanding high-level intent, their performance degrades substantially when translating intent into concrete implementation steps. We also find patches that can be applied but still fail the tests cover fewer reference-reasoning elements (35.7% lower recall) and contain more unsupported reasoning elements (94.1% higher over-prediction) than successful patches. These findings highlight the importance of evaluating repository-level code agents beyond final patch correctness by examining the quality of their reasoning processes.

cs.SE

Multi-tracer mass bias in matched cosmic voids from SDSS DR7 and the ELUCID constrained simulation

Cosmic voids provide a unique environment for studying the relationship between galaxies, subhaloes, and dark matter in the underdense Universe. Using the SDSS galaxy catalogue and the ELUCID constrained simulation, we establish an observationally anchored framework for measuring multi-tracer mass bias within matched cosmic voids. A sample of 102 matched void pairs is constructed to directly compare galaxy, subhalo, and dark matter mass distributions within an observationally constrained realisation of the local Universe. We find that both the galaxy-to-dark matter and subhalo-to-dark matter mass ratios decrease toward void centres, indicating that luminous and halo tracers become increasingly depleted relative to the underlying matter distribution in the deepest underdensities. In contrast, the galaxy-to-subhalo mass ratio exhibits substantially larger statistical uncertainties within the inner void regions ($r/R_{\rm v}\lesssim0.5$). By comparing measurements obtained using independent and common coordinate frameworks, we show that coordinate offsets contribute to the observed scatter but cannot fully account for the large uncertainties. The remaining uncertainty primarily arises from the severe scarcity of massive subhaloes ($\log_{10}(M_{\rm sub}/h^{-1}M_\odot)\ge11.8$) within void interiors, which greatly reduces the number of statistically valid measurements near void centres. Our results provide a direct measurement of multi-tracer mass bias in observationally constrained cosmic environments and highlight the fundamental statistical limitations of multi-tracer studies in extreme underdense regions.

astro-ph.CO

A high-significance detection of primordial tidal torque imprints

Tidal-torque theory predicts that galaxy angular momenta are imprinted by the primordial tidal field acting on proto-structures and that they can retain information about the early Universe through cosmic evolution. Here we test this prediction by comparing observed galaxy angular momentum vectors with those predicted from the primordial density field reconstructed by ELUCID for the nearby Universe. Among the galaxy populations considered, the gas component of central massive elliptical galaxies provides the clearest signal, exhibiting a strong direction correlation at a significance of about $7σ$. These results provide a robust observational evidence for tidal-torque theory and open a window for cosmological measurements of neutrino mass and other cosmological parameters.

astro-ph.CO

Demystifying Solana Bots: From GitHub Blueprints to On-Chain Fingerprints

Solana is an emerging blockchain platform designed for high throughput and low transaction fees, making it inexpensive to submit transactions at scale and, consequently, increasing exposure to bot spamming and related financial exploitation. Solana bots are typically off-chain software systems that operate in a competitive on-chain execution environment by constructing and submitting transactions, and the bot-related transactions on the decentralized exchanges exceed 250 million dollars in daily trading volume in January 2026. Prior studies on Solana have examined system performance, smart-contract security, and specific on-chain phenomena. However, we still lack a systematic understanding of what Solana bots implement in practice and how these implementations manifest as observable on-chain execution fingerprints. To address this gap, we performed a large-scale empirical study of Solana bots from two complementary views: (i) 586 bot repositories collected from GitHub, and (ii) 200 bot addresses on Solana, with over 44 million on-chain transactions. Our study derives an implementation-grounded taxonomy of Solana bots comprising 15 categories grouped into five domains (e.g., Trading Operations, MEV, and On-chain Analytics), identifies a largely shared five-stage operational pipeline manifested in bot implementations, and uncovers systematic variation in on-chain trading behaviors of Solana bots across diverse trading platforms and assets. Based on our findings, we highlight future research directions, and provide recommendations for building and operating bots on the Solana blockchain.

cs.SE

Assessing the Cross-Version Applicability of Java Library Vulnerability Exploits

Open-source software supply chain security relies heavily on assessing affected versions of library vulnerabilities. While prior studies have leveraged exploits for verifying vulnerability affected versions, they point out a key limitation that exploits are version-specific and cannot be directly applied across library versions. Despite being widely acknowledged, this limitation has not been systematically validated at scale, leaving the actual applicability of exploits across versions unexplored. To fill this gap, we conduct the first large-scale empirical study on exploit applicability across library versions. We construct a comprehensive dataset consisting of 259 exploits spanning 128 Java libraries and 28,150 historical versions, covering 61 CWEs that account for 76.33% of vulnerabilities in Maven. Leveraging this dataset, we execute each exploit against the library version history and compare the execution outcomes with our manually annotated ground-truth affected versions. We further investigate the root causes of inconsistencies between exploit execution and ground truth, and explore strategies for exploit migration. Our results (RQ1) show that, even without migration, exploits achieve 83.0% recall and 99.3% precision in identifying affected versions in Java, outperforming most widely used vulnerability databases and assessment tools. Notably, this capability enables us to contribute 796 confirmed missing affected versions to the CPE dictionary. We investigate the remaining exploit failures (RQ2) and find that they mainly stem from compatibility issues introduced by library evolution and changing environmental constraints. Based on these observations, we manually migrate exploits for 1,885 versions and distill a taxonomy of 10 strategies from these successful adaptation cases (RQ3), thereby increasing the overall recall to 96.1%.

cs.SE

Galaxy populations in groups and clusters-II. Conditional luminosity functions at redshifts from z~1 to z~0

Using DESI SV3 spectroscopic group centrals and HSC photometric data, we measure conditional luminosity functions (CLFs) of central and satellite galaxies for red and blue populations in dark matter haloes spanning $M_h\sim10^{12}- 10^{15}M_{\odot}$ and $0<z<1$. HSC depth permits measurements to $M_r \approx -15$ at $0.2 \leqslant z < 0.5$ and $M_r \approx -17$ at $0.5 \leqslant z < 1.0$. We find satellite CLFs evolve weakly over $0<z<1$. Blue satellite CLFs are well described by a single Schechter function across halo masses and redshifts, with a nearly constant slope of $-1.25\lesssim α\lesssim -1.2$. In contrast, red satellite CLFs exhibit a pronounced faint-end upturn in all halo mass and redshift bins, with little evolution in the faint-end slope ($-1.8\lesssim α_f\lesssim -1.7$). The low-mass red sequence was therefore already established in clusters/groups by $z\sim1$. The lack of faint-end-slope evolution favors models where the steep upturn originates from early formation processes at $z\gtrsim2$, rather than environmental quenching after infall. Satellite characteristic magnitudes and central galaxy luminosities fade with time. Red central galaxies are consistent with passive evolution, whereas blue-central luminosity evolution is dominated by ongoing star formation. Satellites evolve more rapidly than predicted by simple stellar population models, highlighting environmental effects. Satellite quenched fractions as a function of stellar mass exhibit a minimum at $M_{*} \sim 10^9M_{\odot}$ that is consistent across halo masses and redshifts. We discuss possible interpretations of these results and their implications for galaxy formation and evolution.

astro-ph.GA

CSST large-scale structure analysis pipeline: IV. Cosmic Voids Identified from Galaxy Group Samples as Probes of the Large-scale Structure

Because groups are directly associated with halos, they allow for considerably simpler theoretical modeling than approaches based on individual galaxies. We therefore propose to use voids identified in galaxy group catalogs, referred to as group-voids, to investigate the cosmic large-scale structure (LSS). Using the reference mock galaxy redshift survey (MGRS) designed for the Chinese Space-station Survey Telescope (CSST), we build two galaxy group catalogs representing ideal and realistic scenarios, derived from galaxy samples with 100\% and roughly 30\% spectroscopic redshift completeness, respectively. We then identify voids in these two mock group catalogs, as well as in the underlying halo catalog, and measure two void statistics, the void size function (VSF) and the void density profile, within five redshift intervals spanning $z=0$ to $1.0$. We compare the statistics obtained from two kinds of voids: those defined by galaxy groups (group-voids) and those defined by dark matter halos (halo-voids). In the void-finding process, we adopt the brightest central galaxy (BCG) as the group center to improve the accuracy of the inferred void centers. Our analysis shows that void statistics derived from group-voids with spectroscopic redshift completeness of at least 40\% can faithfully reproduce the corresponding statistics from halo-voids. Even when the redshift completeness of galaxies falls to as low as 30\%, we can still reliably describe group-voids via halo-voids by incorporating a redshift error term. This indicates that group-voids are a promising tool for probing LSS and offer a valuable complement to standard void studies, which is especially advantageous for emulator-based methods.

astro-ph.CO