arXiv ScienceSearch

arXiv subjects

Inkyu Park

Publications and source records attributed to Inkyu Park.

At least 19 recordsLinked to original sources

Climbing the $N$-point Ladder Part I: Information in the Higher-Order Configuration-Space Clustering of Dark Matter Halos

We quantify the information content of the configuration-space two-, three-, and connected four-point correlation functions of Quijote dark-matter haloes at $z=0$ and fixed number density. We build Fisher forecasts for $\{Ω_m, Ω_b, h, n_s, σ_8, M_ν\}$ in real and redshift space from ${\sim}38{,}000$ GPU-accelerated $N$-point measurements. Treating the statistics as a ladder, $\mathrm{2PCF} \rightarrow +\mathrm{3PCF} \rightarrow +ζ^{(4)}_{\mathrm{conn}}$, we report the information gained at each rung. The 3PCF supplies most of the accessible higher-order information: it tightens every parameter, most strongly $σ_8$ and $M_ν$, whose degeneracy it partially breaks, by factors of $1.6$--$6.3$ over the redshift-space $\{ξ_0,ξ_2\}$ baseline with per-parameter gains consistent with those of the Fourier-space halo bispectrum on the same simulations. The 3PCF gains persist under conservative, modelling-motivated scale cuts: restricted to the tree-level-validated regime with minimum triangle side $\ge40\,h^{-1}{\rm Mpc}$, it still improves $M_ν$ by $3.9\times$. The connected 4PCF adds a further $\sim1.2$--$1.5\times$, an increment insensitive to whether the quadrupole is included in the baseline but traceable to small-scale configurations with sides $\lesssim30\,h^{-1}{\rm Mpc}$. This rung-to-rung increment is stable against derivative-sample noise and compression regularization, whereas the absolute constraints remain limited by the finite simulation ensembles and are reported as preliminary. The configuration-space ladder thus offers an independent and complementary route to the higher-order information probed by the Fourier-space poly-spectra.

astro-ph.CO

KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs

Speech language models (SpeechLMs) have achieved substantial progress by extending large language models (LLMs) to the speech modality. However, SpeechLM evaluation remains heavily centered on English, limiting reliable assessment of multilingual speech capabilities. Straightforward benchmark transfer through ASR, translation, normalization, and TTS can corrupt language-specific instructions, answer constraints, and spoken forms; for audio understanding, transferring source-language audio also fails to preserve target-language speaker attributes, accents, and paralinguistic properties. To address these limitations, we propose two human-agent benchmark-construction frameworks: one transfers source-language SpokenQA benchmarks into target-language SpokenQA benchmarks, and the other converts target-language ASR corpora into audio understanding benchmarks using transcriptions and speaker metadata. Using these frameworks, we construct and publicly release three Korean speech benchmarks: KVoiceBench and KOpenAudioBench for Korean SpokenQA, and KMMAU for Korean audio understanding, comprising 12,345 samples in total. We evaluate eight recent SpeechLMs and find that English-Korean performance gaps vary substantially across models and task families, and that SpokenQA and audio understanding rankings diverge, revealing complementary weaknesses invisible to English-only evaluation.

cs.CL

Cold Neutron Imaging and Efficiency Measurements with a Boron-10 Coated Double-GEM Detector

A ${}^{10}\mathrm{B}$-coated double-GEM neutron detector (BGEM) was developed as a ${}^{3}\mathrm{He}$-free cold-neutron beamline detector using a single $\mathrm{B}_{4}\mathrm{C}$ converter cathode and a 512-channel APV25 orthogonal-strip readout over an active area of $10 \times 10~\mathrm{cm}^{2}$. The detector was tested at the HANARO Bio-REF beamline with a monochromatic $4.5~\mathring{\mathrm{A}}$ beam ($E_{n}=4.03~\mathrm{meV}$). The absolute detection efficiency relative to a ${}^{6}\mathrm{Li}$-based Ce:LiCAF reference detector was $\varepsilon_{\mathrm{BGEM}}=(8.69 \pm 0.20)\%$ (stat.). The pulse-height spectrum was qualitatively consistent with Geant4 energy-deposition simulations, and Cd-mask imaging yielded a Gaussian-equivalent edge-spread width of $σ= 555 \oplus 102~μ\mathrm{m}$. These results establish a cold-neutron beamline benchmark for a single-converter BGEM detector with full-strip APV25 readout.

physics.ins-det

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games

Large Language Model (LLM) agents are reshaping the game industry, by enabling more intelligent and human-preferable characters. Yet, current game benchmarks fall short of practical needs: they lack evaluations of diverse LLM capabilities across various game genres, studies of agentic modules crucial for complex gameplay, and fine-tuning datasets to adapt pre-trained LLMs into gaming agents. To fill these gaps, we present Orak, a benchmark for training and evaluating LLM agents across 12 popular video games spanning all major genres. Using a plug-and-play interface built on Model Context Protocol (MCP), Orak supports systematic and reproducible studies of agentic modules in varied game scenarios. We further release a fine-tuning dataset of expert LLM gameplay trajectories covering multiple genres, turning general LLMs into effective game agents. Orak offers a united evaluation framework, including game leaderboards, LLM battle arenas, and \fix{ablation studies} of input modality, agentic strategies, and fine-tuning effects, establishing a foundation towards versatile gaming agents. Code and datasets are available at https://github.com/krafton-ai/Orak and https://huggingface.co/datasets/KRAFTON/Orak.

cs.AI

Raon-Speech Technical Report

We present Raon-Speech, a top-performing 9B-parameter speech language model (SpeechLM) for English and Korean speech understanding, answering, and generation, and Raon-SpeechChat, a high-performing full-duplex extension for natural real-time conversation. Raon-Speech successfully transforms a pre-trained LLM into a SpeechLM that both understands and generates speech while preserving strong text capabilities. It trains on 1.38M hours of highly curated English and Korean speech and text datasets with the following training stages: (1) speech modules alignment, (2) end-to-end SpeechLM pre-training with knowledge distillation, and (3) multi-task preference optimization-based post-training. Across 42 English and Korean speech and text benchmarks, Raon-Speech establishes the strongest overall profile on speech-centric tasks in our comparison against eight similarly sized recent audio foundation models, including Qwen2.5-Omni and Fun-Audio-Chat, while preserving strong text question answering performance. Building upon it, Raon-SpeechChat enables natural full-duplex conversation by continual training on 119K hours of time-aligned real and synthetic dialogue data. It proceeds through three complementary training stages: (1) causal encoder adaptation, (2) full-duplex pre-training, (3) full-duplex fine-tuning for voice and role-control. On multiple full-duplex benchmarks, Raon-SpeechChat shows its clearest strengths on the turn-taking and interruption-sensitive behaviors covered by FDB v1.0, and remains competitive across the broader full-duplex evaluation suite. We open-source all model checkpoints, the training and inference pipeline, and an interactive demo.

cs.CL

Adversarial Reinforcement Learning Framework for ESP Cheater Simulation

Extra-Sensory Perception (ESP) cheats, which reveal hidden in-game information such as enemy locations, are difficult to detect because their effects are not directly observable in player behavior. The lack of observable evidence makes it difficult to collect reliably labeled data, which is essential for training effective anti-cheat systems. Furthermore, cheaters often adapt their behavior by limiting or disguising their cheat usage, which further complicates detection and detector development. To address these challenges, we propose a simulation framework for controlled modeling of ESP cheaters, non-cheaters, and trajectory-based detectors. We model cheaters and non-cheaters as reinforcement learning agents with different levels of observability, while detectors classify their behavioral trajectories. Next, we formulate the interaction between the cheater and the detector as an adversarial game, allowing both players to co-adapt over time. To reflect realistic cheater strategies, we introduce a structured cheater model that dynamically switches between cheating and non-cheating behaviors based on detection risk. Experiments demonstrate that our framework successfully simulates adaptive cheater behaviors that strategically balance reward optimization and detection evasion. This work provides a controllable and extensible platform for studying adaptive cheating behaviors and developing effective cheat detectors.

cs.LG

Simple Drop-in LoRA Conditioning on Attention Layers Will Improve Your Diffusion Model

Current state-of-the-art diffusion models employ U-Net architectures containing convolutional and (qkv) self-attention layers. The U-Net processes images while being conditioned on the time embedding input for each sampling step and the class or caption embedding input corresponding to the desired conditional generation. Such conditioning involves scale-and-shift operations to the convolutional layers but does not directly affect the attention layers. While these standard architectural choices are certainly effective, not conditioning the attention layers feels arbitrary and potentially suboptimal. In this work, we show that simply adding LoRA conditioning to the attention layers without changing or tuning the other parts of the U-Net architecture improves the image generation quality. For example, a drop-in addition of LoRA conditioning to EDM diffusion model yields FID scores of 1.91/1.75 for unconditional and class-conditional CIFAR-10 generation, improving upon the baseline of 1.97/1.79.

cs.CV

Final parsec problem of black hole mergers and ultralight dark matter

When two galaxies merge, they often produce a supermassive black hole binary (SMBHB) at their center. Numerical simulations with stars and cold dark matter show that SMBHBs typically stall out at a distance of a few parsecs apart and take billions of years to coalesce. This is known as the final parsec problem. We suggest that ultralight dark matter (ULDM) halos around SMBHBs can generate dark matter waves due to dynamical friction. These waves can effectively carry away orbital energy from the black holes, rapidly driving them together. To test this hypothesis, we performed numerical simulations of black hole binaries inside ULDM halos. Due to gravitational cooling and quasi-normal modes, the loss-cone problem can be avoided. The decay time scale gives lower bounds on masses of the ULDM particles and SMBHBs comparable to observational data. Our results imply that ULDM waves can lead to the rapid orbital decay of black hole binaries.

astro-ph.GA

Constructing a Galaxy Cluster Catalog in IllustrisTNG-300 using the Mulguisin Algorithm

We present a new simulated galaxy cluster catalog based on the IllustrisTNG simulation. We use the Mulguisin (MGS) algorithm to identify galaxy overdensities. Our cluster identification differs from the previous FoF cluster identification in two aspects; 1) we identify cluster halos based on the galaxy subhalos instead of unobservable dark matter particles, and 2) we use the MGS algorithm that separates galaxy overdensities hosted by massive galaxies. Our approach provides a cluster catalog constructed similar to the observed cluster catalogs using spectroscopic surveys. The MGS cluster catalog lists 303 halos with M$_{200} > 10^{14}$ M$_{\odot}$, including $\sim 10\%$ more than the FoF. The MGS catalog includes more systems because we separate some independent massive MGS cluster halos that are bundled into a single FoF algorithm. These independent MGS halos are apparently distinguishable in galaxy spatial distribution and the phase-space diagram. Because we constructed a refined cluster catalog that identifies local galaxy overdensities, we evaluate the effect of MGS clusters on the evolution of galaxies better than using the FoF cluster catalog. The MGS halo identification also enables effective identifications of merging clusters by selecting systems with neighboring galaxy overdensities. We thus highlight that the MGS cluster catalog is a useful tool for studying clusters in cosmological simulations and for comparing with the observed cluster samples.

astro-ph.GA

MulGuisin, a Topological Network Finder and its Performance on Galaxy Clustering

We introduce a new clustering algorithm, MulGuisin (MGS), that can identify distinct galaxy over-densities using topological information from the galaxy distribution. This algorithm was first introduced in an LHC experiment as a Jet Finder software, which looks for particles that clump together in close proximity. The algorithm preferentially considers particles with high energies and merges them only when they are closer than a certain distance to create a jet. MGS shares some similarities with the minimum spanning tree (MST) since it provides both clustering and network-based topology information. Also, similar to the density-based spatial clustering of applications with noise (DBSCAN), MGS uses the ranking or the local density of each particle to construct clustering. In this paper, we compare the performances of clustering algorithms using controlled data and some realistic simulation data as well as the SDSS observation data, and we demonstrate that our new algorithm finds networks most efficiently and defines galaxy networks in a way that most closely resembles human vision.

astro-ph.IM

SAiD: Speech-driven Blendshape Facial Animation with Diffusion

Speech-driven 3D facial animation is challenging due to the scarcity of large-scale visual-audio datasets despite extensive research. Most prior works, typically focused on learning regression models on a small dataset using the method of least squares, encounter difficulties generating diverse lip movements from speech and require substantial effort in refining the generated outputs. To address these issues, we propose a speech-driven 3D facial animation with a diffusion model (SAiD), a lightweight Transformer-based U-Net with a cross-modality alignment bias between audio and visual to enhance lip synchronization. Moreover, we introduce BlendVOCA, a benchmark dataset of pairs of speech audio and parameters of a blendshape facial model, to address the scarcity of public resources. Our experimental results demonstrate that the proposed approach achieves comparable or superior performance in lip synchronization to baselines, ensures more diverse lip movements, and streamlines the animation editing process.

cs.CV

The Universe is worth $64^3$ pixels: Convolution Neural Network and Vision Transformers for Cosmology

We present a novel approach for estimating cosmological parameters, $Ω_m$, $σ_8$, $w_0$, and one derived parameter, $S_8$, from 3D lightcone data of dark matter halos in redshift space covering a sky area of $40^\circ \times 40^\circ$ and redshift range of $0.3 < z < 0.8$, binned to $64^3$ voxels. Using two deep learning algorithms, Convolutional Neural Network (CNN) and Vision Transformer (ViT), we compare their performance with the standard two-point correlation (2pcf) function. Our results indicate that CNN yields the best performance, while ViT also demonstrates significant potential in predicting cosmological parameters. By combining the outcomes of Vision Transformer, Convolution Neural Network, and 2pcf, we achieved a substantial reduction in error compared to the 2pcf alone. To better understand the inner workings of the machine learning algorithms, we employed the Grad-CAM method to investigate the sources of essential information in activation maps of the CNN and ViT. Our findings suggest that the algorithms focus on different parts of the density field and redshift depending on which parameter they are predicting. This proof-of-concept work paves the way for incorporating deep learning methods to estimate cosmological parameters from large-scale structures, potentially leading to tighter constraints and improved understanding of the Universe.

astro-ph.CO

Zero-Permutation Jet-Parton Assignment using a Self-Attention Network

In high-energy particle physics events, it can be advantageous to find the jets associated with the decays of intermediate states, for example, the three jets produced by the hadronic decay of the top quark. Typically, a goodness-of-association measure, such as a $χ^2$ related to the mass of the associated jets, is constructed, and the best jet combination is found by optimizing this measure. As this process suffers from a combinatorial explosion with the number of jets, the number of permutations is limited by using only the $n$ highest $p_T$ jets. The self-attention block is a neural network unit used for the neural machine translation problem, which can highlight relationships between any number of inputs in a single iteration without permutations. In this paper, we introduce the Self-Attention for Jet Assignment (SaJa) network. SaJa can take any number of jets for input and outputs probabilities of jet-parton assignment for all jets in a single step. We apply SaJa to find jet-parton assignments of fully-hadronic $t\bar{t}$ events to evaluate the performance. We show that SaJa achieves better performance than a likelihood-based approach.

hep-ex

Analyzing Planar Galactic Halo Distributions with Fuzzy/Cold Dark Matter Models

We perform a numerical comparison between the fuzzy dark matter model and the cold dark matter model, focusing on formation of satellite galaxy planes around massive galaxies. Such galactic dynamics with controlled initial subhalo configurations are investigated using GADGET2 for the cold dark matter and PyUltraLight for the fuzzy dark matter, respectively. We demonstrate that satellite galaxies in the fuzzy dark matter side have a tendency to form more flattened and corotating satellite systems than in the cold dark matter side mainly due to the dissipation by the gravitational cooling effect of the fuzzy dark matter. Our simulations with the fuzzy dark matter typically show the minor-to-major axis ratio $c/a$ of the satellite galaxy planes to be $0.21 \sim 0.30$; This well matches the current observed value for the Milky Way.

astro-ph.CO

Tracking Halo Orbits and Their Mass Evolution around Large-scale Filaments

We have explored the dynamical and mass evolution of halos driven by large-scale filaments using a dark matter-only cosmological simulation with the help of a phase-space analysis. Since a non-negligible number of galaxies is expected to fall into the cluster environment through large-scale filaments, tracking how halos move around large-scale filaments can provide a more comprehensive view on the evolution of cluster galaxies. Halos exhibit orbital motions around filaments, which emerge as specific trajectories in a phase space composed of halos' perpendicular distance and velocity component with respect to filaments. These phase-space trajectories can be represented by three cases according to their current states. We parameterize the trajectories with halos' initial position and velocity, maximum velocity, formation time, and time since first crossing, which are found to be correlated with each other. These correlations are explained well in the context of the large-scale structure formation. The mass evolution and dynamical properties of halos seem to be affected by the density of filaments, which can be shown from the fact that halos around denser filaments are more likely to lose their mass and be bound within large-scale filaments. Finally we reproduce the mass segregation trend around filaments found in observations. It is resulted because halos that formed earlier arrived filaments earlier, and grew efficiently there being more massive. We also found that dynamical friction helps to retain this segregation trend.

astro-ph.CO

Role of Acoustic Phonon Transport in Near- to Asperity-Contact Heat Transfer

Acoustic phonon transport is revealed as a potential radiation-to-conduction transition mechanism for single-digit nanometer vacuum gaps. To show this, we measure heat transfer from a feedback-controlled platinum nanoheater to a laterally oscillating silicon tip as the tip-nanoheater vacuum gap distance is precisely controlled from a single-digit nanometer down to bulk contact in a high-vacuum shear force microscope. The measured thermal conductance shows a gap dependence of $d^{-5.7\pm1.1}$ in the near-contact regime, which is in good agreement with acoustic phonon transport modeling based on the atomistic Green's function framework. The obtained experimental and theoretical results suggest that acoustic phonon transport across a nanoscale vacuum gap can be the dominant heat transfer mechanism in the near- and asperity-contact regimes and can potentially be controlled by an external force stimuli.

cond-mat.mes-hall

Probing Ultra-light Axion Dark Matter from 21cm Tomography using Convolutional Neural Networks

We present forecasts on the detectability of Ultra-light axion-like particles (ULAP) from future 21cm radio observations around the epoch of reionization (EoR). We show that the axion as the dominant dark matter component has a significant impact on the reionization history due to the suppression of small scale density perturbations in the early universe. This behavior depends strongly on the mass of the axion particle. Using numerical simulations of the brightness temperature field of neutral hydrogen over a large redshift range, we construct a suite of training data. This data is used to train a convolutional neural network that can build a connection between the spatial structures of the brightness temperature field and the input axion mass directly. We construct mock observations of the future Square Kilometer Array survey, SKA1-Low, and find that even in the presence of realistic noise and resolution constraints, the network is still able to predict the input axion mass. We find that the axion mass can be recovered over a wide mass range with a precision of approximately 20\%, and as the whole DM contribution, the axion can be detected using SKA1-Low at 68\% if the axion mass is $M_X<1.86 \times10^{-20}$eV although this can decrease to $M_X<5.25 \times10^{-21}$eV if we relax our assumptions on the astrophysical modeling by treating those astrophysical parameters as nuisance parameters.

astro-ph.CO

Measuring $|V_{ts}|$ directly using strange-quark tagging at the LHC

The Cabibbo-Kobayashi-Maskawa (CKM) element $V_{ts}$, representing the coupling between the top and strange quarks, is currently best determined through fits based on the unitarity of the CKM matrix, and measured indirectly through box-diagram oscillations, and loop-mediated rare decays of the $B$ or $K$ mesons. It has been previously proposed to use the tree level decay of the $t$ quark to the $s$ quark to determine $|V_{ts}|$ at the LHC, which has become a top factory. In this paper, we extend the proposal by performing a detailed analysis of measuring $t \to sW$ in dileptonic $t\bar{t}$ events. In particular, we perform detector response simulation, including the reconstruction of $K_S$, which are used for tagging jets produced by $s$ quarks against the dominant $t \to bW$ decay. We show that it should be possible to exclude $|V_{ts}| = 0$ at 6.0$σ$ with the expected High Luminosity LHC luminosity of 3000 fb$^{-1}$.

hep-ph