arXiv ScienceSearch

arXiv subjects

Minseok Kim

Publications and source records attributed to Minseok Kim.

At least 19 recordsLinked to original sources

FaVOR: LLM-Based Agentic Framework for Factor Mining via Empirical Validation

Traditional finance relies on experts to hand-craft factors through a principled process grounded in economic rationale. Recent LLM-based multi-agent systems have automated this process, scaling factor mining far beyond manual effort. However, these automated approaches optimize directly for returns and rarely check whether a generated factor still expresses the economic hypothesis that motivated it. We identify this inconsistency between mathematical form and economic meaning as a structural failure mode of return-oriented automation. The resulting factors blur the line between real signals and spurious correlations and break down across regime shifts. We propose FaVOR (Factor Validation through Observable Reasoning), an agentic framework that restructures factor mining around hypothesis-level evidence rather than return outcomes. In place of the standard hypothesis-to-formula leap, FaVOR enforces a three-stage consistency loop tying mathematical form to economic rationale throughout. (1) Decomposition splits a broad economic hypothesis into independent observable conditions. (2) Validation checks whether each factor reflects its intended condition. (3) Integration merges them into a composite whose structure remains interpretable. On the CSI 500 and S&P 500 in 2025, FaVOR outperforms existing baselines while remaining effective across regimes. FaVOR shows that hypothesis-grounded factor discovery produces signals that are interpretable by construction, regime-robust, and economically faithful. The code is available at https://github.com/damilab/FaVOR.

cs.AI

Real-time feedback control of ELM frequency using divertor gas puffing and its effects on tungsten-induced radiation and plasma performance in KSTAR

The edge-localized mode (ELM) frequency ($f_{\mathrm{ELM}}$) was successfully controlled in real time on KSTAR using a proportional-integral (PI) feedback controller, employing a $\mathrm{D}_2$ divertor gas puff as the actuator under tungsten lower-divertor conditions. The controller accurately tracked a two-step target---a 30 Hz increase in $f_{\mathrm{ELM}}$ for 4 s, followed by a 30 Hz decrease for 3 s---yielding mean and median absolute percentage errors of approximately 13% and 12%, respectively. Compared to a reference discharge, the actively controlled shot did not exhibit a significant drop in volume-integrated core radiation, confirming that excessive gas use merely degrades overall plasma performance. However, when contrasted with the exponential increase in core radiation observed in the absence of divertor gas puffing, these results underscore the critical need for real-time optimization. Specifically, divertor gas commands must be actively managed to maintain an $f_{\mathrm{ELM}}$ sufficient for flushing tungsten from the core while maximizing global plasma performance.

physics.plasm-ph

Multimodal Distribution Matching for Vision-Language Dataset Distillation

Dataset distillation compresses large training sets into compact synthetic datasets while preserving downstream performance. As modern systems increasingly operate on paired vision-language inputs, multimodal distillation must preserve representation quality and cross-modal alignment under tight compute and memory budgets, yet prior methods often require heavy computes and overlook their correlations. To address this, we present Multimodal Distribution Matching (MDM), a geometry-aware framework for efficient and generalizable multimodal distillation. Specifically, MDM integrates complementary components at the data, model, and loss levels. At the data level, it initializes synthetic image-text pairs by sampling from clusters in the joint embedding space. At the model level, it forms a mixed teacher by interpolating independently fine-tuned models in weight space according to their angular deviation from the pretrained anchor. At the loss level, it matches joint distributions on the unit hypersphere using a geometry-aware matching objective that exploits the joint features in the cross-modal agreement and discrepancy directions along with symmetric contrastive learning. Across image-text retrieval benchmarks with cross-architecture evaluation, MDM yields compact synthetic sets that preserve multimodal semantics, substantially reduce distillation cost, and remain robust across architectures.

cs.CV

Finite-temperature spin diffusion in the two-dimensional XY model

We present a combined theory-experiment study to quantify spin diffusion in the square lattice quantum spin-1/2 XY model at finite temperature. On the theory side, we leverage a recently developed dynamical high-temperature expansion method to faithfully capture the long spatiotemporal scales of the hydrodynamic regime. Experimental results are obtained from an optical lattice hard-core boson quantum simulator. The excellent agreement of spin diffusion constants marks a breakthrough in spin-transport beyond one dimension and for the quantitative validation of state-of-the-art quantum simulation platforms. We also provide theory predictions for future experiments on dynamic spin conductivity or anisotropy-induced integrability breaking.

cond-mat.quant-gas

Monotone Neural Policy Iteration for High-Dimensional First-Order Hamilton--Jacobi--Bellman Equations

We analyze a neural semi-discrete method for high-dimensional first-order Hamilton-Jacobi-Bellman (HJB) equations with known or learned dynamics. Centered differences and an artificial viscosity $Nh=O(h)$ define a monotone operator evaluated through $2d+1$ shifted network queries; policy iteration solves the resulting Bellman equation without a tensor grid. At fixed $h$, the sharp componentwise condition $\max_i|f_i|\le2N$ turns every frozen-policy operator into a nearest-neighbor Markov-chain generator with a policy-independent total jump rate. Uniformization gives whole-space well-posedness for measurable feedbacks, an explicit Poisson-tail bound on the numerical domain of dependence, and boundary-free localization. The representation also yields a posteriori policy-evaluation bounds that account for residual and learned-model errors. A greedy-gap analysis controls inexact policy iteration at fixed $h$; a separate consistency estimate connects the semi-discrete equation to the continuous HJB equation. Experiments reproduce the extremal tail, show rates consistent with $O(\sqrt h)$ and nearly $h$-independent exact-policy-iteration decay, and assess empirical estimator effectivity. A nonsmooth example shows that the continuous residual can miss a non-viscosity solution, whereas the shifted residual detects the defect. Further tests provide a structured interval-verified certificate calibration, an early-budget benefit of policy freezing for bang-bang control, and learned-dynamics diagnostics. A structured nonlinear problem with active compact-control constraints is tested against a manufactured semi-discrete reference through $d=1024$.

cs.LG

Environment-Aware Channel Prediction for Vehicular Communications: A Multimodal Visual Feature Fusion Framework

The deep integration of communication with intelligence and sensing, as a defining vision of 6G, renders environment-aware channel prediction a key enabling technology. As a representative 6G application, vehicular communications require accurate and forward-looking channel prediction under stringent reliability, latency, and adaptability demands. Traditional empirical and deterministic models remain limited in balancing accuracy, generalization, and deployability, while the growing availability of onboard and roadside sensing devices offers a promising source of environmental priors. This paper proposes an environment-aware channel prediction framework based on multimodal visual feature fusion. Using GPS data and vehicle-side panoramic RGB images, together with semantic segmentation and depth estimation, the framework extracts semantic, depth, and position features through a three-branch architecture and performs adaptive multimodal fusion via a squeeze-excitation attention gating module. For 360-dimensional angular power spectrum (APS) prediction, a dedicated regression head and a composite multi-constraint loss are further designed. As a result, joint prediction of path loss (PL), delay spread (DS), azimuth spread of arrival (ASA), azimuth spread of departure (ASD), and APS is achieved. Experiments on a synchronized urban V2I measurement dataset yield the best root mean square error (RMSE) of 3.26 dB for PL, RMSEs of 37.66 ns, 5.05 degrees, and 5.08 degrees for DS, ASA, and ASD, respectively, and mean/median APS cosine similarities of 0.9342/0.9571, demonstrating strong accuracy, generalization, and practical potential for intelligent channel prediction in 6G vehicular communications.

cs.CV

HarassGuard: Detecting Harassment Behaviors in Social Virtual Reality with Vision-Language Models

Social Virtual Reality (VR) platforms provide immersive social experiences but also expose users to serious risks of online harassment. Existing safety measures are largely reactive, while proactive solutions that detect harassment behavior during an incident often depend on sensitive biometric data, raising privacy concerns. In this paper, we present HarassGuard, a vision-language model (VLM) based system that detects physical harassment in social VR using only visual input. We construct an IRB-approved harassment vision dataset, apply prompt engineering, and fine-tune VLMs to detect harassment behavior by considering contextual information in social VR. Experimental results demonstrate that HarassGuard achieves competitive performance compared to state-of-the-art baselines (i.e., LSTM/CNN, Transformer), reaching an accuracy of up to 88.09% in binary classification and 68.85% in multi-class classification. Notably, HarassGuard matches these baselines while using significantly fewer fine-tuning samples (200 vs. 1,115), offering unique advantages in contextual reasoning and privacy-preserving detection.

cs.CV

Aligning Paralinguistic Understanding and Generation in Speech LLMs via Multi-Task Reinforcement Learning

Speech large language models (LLMs) observe paralinguistic cues such as prosody, emotion, and non-verbal sounds--crucial for intent understanding. However, leveraging these cues faces challenges: limited training data, annotation difficulty, and models exploiting lexical shortcuts over paralinguistic signals. We propose multi-task reinforcement learning (RL) with chain-of-thought prompting that elicits explicit affective reasoning. To address data scarcity, we introduce a paralinguistics-aware speech LLM (PALLM) that jointly optimizes sentiment classification from audio and paralinguistics-aware response generation via a two-stage pipeline. Experiments demonstrate that our approach improves paralinguistics understanding over both supervised baselines and strong proprietary models (Gemini-2.5-Pro, GPT-4o-audio) by 8-12% on Expresso, IEMOCAP, and RAVDESS. The results show that modeling paralinguistic reasoning with multi-task RL is crucial for building emotionally intelligent speech LLMs.

cs.CL

A Physics-Informed, Global-in-Time Neural Particle Method for the Spatially Homogeneous Landau Equation

We propose a physics-informed neural particle method (PINN--PM) for the spatially homogeneous Landau equation. The method adopts a Lagrangian interacting-particle formulation and jointly parameterizes the time-dependent score and the characteristic flow map with neural networks. Instead of advancing particles through explicit time stepping, the Landau dynamics is enforced via a continuous-time residual defined along particle trajectories. This design removes time-discretization error and yields a mesh-free solver that can be queried at arbitrary times without sequential integration. We establish a rigorous stability analysis in an $L^2_v$ framework. The deviation between learned and exact characteristics is controlled by three interpretable sources: (i) score approximation error, (ii) empirical particle approximation error, and (iii) the physics residual of the neural flow. This trajectory estimate propagates to density reconstruction, where we derive an $L^2_v$ error bound for kernel density estimators combining classical bias--variance terms with a trajectory-induced contribution. Using Hyvarinen's identity, we further relate the oracle score-matching gap to the $L^2_v$ score error and show that the empirical loss concentrates at the Monte Carlo rate, yielding computable a posteriori accuracy certificates. Numerical experiments on analytical benchmarks, including the two- and three-dimensional BKW solutions, as well as reference-free configurations, demonstrate stable transport, preservation of macroscopic invariants, and competitive or improved accuracy compared with time-stepping score-based particle and blob methods while using significantly fewer particles.

math.NA

Loss-Optimized Reconfigurable Nonlocal Metasurface-aided Cavity Antenna

This paper presents the design and experimental demonstration of a reconfigurable cavity excited nonlocal metasurface antenna capable of wide angle dynamic beam steering. The antenna is synthesized using a volume surface integral equation based framework that rigorously captures nonlocal mutual coupling among metasurface unit cells. To ensure physical consistency, the numerically characterized resistance and reactance relationship of the tunable unit cells is directly incorporated into the synthesis, enabling precise far-field synthesis while minimizing Ohmic losses. The proposed approach is applied to a 10 GHz cavity fed metasurface antenna composed of 24 independently controlled varactor-loaded unit cells. Numerical simulations and near-field measurements demonstrate stable beam steering with a range of 80 degrees across broadside with excellent agreement between measured and simulated radiation patterns. These results confirm the effectiveness of the proposed framework for the realization of compact, reconfigurable cavity-excited metasurface antennas.

physics.app-ph

Fool Me If You Can: On the Robustness of Binary Code Similarity Detection Models against Semantics-preserving Transformations

Binary code analysis plays an essential role in cybersecurity, facilitating reverse engineering to reveal the inner workings of programs in the absence of source code. Traditional approaches, such as static and dynamic analysis, extract valuable insights from stripped binaries, but often demand substantial expertise and manual effort. Recent advances in deep learning have opened promising opportunities to enhance binary analysis by capturing latent features and disclosing underlying code semantics. Despite the growing number of binary analysis models based on machine learning, their robustness to adversarial code transformations at the binary level remains underexplored. We evaluate the robustness of deep learning models for the task of binary code similarity detection (BCSD) under semantics-preserving transformations. The unique nature of machine instructions presents distinct challenges compared to the typical input perturbations found in other domains. We introduce asmFooler, a system that evaluates the resilience of BCSD models using a diverse set of adversarial code transformations that preserve functional semantics. We construct a dataset of 9,565 binary variants from 620 baseline samples by applying eight semantics-preserving transformations across six representative BCSD models. Our major findings highlight several key insights: i) model robustness relies on the processing pipeline, including code pre-processing, architecture, and feature selection; ii) adversarial transformation effectiveness is bounded by a budget shaped by model-specific constraints like input size and instruction expressive capacity; iii) well-crafted transformations can be highly effective with minimal perturbations; and iv) such transformations efficiently disrupt model decisions (e.g., misleading to false positives or false negatives) by focusing on semantically significant instructions.

cs.CR

Completing Missing Annotation: Multi-Agent Debate for Accurate and Scalable Relevant Assessment for IR Benchmarks

Information retrieval (IR) evaluation remains challenging due to incomplete IR benchmark datasets that contain unlabeled relevant chunks. While LLMs and LLM-human hybrid strategies reduce costly human effort, they remain prone to LLM overconfidence and ineffective AI-to-human escalation. To address this, we propose DREAM, a multi-round debate-based relevance assessment framework with LLM agents, built on opposing initial stances and iterative reciprocal critique. Through our agreement-based debate, it yields more accurate labeling for certain cases and more reliable AI-to-human escalation for uncertain ones, achieving 95.2% labeling accuracy with only 3.5% human involvement. Using DREAM, we build BRIDGE, a refined benchmark that mitigates evaluation bias and enables fairer retriever comparison by uncovering 29,824 missing relevant chunks. We then re-benchmark IR systems and extend evaluation to RAG, showing that unaddressed holes not only distort retriever rankings but also drive retrieval-generation misalignment. The relevance assessment framework is available at https: //github.com/DISL-Lab/DREAM-ICLR-26; and the BRIDGE dataset is available at https://github.com/DISL-Lab/BRIDGE-Benchmark.

cs.CL

Synthesized-Isotropic Narrowband Channel Parameter Extraction from Angle-Resolved Wideband Channel Measurements

Angle-resolved channel sounding using antenna arrays or mechanically steered high-gain antennas is widely employed at millimeter-wave and terahertz bands. To extract antenna-independent large-scale channel parameters such as path loss, delay spread, and angular spread, the radiation-pattern effects embedded in the measured responses must be properly compensated. This paper revisits the technical challenges of path loss/path gain calculation from angle-resolved wideband measurements, with emphasis on angular-domain power integration where the scan beams are inherently non-orthogonal and simple power summation leads to biased isotropic-equivalent power estimates. We first formulate the synthesized-isotropic narrowband power in a unified matrix form and introduce a beam-accumulation correction factor, including an offset-averaged variant to mitigate scalloping due to off-grid angles. The proposed framework is validated through simulations using channel models and 154~GHz corridor measurements.

eess.SP

Non-Local Metasurface-aided Leaky-Wave Antennas

This work presents a non-local terahertz metasurface integrated into a leaky-wave antenna for robust, wide-angle beam steering. The metasurface encodes a holographic pattern by explicitly inducing tangential and normal susceptibilities, along with magnetoelectric coupling. This design maintains stable radiation performance even when the longitudinal wavenumber of the incident guided mode - and thus its effective impinging angle - varies as a function of frequency. In particular, we show that there exists a limit to achieving exact angular insensitivity and propose an optimization-based framework to obtain the required susceptibilities that closely approximate near angle-insensitive performance for stable beam-steering performance. Additionally, an iterative synthesis approach is introduced that maps abstract susceptibilities to physically realizable structures. Full-wave simulations demonstrate a beam-scanning range of nearly 50 degrees over the 2.0-2.7 THz band - a more than threefold improvement over conventional local-metasurface designs.

physics.app-ph

Engineering Spatial Dispersion to Synthesize Arbitrary Spatial Filters Based on Metagratings

This paper presents a design framework for synthesizing angularly selective spatial filters using non-uniform metagratings. While traditional metagratings focus on channeling energy into higher-order Floquet modes for a fixed incidence angle, we leverage the fundamental mode as a versatile degree of freedom to engineer spatial dispersion over a continuous angular spectrum. By strategically distributing non-uniformly loaded metallic wires and rigorously modeling their mutual interactions through an impedance-matrix formulation, we realize prescribed angular transfer functions with high efficiency. In particular, the framework is validated at 3.5 GHz through full-wave simulations of (i) low-pass, (ii) high-pass, and (iii) all-pass spatial filters. The results demonstrate that fundamental-mode engineering in non-uniform metagratins offers a highly efficient platform for advanced spatial wave manipulation.

physics.app-ph

Nonlocal Dual-Band Reconfigurable Intelligent Surfaces for Precise Full-Space Beamforming

This paper introduces a nonlocal, dual-band reconfigurable intelligent surface (RIS) designed for full-space beam synthesis at 4.0 GHz and 6.3 GHz. The constituent unit cells comprise a pair of interleaved sub-cells that are specifically engineered to operate independently at their respective target frequencies. This hardware-level decoupling facilitates an efficient synthesis framework based on microwave network theory (MNT) that rigorously accounts for mutual coupling within both bands. Under this framework, the optimal biasing for sub-cells is determined to achieve precise full-space beam synthesis at both frequencies. The proposed method is numerically and experimentally validated with an RIS comprising 14 X 14 varactor-loaded unit cells that can be individually biased. We experimentally demonstrate arbitrary beam profile synthesis beyond simple beam steering, including dual-beam and sector patterns in full space. Experimental and simulation results show good agreement with the MNT model, confirming the effectiveness of the proposed method.

physics.app-ph

Transient Pauli blocking in a InN film as a mechanism for broadband ultrafast optical switching

The transient Pauli blocking effect offers a promising route for achieving ultrafast optical switching in semiconductors, enabling a rapid switching from an initially opaque state to a relatively transparent state upon photoexcitation. Herein, we demonstrate broadband ultrafast optical switching in degenerate InN thin films, spanning the visible to near-infrared spectral range, using pump-probe transient transmittance measurements. To elucidate the underlying physical mechanism, we perform probe-energy-resolved analysis for ultrafast dynamics, and develop a theoretical model based on a quasi-equilibrium Fermi-Dirac distribution. The model successfully captures the experimental transients and yields an electron-phonon coupling constant of $1.0\times10^{17}\,\mathrm{W\,m^{-3}\,K^{-1}}$, along with an electronic specific heat coefficient ranging from 1.52 to 2.02 $\mathrm{mJ\,mol^{-1}\,K^{-2}}$, which allow direct prediction of the spectral switching window. Notably, we demonstrate that the Pauli blocking effect can be induced solely by a laser-excitation driven rise in electronic temperature, without requiring significant carrier injection into the conduction band in degenerate semiconductors. These findings offer new insights for designing ultrafast optical modulators, shutters, and photonic devices for next-generation communication and computing technologies.

physics.optics

Urban Macro/Microcellular Channel Characterization at 4.85 GHz With Literature-Referenced Upper FR1-to-FR3 Cross-Band Analysis

The transition from 5G to 6G requires frequency-dependent, physically consistent radio channel models across the upper-FR1/FR3 transition region, particularly in the under-explored $4$--$8$~GHz region targeted in the current WRC-$27$ studies, where outdoor urban channel measurements and characterizations remain scarce. This paper presents a $4.85$~GHz measurement-anchored study of urban channels and a literature-referenced cross-band analysis. Double-directional measurements were conducted at $4.85$~GHz in urban macrocell (UMa) and urban microcell (UMi) routes in Yokohama, Japan, from which path loss, delay spread (DS), azimuth spread of arrival/departure (ASA/ASD), $K$-factor, and route-dependent spatial-consistency statistics were extracted. To align these results in a broader cross-band context, the measured $4.85$~GHz large-scale parameter (LSP) means were combined with scenario-matched literature anchors to derive log-log trends for DS, ASA, and ASD over an approximately $4$--$28$~GHz range around the $7.125$~GHz upper-FR1/FR3 cross-band boundary. The resulting trends were compared with 3GPP UMa/UMi reference parameterizations over the same interval, and the sensitivity of the UMi DS fit was examined via leave-one-out analysis. Because the cross-band analysis still relies on a single in-house measurement band alongside heterogeneous anchors from different campaigns, it is presented as measurement-informed and indicative rather than as a definitive multi-band model. The paper therefore contributes both a detailed, parameterized $4.85$~GHz urban measurement reference and a bounded literature-referenced view of channel behavior near the upper-FR1/FR3 transition.

eess.SP