arXiv ScienceSearch

arXiv subjects

Jongwon Lee

Publications and source records attributed to Jongwon Lee.

At least 19 recordsLinked to original sources

Toward Compact Fiber In-line Nonlinear Devices via Highly Efficient Nanophotonic Cavity Interface

Compact and efficient frequency conversion within optical fibers is highly desirable for nonlinear and quantum photonic technologies, yet it remains challenging due to weak nonlinear interactions and limited coupling efficiencies onto optical fibers. Here, we demonstrate resonantly enhanced second-harmonic generation (SHG) through the all-fiber integration of a gallium nitride (GaN) hole-type circular Bragg grating (h-CBG) cavity, directly transferred onto a standard optical fiber. Together with the large second-order nonlinear susceptibility and wide optical transparency window of GaN, the fabricated h-CBG membrane cavity on GaN enables strong field confinement and vertically directional out-coupling of the generated SHG signal. As a result, we observe drastically enhanced SHG signals from the h-CBG device compared with the bulk GaN and the unpatterned freestanding GaN membrane. Using a deterministic pick-and-place transfer technique, we demonstrate robust and precise fiber integration of the GaN cavity device, enabling in-line SHG generation from a conventional fiber platform. This work establishes a compact and scalable approach for incorporating optical nonlinearity into fiber-based photonic systems.

physics.optics

Test-Beam Performance of the AstroPix Silicon Sensor for Imaging Calorimetry

AstroPix is a high-voltage CMOS HVCMOS monolithic active pixel sensor MAPS developed for future space-based gamma-ray missions. It is also a candidate technology for the imaging layer of the Barrel Imaging Calorimeter BIC in the ePIC experiment at the future Electron-Ion Collider EIC. We report the first AstroPix test-beam results obtained at the KEK Photon Factory Advanced Ring PF-AR and the CERN Proton Synchrotron PS T10 beam line in 2025, using the third prototype AstroPix-v3. AstroPix-v3 sensors were operated as both standalone tracking layers and imaging layers interleaved with prototype lead/scintillating-fiber Pb/SciFi calorimeter modules, using electron and hadron beams in the few-GeV/c momentum range. Event synchronization between the continuous readout of AstroPix-v3 and the trigger-based readout of the Pb/SciFi calorimeter was achieved using a common timestamp. The AstroPix-v3 sensors exhibit stable performance, reaching a maximum hit efficiency of 68 percent at a bias voltage of -400 V under pion-dominated beam conditions. When combined with the Pb/SciFi calorimeter, the AstroPix layers successfully capture the development of electromagnetic showers. Using Cherenkov-based particle identification, electron-induced events exhibit significantly higher hit multiplicities and broader spatial distributions than pion-induced events, thereby providing clear discrimination between electromagnetic and hadronic showers. These results demonstrate that AstroPix-v3 provides effective, high-granularity imaging of shower development and is well suited as an imaging layer in future calorimeter systems for both collider and space-based experiments.

physics.ins-det

FINEST: Improving LLM Responses to Sensitive Topics Through Fine-Grained Evaluation

Large Language Models (LLMs) often generate overly cautious and vague responses on sensitive topics, sacrificing helpfulness for safety. Existing evaluation frameworks lack systematic methods to identify and address specific weaknesses in responses to sensitive topics, making it difficult to improve both safety and helpfulness simultaneously. To address this, we introduce FINEST, a FINE-grained response evaluation taxonomy for Sensitive Topics, which breaks down helpfulness and harmlessness into errors across three main categories: Content, Logic, and Appropriateness. Experiments on a Korean-sensitive question dataset demonstrate that our score- and error-based improvement pipeline, guided by FINEST, significantly improves the model responses across all three categories, outperforming refinement without guidance. Notably, score-based improvement -- providing category-specific scores and justifications -- yields the most significant gains, reducing the error sentence ratio for Appropriateness by up to 33.09%. This work lays the foundation for a more explainable and comprehensive evaluation and improvement of LLM responses to sensitive questions.

cs.CL

Evaluating the Pre-Consultation Ability of LLMs using Diagnostic Guidelines

We introduce EPAG, a benchmark dataset and framework designed for Evaluating the Pre-consultation Ability of LLMs using diagnostic Guidelines. LLMs are evaluated directly through HPI-diagnostic guideline comparison and indirectly through disease diagnosis. In our experiments, we observe that small open-source models fine-tuned with a well-curated, task-specific dataset can outperform frontier LLMs in pre-consultation. Additionally, we find that increased amount of HPI (History of Present Illness) does not necessarily lead to improved diagnostic performance. Further experiments reveal that the language of pre-consultation influences the characteristics of the dialogue. By open-sourcing our dataset and evaluation pipeline on https://github.com/seemdog/EPAG, we aim to contribute to the evaluation and further development of LLM applications in real-world clinical settings.

cs.CL

GSFeatLoc: Visual Localization Using Feature Correspondence on 3D Gaussian Splatting

In this paper, we present a method for localizing a query image with respect to a precomputed 3D Gaussian Splatting (3DGS) scene representation. First, the method uses 3DGS to render a synthetic RGBD image at some initial pose estimate. Second, it establishes 2D-2D correspondences between the query image and this synthetic image. Third, it uses the depth map to lift the 2D-2D correspondences to 2D-3D correspondences and solves a perspective-n-point (PnP) problem to produce a final pose estimate. Results from evaluation across three existing datasets with 38 scenes and over 2,700 test images show that our method significantly reduces both inference time (by over two orders of magnitude, from more than 10 seconds to as fast as 0.1 seconds) and estimation error compared to baseline methods that use photometric loss minimization. Results also show that our method tolerates large errors in the initial pose estimate of up to 55{\deg} in rotation and 1.1 units in translation (normalized by scene scale), achieving final pose errors of less than 5{\deg} in rotation and 0.05 units in translation on 90% of images from the Synthetic NeRF and Mip-NeRF360 datasets and on 42% of images from the more challenging Tanks and Temples dataset.

cs.CV

Efficient Extrinsic Self-Calibration of Multiple IMUs using Measurement Subset Selection

This paper addresses the problem of choosing a sparse subset of measurements for quick calibration parameter estimation. A standard solution to this is selecting a measurement only if its utility -- the difference between posterior (with the measurement) and prior information (without the measurement) -- exceeds some threshold. Theoretically, utility, a function of the parameter estimate, should be evaluated at the estimate obtained with all measurements selected so far, hence necessitating a recalibration with each new measurement. However, we hypothesize that utility is insensitive to changes in the parameter estimate for many systems of interest, suggesting that evaluating utility at some initial parameter guess would yield equivalent results in practice. We provide evidence supporting this hypothesis for extrinsic calibration of multiple inertial measurement units (IMUs), showing the reduction in calibration time by two orders of magnitude by forgoing recalibration for each measurement.

cs.RO

How language models extrapolate outside the training data: A case study in Textualized Gridworld

Language models' ability to extrapolate learned behaviors to novel, more complex environments beyond their training scope is highly unknown. This study introduces a path planning task in a textualized Gridworld to probe language models' extrapolation capabilities. We show that conventional approaches, including next token prediction and Chain of Thought (CoT) finetuning, fail to extrapolate in larger, unseen environments. Inspired by human cognition and dual process theory, we propose cognitive maps for path planning, a novel CoT framework that simulates humanlike mental representations. Our experiments show that cognitive maps not only enhance extrapolation to unseen environments but also exhibit humanlike characteristics through structured mental simulation and rapid adaptation. Our finding that these cognitive maps require specialized training schemes and cannot be induced through simple prompting opens up important questions about developing general-purpose cognitive maps in language models. Our comparison with exploration-based methods further illuminates the complementary strengths of offline planning and online exploration.

cs.CL

The Use of Multi-Scale Fiducial Markers To Aid Takeoff and Landing Navigation by Rotorcraft

This paper quantifies the performance of visual SLAM that leverages multi-scale fiducial markers (i.e., artificial landmarks that can be detected at a wide range of distances) to show its potential for reliable takeoff and landing navigation in rotorcraft. Prior work has shown that square markers with a black-and-white pattern of grid cells can be used to improve the performance of visual SLAM with color cameras. We extend this prior work to allow nested marker layouts. We evaluate performance during semi-autonomous takeoff and landing operations in a variety of environmental conditions by a DJI Matrice 300 RTK rotorcraft with two FLIR Blackfly color cameras, using RTK GNSS to obtain ground truth pose estimates. Performance measures include absolute trajectory error and the fraction of the number of estimated poses to the total frame. We release all of our results -- our dataset and the code of the implementation of the visual SLAM with fiducial markers -- to the public as open-source.

cs.CV

Comparative Study of Visual SLAM-Based Mobile Robot Localization Using Fiducial Markers

This paper presents a comparative study of three modes for mobile robot localization based on visual SLAM using fiducial markers (i.e., square-shaped artificial landmarks with a black-and-white grid pattern): SLAM, SLAM with a prior map, and localization with a prior map. The reason for comparing the SLAM-based approaches leveraging fiducial markers is because previous work has shown their superior performance over feature-only methods, with less computational burden compared to methods that use both feature and marker detection without compromising the localization performance. The evaluation is conducted using indoor image sequences captured with a hand-held camera containing multiple fiducial markers in the environment. The performance metrics include absolute trajectory error and runtime for the optimization process per frame. In particular, for the last two modes (SLAM and localization with a prior map), we evaluate their performances by perturbing the quality of prior map to study the extent to which each mode is tolerant to such perturbations. Hardware experiments show consistent trajectory error levels across the three modes, with the localization mode exhibiting the shortest runtime among them. Yet, with map perturbations, SLAM with a prior map maintains performance, while localization mode degrades in both aspects.

cs.RO

Dispersive Decay Bound of Small Data Solutions to Kawahara Equation in a Finite Time Scale

In this article, we prove that small localized data yield solutions to Kawahara type equation which have linear dispersive decay on a finite time. We use the similar method used to derive the dispersive decay bound of the solutions to the KdV equation, with some steps being simpler. This result is expected to be the first result of the small data global bounds of the fifth-order dispersive equations with quadratic nonlinearity.

math.AP

Extrinsic Calibration of Multiple Inertial Sensors from Arbitrary Trajectories

We present a method of extrinsic calibration for a system of multiple inertial measurement units (IMUs) that estimates the relative pose of each IMU on a rigid body using only measurements from the IMUs themselves, without the need to prescribe the trajectory. Our method is based on solving a nonlinear least-squares problem that penalizes inconsistency between measurements from pairs of IMUs. We validate our method with experiments both in simulation and in hardware. In particular, we show that it meets or exceeds the performance -- in terms of error, success rate, and computation time -- of an existing, state-of-the-art method that does not rely only on IMU measurements and instead requires the use of a camera and a fiducial marker. We also show that the performance of our method is largely insensitive to the choice of trajectory along which IMU measurements are collected.

cs.RO

KOLD: Korean Offensive Language Dataset

Recent directions for offensive language detection are hierarchical modeling, identifying the type and the target of offensive language, and interpretability with offensive span annotation and prediction. These improvements are focused on English and do not transfer well to other languages because of cultural and linguistic differences. In this paper, we present the Korean Offensive Language Dataset (KOLD) comprising 40,429 comments, which are annotated hierarchically with the type and the target of offensive language, accompanied by annotations of the corresponding text spans. We collect the comments from NAVER news and YouTube platform and provide the titles of the articles and videos as the context information for the annotation process. We use these annotated comments as training data for Korean BERT and RoBERTa models and find that they are effective at offensiveness detection, target classification, and target span detection while having room for improvement for target group classification and offensive span detection. We discover that the target group distribution differs drastically from the existing English datasets, and observe that providing the context information improves the model performance in offensiveness detection (+0.3), target classification (+1.5), and target group classification (+13.1). We publicly release the dataset and baseline models.

cs.CL

You Only Need One Model for Open-domain Question Answering

Recent approaches to Open-domain Question Answering refer to an external knowledge base using a retriever model, optionally rerank passages with a separate reranker model and generate an answer using another reader model. Despite performing related tasks, the models have separate parameters and are weakly-coupled during training. We propose casting the retriever and the reranker as internal passage-wise attention mechanisms applied sequentially within the transformer architecture and feeding computed representations to the reader, with the hidden representations progressively refined at each stage. This allows us to use a single question answering model trained end-to-end, which is a more efficient use of model capacity and also leads to better gradient flow. We present a pre-training method to effectively train this architecture and evaluate our model on the Natural Questions and TriviaQA open datasets. For a fixed parameter budget, our model outperforms the previous state-of-the-art model by 1.0 and 0.7 exact match scores.

cs.CL

KLUE: Korean Language Understanding Evaluation

We introduce Korean Language Understanding Evaluation (KLUE) benchmark. KLUE is a collection of 8 Korean natural language understanding (NLU) tasks, including Topic Classification, SemanticTextual Similarity, Natural Language Inference, Named Entity Recognition, Relation Extraction, Dependency Parsing, Machine Reading Comprehension, and Dialogue State Tracking. We build all of the tasks from scratch from diverse source corpora while respecting copyrights, to ensure accessibility for anyone without any restrictions. With ethical considerations in mind, we carefully design annotation protocols. Along with the benchmark tasks and data, we provide suitable evaluation metrics and fine-tuning recipes for pretrained language models for each task. We furthermore release the pretrained language models (PLM), KLUE-BERT and KLUE-RoBERTa, to help reproducing baseline models on KLUE and thereby facilitate future research. We make a few interesting observations from the preliminary experiments using the proposed KLUE benchmark suite, already demonstrating the usefulness of this new benchmark suite. First, we find KLUE-RoBERTa-large outperforms other baselines, including multilingual PLMs and existing open-source Korean PLMs. Second, we see minimal degradation in performance even when we replace personally identifiable information from the pretraining corpus, suggesting that privacy and NLU capability are not at odds with each other. Lastly, we find that using BPE tokenization in combination with morpheme-level pre-tokenization is effective in tasks involving morpheme-level tagging, detection and generation. In addition to accelerating Korean NLP research, our comprehensive documentation on creating KLUE will facilitate creating similar resources for other languages in the future. KLUE is available at https://klue-benchmark.com.

cs.CL

Alternative Pathways for Multiple Exciton Generation Solar Cells by Tandem Configurations

Multiple exciton generation solar cells (MEGSCs) undergo low efficiency due to material imperfections such as nonradiative recombination This paper introduces alternative approaches for realizing photovoltaic (PV) devices similar to MEGSCs. Furthermore, we reorganize the detailed balance equation of MEGSCs such that it is similar to that of independent connection tandem solar cells. This is possible because of the spectral dependence of the ideal QY. Finally, we compare these two similar equations and propose alternative approaches for realizing MEGSC-like tandem solar cells. We explain the difficulty in fabricating MEGSCs, which arises from the high rate of non-idealities. In this regard, the deconstruction of the detailed balance equation of MEGSCs can reveal alternative paths for replacing MEGSCs with tandem solar cell configurations.

physics.app-ph

Multiple Exciton Generation Solar Cells: Numerical Approach of Quantum Yield Extraction and its Limiting Efficiencies

Multiple exciton generation solar cells exhibit a low power conversion efficiency owing to nonradiative recombination even if numerous electron and hole pairs are generated per incident photon. This paper elucidates the non-idealities of multiple exciton generation solar cells (MEGSCs) and alternative approaches for realizing photovoltaic (PV) devices similar to MEGSCs. First, we present mathematical approaches for determining the quantum yield (QY) to discuss the non-idealities of MEGSCs by adjusting the delta function. In particular, we employ the Gaussian distribution function to present the occupancy status of carriers at each energy state by Dirac delta function. By adjusting the Gaussian distribution function for each energy state, we obtain the ideal and non-ideal QYs. Through this approach, we discuss the material imperfections of MEGSCs by analyzing the mathematically obtained QYs. By calculating the ratio between the radiative and nonradiative recombination, we can discuss the status of radiative recombination calculate Furthermore, we apply this approach into the detailed balance limit of MEGSC to investigate the practical limit of MEGSC.

physics.app-ph