arXiv ScienceSearch

arXiv subjects

Insung Hwang

Publications and source records attributed to Insung Hwang.

9 recordsLinked to original sources

WSPolypNet: Weakly Supervised Polyp Localization in Colonoscopy Videos

Because dense frame-level annotation of colonoscopy videos is costly, we propose WSPolypNet, a weakly supervised framework for polyp localization using only video-level labels. WSPolypNet employs a 3D convolutional neural network trained with video-level supervision to generate class activation maps (CAMs), which identify candidate polyp regions without requiring frame-level spatial annotations. The CAM-derived localization cues are further enhanced using a multi-view strategy and provided to MedSAM2 as point prompts. MedSAM2 then propagates segmentation masks across the video, refining the coarse localization cues according to polyp boundaries. WSPolypNet achieved CorLoc scores of 47.80%, 43.68%, and 35.01% at IoU thresholds of 0.3, 0.5, and 0.7, respectively, compared with 36.87%, 33.72%, and 27.94% in the single-view setting. For small polyps, the multi-view strategy improved CorLoc@0.5 from 16.01% to 30.97%. The framework also achieved a recall of 94.51%. These results demonstrate the potential of weakly supervised spatiotemporal learning to substantially reduce spatial annotation requirements for polyp localization in colonoscopy videos.

cs.CV

Readout electronics for SUBMET

A dedicated data acquisition (DAQ) system has been developed for the SUB-Millicharge ExperimenT (SUBMET) at the Japan Proton Accelerator Research Complex (J-PARC), a search for particles carrying a fractional electric charge $Q = \epsilon e$ with $\epsilon$ below $\mathcal{O}(10^{-3})$, hereafter referred to as millicharged particles (mCPs). Because such particles are expected to produce at most a few scintillation photons, the system is optimized for single-photoelectron detection from the photomultiplier tubes (PMTs), combining high-speed waveform digitization with precise timing. To capture eight consecutive proton bunches of the 30 GeV J-PARC beam within a single trigger, the eight channels of the Domino Ring Sampler 4 (DRS4) chip are cascaded in groups of four to form two readout inputs, each sampling 4096 points continuously at 820.5 MHz over an effective time window of 5 us. After calibration, timing differences between channels are within 1 ns on the same DRS4 chip, 2 ns on the same board, and 8 ns across different boards, well within the 30 ns coincidence window of the experiment. The front-end electronics achieve an RMS noise below 0.4 mV. The baseline is deliberately offset upward such that the negative-going pulses span a larger fraction of the digitizer range, improving voltage resolution and dynamic range. A trigger control board aggregates data from multiple readout boards and sustains the data-transfer rate required for beam operation. The measured performance confirms that the DAQ system meets the timing, noise, and throughput requirements of the experiment.

hep-ex

NEMESIS: Noise-suppressed Efficient MAE with Enhanced Superpatch Integration Strategy

Volumetric CT imaging is essential for clinical diagnosis, yet annotating 3D volumes is expensive and time-consuming, motivating self-supervised learning (SSL) from unlabeled data. However, applying SSL to 3D CT remains challenging due to the high memory cost of full-volume transformers and the anisotropic spatial structure of CT data, which is not well captured by conventional masking strategies. We propose NEMESIS, a masked autoencoder (MAE) framework that operates on local 128x128x128 superpatches, enabling memory-efficient training while preserving anatomical detail. NEMESIS introduces three key components: (i) noise-enhanced reconstruction as a pretext task, (ii) Masked Anatomical Transformer Blocks (MATB) that perform dual-masking through parallel plane-wise and axis-wise token removal, and (iii) NEMESIS Tokens (NT) for cross-scale context aggregation. On the BTCV multi-organ classification benchmark, NEMESIS with a frozen backbone and a linear classifier achieves a mean AUROC of 0.9633, surpassing fully fine-tuned SuPreM (0.9493) and VoCo (0.9387). Under a low-label regime with only 10% of available annotations, it retains an AUROC of 0.9075, demonstrating strong label efficiency. Furthermore, the superpatch-based design reduces computational cost to 31.0 GFLOPs per forward pass, compared to 985.8 GFLOPs for the full-volume baseline, providing a scalable and robust foundation for 3D medical imaging.

cs.CV

Dedicated Searches for Millicharged Particles at Intensity-Frontier Facilities: SpinQuest and SHiP

We conduct a dedicated study of searches for millicharged particles (mCPs) utilizing scintillator-based detectors at high-intensity fixed-target experiments, with particular focus on the SpinQuest and forthcoming Search for Hidden Particles experiment (SHiP) facilities. The analysis incorporates the three primary production mechanisms: meson decays, Drell-Yan processes, and proton bremsstrahlung. In particular, our updated analysis reveals that proton bremsstrahlung dominates the production rate in the sub-GeV mass regime. Detailed detector simulations and background evaluations are performed to obtain realistic sensitivity estimates. Our results demonstrate that future experiments located in the SpinQuest and SHiP facilities can achieve substantial improvements in discovery potential, enhancing sensitivity to the mCP charge parameter $\epsilon=q_{\chi}/e$ (with $q_\chi$ denoting the mCP electric charge) by up to two orders of magnitude relative to existing limits.

hep-ph

Afterpulse prediction for SUBMET experiment

The SUB-Millicharge ExperimenT (SUBMET) investigates an unexplored parameter space of millicharged particles with mass $m_\chi < $ 1.6 GeV/c$^2$ and charge $Q_\chi < 10^{-3}e$. The detector consists of an Eljen-200 plastic scintillator coupled to a Hamamatsu Photonics R7725 photomultiplier tube (PMT). PMT afterpulses, delayed pulses produced after an energetic pulse, have been observed in the SUBMET readout system, especially following primary pulses with a large area. We present a prediction method for afterpulse rates based on measurable parameters, which reproduces the observed rate with approximately 20\% precision. This approach enables a better understanding of afterpulse contributions and, consequently, improves the reliability of background predictions.

physics.ins-det

Design and Mechanical Integration of Scintillation Modules for SUB-Millicharge ExperimenT (SUBMET)

We present a detailed description of the detector design for the SUB-Millicharge ExperimenT (SUBMET), developed to search for millicharged particles. The experiment probes a largely unexplored region of the charge-mass parameter space, focusing on particles with mass $m_\chi < 1.6~\textrm{GeV}/c^2$ and electric charge $Q < 10^{-3}e$. The detector has been optimized to achieve high sensitivity to interactions of such particles while maintaining effective discrimination against background events. We provide a comprehensive overview of the key detector components, including scintillation modules, photomultiplier tubes, and the mechanical support structure.

physics.ins-det

LANSCE-mQ: Dedicated search for milli/fractionally charged particles at LANL

In this paper, we propose an experiment, LANSCE-mQ, aiming to detect fractionally charged and millicharged particles (mCP) using an 800 MeV proton beam fixed target at the Los Alamos Neutron Science Center (LANSCE) facility. This search can shed new light on numerous fundamental questions, including charge quantization, the predictions of string theories and grand unification theories, the gauge symmetry of the Standard Model, dark sector models, and the tests of cosmic reheating. We propose to install two-layer scintillation detectors made of plastic (such as EJ-200) or CeBr3 to search for mCPs. Dedicated Geant4 detector simulations and in situ measurements have been conducted to obtain a preliminary determination of the background rate. The dominant backgrounds are beam-induced neutrons and coincident dark current signals from the photomultiplier tubes, while beam-induced gammas and cosmic muons are subdominant. We determined that LANSCE-mQ, the dedicated mCP experiment, has the leading mCP sensitivity for mass between ~ 1 MeV to 300 MeV.

hep-ph

Co-salient Object Detection Based on Deep Saliency Networks and Seed Propagation over an Integrated Graph

This paper presents a co-salient object detection method to find common salient regions in a set of images. We utilize deep saliency networks to transfer co-saliency prior knowledge and better capture high-level semantic information, and the resulting initial co-saliency maps are enhanced by seed propagation steps over an integrated graph. The deep saliency networks are trained in a supervised manner to avoid online weakly supervised learning and exploit them not only to extract high-level features but also to produce both intra- and inter-image saliency maps. Through a refinement step, the initial co-saliency maps can uniformly highlight co-salient regions and locate accurate object boundaries. To handle input image groups inconsistent in size, we propose to pool multi-regional descriptors including both within-segment and within-group information. In addition, the integrated multilayer graph is constructed to find the regions that the previous steps may not detect by seed propagation with low-level descriptors. In this work, we utilize the useful complementary components of high-, low-level information, and several learning-based steps. Our experiments have demonstrated that the proposed approach outperforms comparable co-saliency detection methods on widely used public databases and can also be directly applied to co-segmentation tasks.

cs.CV

A New Convolutional Network-in-Network Structure and Its Applications in Skin Detection, Semantic Segmentation, and Artifact Reduction

The inception network has been shown to provide good performance on image classification problems, but there are not much evidences that it is also effective for the image restoration or pixel-wise labeling problems. For image restoration problems, the pooling is generally not used because the decimated features are not helpful for the reconstruction of an image as the output. Moreover, most deep learning architectures for the restoration problems do not use dense prediction that need lots of training parameters. From these observations, for enjoying the performance of inception-like structure on the image based problems we propose a new convolutional network-in-network structure. The proposed network can be considered a modification of inception structure where pool projection and pooling layer are removed for maintaining the entire feature map size, and a larger kernel filter is added instead. Proposed network greatly reduces the number of parameters on account of removed dense prediction and pooling, which is an advantage, but may also reduce the receptive field in each layer. Hence, we add a larger kernel than the original inception structure for not increasing the depth of layers. The proposed structure is applied to typical image-to-image learning problems, i.e., the problems where the size of input and output are same such as skin detection, semantic segmentation, and compression artifacts reduction. Extensive experiments show that the proposed network brings comparable or better results than the state-of-the-art convolutional neural networks for these problems.

cs.CV