arXiv ScienceSearch

arXiv subjects

Bofan Song

Publications and source records attributed to Bofan Song.

4 recordsLinked to original sources

Explainable Multimodal Deep Learning Integrating Imaging and Clinical Data for Oral Potentially Malignant Disorder Detection

Oral potentially malignant disorders (OPMDs) are critical precursors to oral cancer, yet clinical detection remains challenging because of substantial phenotypic heterogeneity and overlap with benign conditions. Although image-based deep learning shows promise for automated screening, visual information alone may be insufficient in real-world settings, where diagnostic decisions also rely on patient-specific risk factors. We developed M2-OPMDNet, a multimodal deep learning framework that integrates co-registered white-light and autofluorescence intraoral images with structured clinical information for OPMD detection. A customized questionnaire was designed to capture clinically relevant risk factors and symptoms in a standardized, reproducible format for integration with image-derived features. Multiple image encoders, including conventional convolutional neural networks and foundation model-based architectures, were evaluated using a prospectively collected dataset reflecting real-world screening conditions. Model interpretability was assessed using SHapley Additive exPlanations (SHAP) to quantify feature- and modality-level contributions. M2-OPMDNet achieved an AUC of 0.952, outperforming unimodal approaches and showing improved performance for visually subtle lesions. SHAP analysis demonstrated that structured clinical variables contributed substantially to risk estimation and complemented imaging features. These results demonstrate that explainable multimodal learning combining white-light and autofluorescence imaging with structured clinical data can provide accurate, transparent, and clinically grounded OPMD detection. M2-OPMDNet offers a scalable framework for real-world oral cancer screening and decision support.

eess.IV

Fully Fiber-Integrated 3D-Printed Probes for Plug-and-Play Endoscopic Optical Coherence Tomography

Miniature probes extend optical coherence tomography (OCT) into lumens and other confined spaces, but conventional implementations often rely on multiple distal components, fiber processing, and probe-specific interferometer matching. Here, a fully fiber-integrated OCT (F2I-OCT) architecture is demonstrated in which two-photon microfabrication defines not only the terminal imaging optic, but the complete distal optical and interferometric architecture within a single fiber-mounted element. Beam expansion, side-view redirection, common-path reference generation, and terminal imaging are physically integrated while remaining independently designable. This architecture transfers complexity from component fabrication and interferometer matching into three-dimensional optical design, reducing assembly to a print-and-bond process and enabling plug-and-play exchange on the same OCT platform. Terminal optics can be adapted to different working distances, surrounding media, and wavefront transformations without redesigning the upstream architecture. Twenty assembled probes exhibit returned-reference and side-viewing-output power standard deviations of 0.149 dB and 0.071 dB, respectively. The system achieves 93.1 dB sensitivity and enables ex vivo imaging of airway and dental structures. The combination of a common plug-and-play architecture with broad optical design freedom provides a versatile platform for endoscopic and confined-space imaging across diverse biomedical and technical environments.

physics.optics

Label-Free Deep-Tissue Peripheral Nerve Detection with a Handheld Multimodal OCT Probe and NerveDetNet

Peripheral nerves buried beneath intact tissue are difficult to visualize during surgery and remain inaccessible to white light wide-field imaging and other surface optical imaging methods. Existing OCT nerve studies have largely relied on exposed nerves or polarization contrast with limited depth penetration, restricting their value for subsurface intraoperative guidance. Here, we introduce, to our knowledge, the first label-free framework for detecting peripheral nerves beneath unopened tissue and resolving their depth using intensity-based OCT structural signatures alone. The framework combines a handheld multimodal probe, integrating swept-source OCT with co-registered white light and autofluorescence imaging, with a ``confirm-then-capture'' workflow designed for practical surgical use. To enable efficient analysis of sparsely sampled OCT volumes, we develop NerveDetNet, a lightweight 2.5D segmentation network that recovers weak and spatially displaced nerve signals by incorporating spatial context, frame-order information, and shift-tolerant correlations across frames through a dedicated nerve feature correlation module. In ex vivo tissue experiments, NerveDetNet consistently outperformed six representative 2D baselines across all frame spacings, achieving a Dice score of 0.725 under the sparsest sampling condition while using approximately half the model parameters. End-to-end validation demonstrated localization of nerves invisible at the surface and depth-resolved detection up to 1.3--1.4~mm below the tissue surface, with OCT derived depth maps overlaid directly onto the surgical view. Together, these results establish a practical label-free approach for subsurface nerve visualization that supports intraoperative compatibility, enables efficient sparse-volume analysis, and provides depth-resolved guidance without tissue opening, contrast agents, or nerve exposure.

physics.optics

Structured light dark-field microscope

A resolution-enhanced dark-field microscope by structured light illumination is proposed to improve resolution and contrast. A set of phase-shifted fringes are projected to the sample plane at large angle to capture modulated dark-field images, from which resolution- and contrast-enhanced dark-field image, as well as sectioned dark-field image, can be obtained. Human tissue samples are tested to demonstrate the resolution and contrast enhancement. The system can be implemented in transmission-mode and reflectance-mode, with potential applications ranging from defect detection to biomedical imaging.

eess.IV