arXiv ScienceSearch

arXiv subjects

Mei Liu

Publications and source records attributed to Mei Liu.

28 records · Page 2Linked to original sources

Contrastive Learning of Temporal Distinctiveness for Survival Analysis in Electronic Health Records

Survival analysis plays a crucial role in many healthcare decisions, where the risk prediction for the events of interest can support an informative outlook for a patient's medical journey. Given the existence of data censoring, an effective way of survival analysis is to enforce the pairwise temporal concordance between censored and observed data, aiming to utilize the time interval before censoring as partially observed time-to-event labels for supervised learning. Although existing studies mostly employed ranking methods to pursue an ordering objective, contrastive methods which learn a discriminative embedding by having data contrast against each other, have not been explored thoroughly for survival analysis. Therefore, in this paper, we propose a novel Ontology-aware Temporality-based Contrastive Survival (OTCSurv) analysis framework that utilizes survival durations from both censored and observed data to define temporal distinctiveness and construct negative sample pairs with adjustable hardness for contrastive learning. Specifically, we first use an ontological encoder and a sequential self-attention encoder to represent the longitudinal EHR data with rich contexts. Second, we design a temporal contrastive loss to capture varying survival durations in a supervised setting through a hardness-aware negative sampling mechanism. Last, we incorporate the contrastive task into the time-to-event predictive task with multiple loss components. We conduct extensive experiments using a large EHR dataset to forecast the risk of hospitalized patients who are in danger of developing acute kidney injury (AKI), a critical and urgent medical condition. The effectiveness and explainability of the proposed model are validated through comprehensive quantitative and qualitative studies.

cs.LG

An Open Natural Language Processing Development Framework for EHR-based Clinical Research: A case demonstration using the National COVID Cohort Collaborative (N3C)

While we pay attention to the latest advances in clinical natural language processing (NLP), we can notice some resistance in the clinical and translational research community to adopt NLP models due to limited transparency, interpretability, and usability. In this study, we proposed an open natural language processing development framework. We evaluated it through the implementation of NLP algorithms for the National COVID Cohort Collaborative (N3C). Based on the interests in information extraction from COVID-19 related clinical notes, our work includes 1) an open data annotation process using COVID-19 signs and symptoms as the use case, 2) a community-driven ruleset composing platform, and 3) a synthetic text data generation workflow to generate texts for information extraction tasks without involving human subjects. The corpora were derived from texts from three different institutions (Mayo Clinic, University of Kentucky, University of Minnesota). The gold standard annotations were tested with a single institution's (Mayo) ruleset. This resulted in performances of 0.876, 0.706, and 0.694 in F-scores for Mayo, Minnesota, and Kentucky test datasets, respectively. The study as a consortium effort of the N3C NLP subgroup demonstrates the feasibility of creating a federated NLP algorithm development and benchmarking platform to enhance multi-institution clinical NLP study and adoption. Although we use COVID-19 as a use case in this effort, our framework is general enough to be applied to other domains of interest in clinical NLP.

cs.CL

Activated Gradients for Deep Neural Networks

Deep neural networks often suffer from poor performance or even training failure due to the ill-conditioned problem, the vanishing/exploding gradient problem, and the saddle point problem. In this paper, a novel method by acting the gradient activation function (GAF) on the gradient is proposed to handle these challenges. Intuitively, the GAF enlarges the tiny gradients and restricts the large gradient. Theoretically, this paper gives conditions that the GAF needs to meet, and on this basis, proves that the GAF alleviates the problems mentioned above. In addition, this paper proves that the convergence rate of SGD with the GAF is faster than that without the GAF under some assumptions. Furthermore, experiments on CIFAR, ImageNet, and PASCAL visual object classes confirm the GAF's effectiveness. The experimental results also demonstrate that the proposed method is able to be adopted in various deep neural networks to improve their performance. The source code is publicly available at https://github.com/LongJin-lab/Activated-Gradients-for-Deep-Neural-Networks.

cs.CV

Deforming the Loss Surface to Affect the Behaviour of the Optimizer

In deep learning, it is usually assumed that the optimization process is conducted on a shape-fixed loss surface. Differently, we first propose a novel concept of deformation mapping in this paper to affect the behaviour of the optimizer. Vertical deformation mapping (VDM), as a type of deformation mapping, can make the optimizer enter a flat region, which often implies better generalization performance. Moreover, we design various VDMs, and further provide their contributions to the loss surface. After defining the local M region, theoretical analyses show that deforming the loss surface can enhance the gradient descent optimizer's ability to filter out sharp minima. With visualizations of loss landscapes, we evaluate the flatnesses of minima obtained by both the original optimizer and optimizers enhanced by VDMs on CIFAR-100. The experimental results show that VDMs do find flatter regions. Moreover, we compare popular convolutional neural networks enhanced by VDMs with the corresponding original ones on ImageNet, CIFAR-10, and CIFAR-100. The results are surprising: there are significant improvements on all of the involved models equipped with VDMs. For example, the top-1 test accuracy of ResNet-20 on CIFAR-100 increases by 1.46%, with insignificant additional computational overhead.

cs.LG

Deforming the Loss Surface

In deep learning, it is usually assumed that the shape of the loss surface is fixed. Differently, a novel concept of deformation operator is first proposed in this paper to deform the loss surface, thereby improving the optimization. Deformation function, as a type of deformation operator, can improve the generalization performance. Moreover, various deformation functions are designed, and their contributions to the loss surface are further provided. Then, the original stochastic gradient descent optimizer is theoretically proved to be a flat minima filter that owns the talent to filter out the sharp minima. Furthermore, the flatter minima could be obtained by exploiting the proposed deformation functions, which is verified on CIFAR-100, with visualizations of loss landscapes near the critical points obtained by both the original optimizer and optimizer enhanced by deformation functions. The experimental results show that deformation functions do find flatter regions. Moreover, on ImageNet, CIFAR-10, and CIFAR-100, popular convolutional neural networks enhanced by deformation functions are compared with the corresponding original models, where significant improvements are observed on all of the involved models equipped with deformation functions. For example, the top-1 test accuracy of ResNet-20 on CIFAR-100 increases by 1.46%, with insignificant additional computational overhead.

cs.CV

COVID-19 SignSym: a fast adaptation of a general clinical NLP tool to identify and normalize COVID-19 signs and symptoms to OMOP common data model

The COVID-19 pandemic swept across the world rapidly, infecting millions of people. An efficient tool that can accurately recognize important clinical concepts of COVID-19 from free text in electronic health records (EHRs) will be valuable to accelerate COVID-19 clinical research. To this end, this study aims at adapting the existing CLAMP natural language processing tool to quickly build COVID-19 SignSym, which can extract COVID-19 signs/symptoms and their 8 attributes (body location, severity, temporal expression, subject, condition, uncertainty, negation, and course) from clinical text. The extracted information is also mapped to standard concepts in the Observational Medical Outcomes Partnership common data model. A hybrid approach of combining deep learning-based models, curated lexicons, and pattern-based rules was applied to quickly build the COVID-19 SignSym from CLAMP, with optimized performance. Our extensive evaluation using 3 external sites with clinical notes of COVID-19 patients, as well as the online medical dialogues of COVID-19, shows COVID-19 Sign-Sym can achieve high performance across data sources. The workflow used for this study can be generalized to other use cases, where existing clinical natural language processing tools need to be customized for specific information needs within a short time. COVID-19 SignSym is freely accessible to the research community as a downloadable package (https://clamp.uth.edu/covid/nlp.php) and has been used by 16 healthcare organizations to support clinical research of COVID-19.

cs.CL

Partial-fraction Expansion of Lossless Negative Imaginary Property and A Generalized Lossless Negative Imaginary Lemma

This paper studies a partial-fraction expansion for lossless negative imaginary systems and presents a generalized lossless negative imaginary lemma by allowing poles at zero. First, a necessary and sufficient condition for a system to be non-proper lossless negative imaginary is developed, and a minor partial-fraction expansion of lossless negative imaginary property is studied. Second, according to the minor decomposition properties, two different and new relationships between lossless positive real and lossless negative imaginary systems are established. Third, according to one of the relationships, a generalized lossless negative imaginary lemma in terms of a minimal state-space realization is derived by allowing poles at zero. Some important properties of lossless negative imaginary systems are also studied in this paper, and three numerical examples are provided to illustrate the developed theory

math.OC

Exciton-Polaritons in Hybrid Inorganic-organic Perovskite Fabry-P\'erot Microcavity

Exciton-polaritons in semiconductor microcavities generate fascinating effects such as long-range spatial coherence and Bose-Einstein Condensation (BEC), which are attractive for their potential use in low threshold lasers, vortices and slowing light, etc. However, currently most of exciton-polariton effects either occur at cryogenic temperature or rely on expensive cavity fabrication procedures. Further exploring new semiconductor microcavities with stronger exciton photon interaction strength is extensively needed. Herein, we demonstrate room temperature photon exciton strong coupling in hybrid inorganic-organic CH3NH3PbBr3 Fabry-P\'erot microcavities for the first time. The vacuum Rabi splitting energy is up to ~390 meV, which is ascribed to large oscillator strength and photon confinement in reduced dimension of optical microcavities. With increasing pumping energy, exciton-photon coupling strength is weakened due to carrier screening effect, leading to occurrence of photonic lasing instead of polartion lasing. The demonstrated strong coupling between photons and excitons in perovskite microcavities would be helpful for development of high performance polariton-based incoherent and coherent light sources, nonlinear optics, and slow light applications.

cond-mat.mes-hall

Investigation of GEM-Micromegas Detector on X-ray Beam of Synchrotron Radiation

To solve the discharge of the standard Bulk Micromegas and GEM detector, the GEM-Micromegas detector was developed in Institute of High Energy Physics. Taking into account the advantages of the two detectors, one GEM foil was set as a preamplifier on the mesh of Micromegas in the structure and the GEM preamplification decreased the working voltage of Micromegas to reduce the effect of the discharge significantly. In the paper, the performance of detector in X-ray beam was studied at 1W2B laboratory of Beijing Synchrotron Radiation Facility. Finally, the result of the energy resolution under various X-ray energies was given in different working gases. It indicated that the GEM-Micromegas detector had the energy response capability in all the energy range and it could work better than the standard Bulk-Micromegas.

physics.ins-det