arXiv Science⌕ Search

arXiv subjects

Lawrence Y. L. Cheung

Publications and source records attributed to Lawrence Y. L. Cheung.

2 recordsLinked to original sources

Entropy-Guided Reverse-Causal AI to Identify Upstream Bottleneck Genes for Alzheimer's Drug Discovery

Identifying upstream regulators that connect several disease processes to therapeutic interventions is a central objective in Alzheimer's disease drug discovery. We propose an entropy-guided reverse-causal framework that makes candidate bottleneck genes the organizing link between disease mechanisms, pathways, molecular targets and drugs. The methodology integrates five stages: an Alzheimer's-specific knowledge graph with language-model assistance and expert review; reverse tracing from drugs to candidate genes; entropy-guided prioritization; forward propagation to drugs and complementary combinations; and staged validation with evidence feedback. The novelty lies in integrating upstream bottleneck identification, entropy-guided prioritization and iterative therapeutic selection within a dynamic, bidirectional discovery architecture. We demonstrate its molecular tracing and gene-prioritization components in a computational feasibility study using DeepDrug2 and MSigDB pathway annotations. Tracing amlodipine, indapamide and atorvastatin through a network of 11,300 molecular and drug nodes identifies 46 routes to nine genes. EGFR is the leading candidate, supported by 26 routes from all three drugs; MME and MAF rank next. These results show how pharmacological starting points can identify shared candidate genes with defined molecular connections. The framework's scientific significance lies in connecting convergent disease mechanisms to systematic intervention selection, with preservation of cognition and independence as the translational objective.

cs.CE↗

DECT: Harnessing LLM-assisted Fine-Grained Linguistic Knowledge and Label-Switched and Label-Preserved Data Generation for Diagnosis of Alzheimer's Disease

Alzheimer's Disease (AD) is an irreversible neurodegenerative disease affecting 50 million people worldwide. Low-cost, accurate identification of key markers of AD is crucial for timely diagnosis and intervention. Language impairment is one of the earliest signs of cognitive decline, which can be used to discriminate AD patients from normal control individuals. Patient-interviewer dialogues may be used to detect such impairments, but they are often mixed with ambiguous, noisy, and irrelevant information, making the AD detection task difficult. Moreover, the limited availability of AD speech samples and variability in their speech styles pose significant challenges in developing robust speech-based AD detection models. To address these challenges, we propose DECT, a novel speech-based domain-specific approach leveraging large language models (LLMs) for fine-grained linguistic analysis and label-switched label-preserved data generation. Our study presents four novelties: We harness the summarizing capabilities of LLMs to identify and distill key Cognitive-Linguistic information from noisy speech transcripts, effectively filtering irrelevant information. We leverage the inherent linguistic knowledge of LLMs to extract linguistic markers from unstructured and heterogeneous audio transcripts. We exploit the compositional ability of LLMs to generate AD speech transcripts consisting of diverse linguistic patterns to overcome the speech data scarcity challenge and enhance the robustness of AD detection models. We use the augmented AD textual speech transcript dataset and a more fine-grained representation of AD textual speech transcript data to fine-tune the AD detection model. The results have shown that DECT demonstrates superior model performance with an 11% improvement in AD detection accuracy on the datasets from DementiaBank compared to the baselines.

cs.CL↗