arXiv ScienceSearch

arXiv subjects

Daniel Rose

Publications and source records attributed to Daniel Rose.

9 recordsLinked to original sources

NEAT-POCKET: Pocket-Conditioned Autoregressive 3D Molecular Generation with a Neighborhood-Guided Set Transformer

AI-driven de novo molecular design offers a promising route to accelerate early-stage drug discovery by generating novel ligands directly within target protein binding pockets. We present NEAT-POCKET, a pocket-conditioned extension of the autoregressive NEAT model for 3D molecular generation. NEAT-POCKET generates molecules atom by atom in protein pocket environments while preserving atom permutation invariance and explicitly modeling hydrogen atoms. Benchmarks on the CrossDocked and SPINDR datasets show that NEAT-POCKET achieves competitive structure-based generation performance while sampling substantially faster than existing baselines. Beyond full-molecule generation, NEAT-POCKET naturally enables pocket-conditioned fragment completion, a task directly relevant to lead optimization and scaffold elaboration. These results position NEAT-POCKET as a fast, flexible, and practical framework for structure-based drug design.

cs.LG

BPMN4CAI: A BPMN Extension for Modeling Dynamic Conversational AI

Conversational AI systems, such as chatbots and virtual assistants, are becoming increasingly important to digital business processes. However, the established Business Process Model and Notation (BPMN) standard faces challenges when representing dynamic, context-sensitive interactions. This paper addresses this methodological and practical research gap by developing a standard-compliant BPMN extension (BPMN4CAI). Using Design Science Research methodology, this paper develops an approach that systematically extends existing BPMN elements and incorporates specialized components. The applicability and relevance of the BPMN4CAI framework are demonstrated and evaluated through a case study. The results show that the BPMN4CAI extension facilitates adaptive decision-making processes, robust context management, and transparent interactions for Conversational AI within business processes.

cs.AI

Disc Candidates in IC 2395: A WISE Survey of the Kinematically Confirmed Membership

Circumstellar disc dissipation timescales constrain the window available for planet formation. At approximately 9 Myr, IC 2395 lies at a critical epoch in primordial disc evolution, yet no kinematically selected disc census exists beyond the inner cluster core. A wide-field survey of IC 2395, extending to a 2 degree radius, establishes a kinematically selected disc census and evaluates the frequency of planet-forming environments in the outer cluster field. Using a high-purity catalogue of 173 members identified via Gaia DR3 kinematics, mid-infrared excesses are identified using AllWISE photometry. Candidates are classified into two tiers: Crossvalidated (corroborated by Gaia variability) and Single-band detections. Twenty-one disc candidates are identified, 90% of which lie beyond the footprint of prior Spitzer surveys. The candidates are concentrated among low-mass members (0.31-0.68M), with no significant correlation between mass and excess (Spearman rho = +0.07, p = 0.84). A secure disc fraction of 4.0 p/m 1.5% is established by the Cross-validated subsample, independently confirmed by Gaia DR3 YSO variability classification; a photometric 3 sigma significance cut yields a consistent estimate of 2.9 p/m 1.3%. An upper bound of 12.1 p/m 2.5% is derived from all 21 W1-W2 excess candidates, with a background subtracted value of approximately 10.1%. Gaia DR3 epoch photometry reveals diverse variability morphologies, including a dramatic dipper and multiple bursters, confirming ongoing magnetospheric accretion. By extending the survey radius, this work demonstrates that infrared excess sources indicative of inner-disc emission in IC 2395 are more abundant and widely distributed than previously recognised. The survival of these discs in the lower-density outer field provides a well-characterised sample for studying the final stages of disc evolution and planet formation.

astro-ph.GA

Star Formation in the NE-Circinus Molecular Cloud Complex

This paper presents the first dedicated characterization of the NE-Circinus Complex (TGU H1984, DCld 320.7-03.6), a previously unstudied star-forming dark cloud complex serendipitously identified in archival Herschel SPIRE observations targeting the foreground Bok globules BHR 99 and BHR 100. Using Herschel SPIRE photometry, Planck Galactic Cold Clumps data, near-infrared 2MASS colour excess mapping, AllWISE photometry, Gaia DR3 photometry, and a 3D dust extinction map, the complex, and its embedded YSO population are characterized for the first time. Two independent photometric methods place the complex at 750 +/- 50 pc. Three morphologically distinct components are identified: a dense main cloud body, a diffuse eastern component, and a northern filamentary extension. The main cloud body has a SPIRE-derived dust temperature of 14.5 +/- 0.5 K. Near-infrared H-K colour excess mapping yields a core gas mass of ~277 Msun and a total gas mass of ~439 Msun, consistent with the Planck PGCC column density. An AllWISE YSO census identifies 39 candidates across all three components, confirming active star formation throughout the complex. Multi-epoch NEOWISE-R photometry reveals three significantly variable sources, including one aperiodic dipper - a previously uncatalogued YSO exhibiting dimming consistent with inner-disc occultation. Two previously uncatalogued compact Bok globules are identified on the western edge of the northern extension, both detected in SPIRE continuum emission. That this complex went uncharacterized despite Herschel data being publicly available since 2013 underscores the scientific value of archival examination and the incomplete state of southern sky molecular cloud inventories.

astro-ph.GA

NEAT: Neighborhood-Guided, Efficient, Autoregressive Set Transformer for 3D Molecular Generation

Transformer-based autoregressive models offer an efficient alternative to diffusion- and flow-matching-based approaches for generating 3D molecules. One challenge remains: standard transformer architectures require a sequential ordering of tokens, which is not inherently defined for the atoms in a molecule. Prior works have addressed this by using canonical atom orderings. However, these approaches are not permutation invariant w.r.t. atoms and bias next-token prediction towards ordering conventions. We overcome this limitation by introducing a novel neighborhood-guided training strategy. Our model, NEAT (Neighborhood-Guided, Efficient, Autoregressive Set Transformer) treats molecular graphs as sets of atoms and learns an order-agnostic distribution over admissible tokens at the graph boundary, thereby ensuring atom-level permutation invariance. NEAT achieves state-of-the-art generation quality on the QM9 and GEOM-Drugs datasets while offering a significant speed advantage over existing baselines.

cs.LG

MEDDxAgent: A Unified Modular Agent Framework for Explainable Automatic Differential Diagnosis

Differential Diagnosis (DDx) is a fundamental yet complex aspect of clinical decision-making, in which physicians iteratively refine a ranked list of possible diseases based on symptoms, antecedents, and medical knowledge. While recent advances in large language models (LLMs) have shown promise in supporting DDx, existing approaches face key limitations, including single-dataset evaluations, isolated optimization of components, unrealistic assumptions about complete patient profiles, and single-attempt diagnosis. We introduce a Modular Explainable DDx Agent (MEDDxAgent) framework designed for interactive DDx, where diagnostic reasoning evolves through iterative learning, rather than assuming a complete patient profile is accessible. MEDDxAgent integrates three modular components: (1) an orchestrator (DDxDriver), (2) a history taking simulator, and (3) two specialized agents for knowledge retrieval and diagnosis strategy. To ensure robust evaluation, we introduce a comprehensive DDx benchmark covering respiratory, skin, and rare diseases. We analyze single-turn diagnostic approaches and demonstrate the importance of iterative refinement when patient profiles are not available at the outset. Our broad evaluation demonstrates that MEDDxAgent achieves over 10% accuracy improvements in interactive DDx across both large and small LLMs, while offering critical explainability into its diagnostic reasoning process.

cs.CL

PharmacoMatch: Efficient 3D Pharmacophore Screening via Neural Subgraph Matching

The increasing size of screening libraries poses a significant challenge for the development of virtual screening methods for drug discovery, necessitating a re-evaluation of traditional approaches in the era of big data. Although 3D pharmacophore screening remains a prevalent technique, its application to very large datasets is limited by the computational cost associated with matching query pharmacophores to database molecules. In this study, we introduce PharmacoMatch, a novel contrastive learning approach based on neural subgraph matching. Our method reinterprets pharmacophore screening as an approximate subgraph matching problem and enables efficient querying of conformational databases by encoding query-target relationships in the embedding space. We conduct comprehensive investigations of the learned representations and evaluate PharmacoMatch as pre-screening tool in a zero-shot setting. We demonstrate significantly shorter runtimes and comparable performance metrics to existing solutions, providing a promising speed-up for screening very large datasets.

cs.LG

Let's Think Frame by Frame with VIP: A Video Infilling and Prediction Dataset for Evaluating Video Chain-of-Thought

Despite exciting recent results showing vision-language systems' capacity to reason about images using natural language, their capacity for video reasoning remains under-explored. We motivate framing video reasoning as the sequential understanding of a small number of keyframes, thereby leveraging the power and robustness of vision-language while alleviating the computational complexities of processing videos. To evaluate this novel application, we introduce VIP, an inference-time challenge dataset designed to explore models' reasoning capabilities through video chain-of-thought. Inspired by visually descriptive scene plays, we propose two formats for keyframe description: unstructured dense captions and structured scene descriptions that identify the focus, action, mood, objects, and setting (FAMOuS) of the keyframe. To evaluate video reasoning, we propose two tasks: Video Infilling and Video Prediction, which test abilities to generate multiple intermediate keyframes and predict future keyframes, respectively. We benchmark GPT-4, GPT-3, and VICUNA on VIP, demonstrate the performance gap in these complex video reasoning tasks, and encourage future work to prioritize language models for efficient and generalized video reasoning.

cs.CL

Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Recent advances in large language models elicit reasoning in a chain-of-thought that allows models to decompose problems in a human-like fashion. Though this paradigm improves multi-step reasoning ability in language models, it is limited by being unimodal and applied mainly to question-answering tasks. We claim that incorporating visual augmentation into reasoning is essential, especially for complex, imaginative tasks. Consequently, we introduce VCoT, a novel method that leverages chain-of-thought prompting with vision-language grounding to recursively bridge the logical gaps within sequential data. Our method uses visual guidance to generate synthetic multimodal infillings that add consistent and novel information to reduce the logical gaps for downstream tasks that can benefit from temporal reasoning, as well as provide interpretability into models' multi-step reasoning. We apply VCoT to the Visual Storytelling and WikiHow summarization datasets and demonstrate through human evaluation that VCoT offers novel and consistent synthetic data augmentation beating chain-of-thought baselines, which can be used to enhance downstream performance.

cs.CL