arXiv ScienceSearch

arXiv subjects

Ruth E. Baker

Publications and source records attributed to Ruth E. Baker.

At least 19 recordsLinked to original sources

Reliable mechanistic operator recovery with biologically-informed neural networks: principles for architecture and optimisation design

Many biological processes are governed by complex dynamical mechanisms that remain incompletely understood despite increasing volumes of experimental data. Biologically-informed neural networks (BINNs) seek to address this challenge by embedding mechanistic differential equations into neural network training, enabling interpretable constitutive operators to be recovered directly from sparse and noisy observations. However, reliable operator recovery depends sensitively on network architecture, optimisation strategy, and data informativeness. Here, we present a systematic empirical study of how these factors influence mechanistic inference using BINNs applied to canonical one-dimensional advection-diffusion-reaction partial differential equation models. Across a suite of benchmark problems, we investigate how network expressivity, learning rate, loss weighting, and batch size influence optimisation behaviour and operator recovery. We show that successful mechanistic inference depends on balancing competing objectives rather than maximising any single aspect of the model or optimisation. Moderately expressive architectures outperform overly complex networks, intermediate learning rates improve optimisation stability, balanced data and PDE losses are essential for accurate operator recovery, and intermediate batch sizes provide the best compromise between computational efficiency and reproducibility. We further identify practical diagnostics for recognising common failure modes, including over-fitting, unstable optimisation, and poor mechanistic recovery when the ground truth is unavailable. Together, these findings provide evidence-based guidelines for deploying BINNs as credible tools for biological model discovery.

q-bio.QM

A likelihood-based framework for simultaneously learning both noise and growth dynamics using biologically-informed neural networks

In recent years, neural ordinary differential equation frameworks such as Biologically-Informed Neural Networks (BINNs) have shown promise for learning mechanistic laws from sparse data. However, most existing approaches implicitly assume homoscedastic Gaussian noise, and therefore do not account for potentially meaningful structure in biological variability. Here, we present an extension to the existing BINNs framework that includes a learnable noise model, allowing discovery of the noise model directly from data. Using population growth as an example, we demonstrate that the framework accurately recovers the underlying noise structure and improves predictions of the underlying growth laws compared to existing approaches. As such, this work establishes a general likelihood-based framework for jointly learning dynamics and heteroscedastic noise within mechanistic neural network approaches.

q-bio.QM

Reliable model selection in the presence of parameter non-identifiability

Mathematical models are invaluable for understanding and predicting how biological systems behave, although their construction requires specifying mechanisms and relationships that are often not perfectly known. In the presence of multiple competing models, model uncertainty should be accounted for when performing inference based on available data. Bayesian model selection is a framework for testing mechanistic hypotheses and generating predictions under model uncertainty, which generally requires computation of the model evidence. In this work, we investigate the reliability of evidence computation methods when parameter non-identifiability -- the inability to distinguish between parameter values given available data -- is present, and find that deterministic evidence approximations can produce misleading model selection results because their underlying assumptions are violated. We propose a novel implementation of adaptive multiple importance sampling for evidence estimation, and demonstrate its robustness against non-identifiability. We use ecological case studies to demonstrate how simple model selection methods fail to produce accurate results, whereas our method yields model selection results that are comparable to those obtained by Markov chain Monte Carlo methods at substantially lower computational cost. Given the pervasiveness of parameter non-identifiability in mathematical biology, this work provides a practical approach to reliable model selection in the presence of poorly identified parameters.

stat.ME

Structural identifiability of partially-observed stochastic processes: from single-particle trajectories to total particle density data

The increasing availability of experimental data has intensified interest in calibrating stochastic models, raising fundamental questions about parameter identifiability. Structural identifiability determines whether parameters can be uniquely recovered from idealised, noise-free data, a prerequisite to allow for parameter estimation in real-world scenarios. However, existing methods to assess structural identifiability are not generally applicable to stochastic processes. We develop a methodology to analyse structural identifiability for a class of stochastic processes. We investigate how structural identifiability depends on the type of available data, distinguishing between single-particle trajectories and total particle density measurements. For trajectory data, we use the particle-based model description that explicitly represents single-particle dynamics. For population-level data, we derive a partial differential equation model representation, that describes the evolution of total particle density, and apply a differential algebra approach, common to ordinary differential equation analysis. We further introduce a method to study information arising from the initial condition, based on using the characteristic equations to construct a Taylor expansion of the particle density evolution. We apply our methodology to an example model and show that it is structurally identifiable from single-particle trajectory data but not from total particle density data, demonstrating that parameter identifiability depends on the type of data available.

stat.ME

Cell-cell adhesion cannot sustain extended follower streams in a minimal non-local model of leader-follower migration

Cell-cell adhesion is widely hypothesised to maintain cohesion within the long streams of follower cells that trail leader subpopulations during collective migration, including in neural crest cell migration, angiogenesis, and cancer cell invasion. Mathematically, non-local advection-diffusion equations provide the canonical continuum framework within which to study such adhesive cell-cell interactions. Here, we study a minimal model of leader-follower migration within this framework, in which leaders migrate at constant velocity while followers are attracted to leaders and to one another over a finite spatial interaction range. Numerical simulations reveal that, although the model can maintain small cohorts of travelling follower cells, the size of these cohorts is limited by the adhesive interaction lengthscale, and is far below what is needed to reproduce the extended streams observed in vivo. This points to a structural limitation of the standard non-local adhesion formulation and highlights the need for the development of extended continuum models capable of sustaining long, coherent migratory streams through purely mass-conserving collective cell movement.

q-bio.CB

An optimal control approach to nonlinear wave speed selection in reaction-diffusion equations

Travelling wave solutions of reaction-diffusion equations are widely used to model the spatial spread of populations and other phenomena in biology and physics. In this article, we reinterpret the classical variational principle approach through an optimal control formulation, in order to obtain a lower bound on the invasion speed of travelling wave solutions in systems of nonlinear partial differential equations. We begin by analysing single-species models, where the evolution of the density is governed by a scalar equation with a density-dependent diffusion term and a nonlinear reaction term. We show that for any admissible test function, maximising with respect to the parameter of interest yields a bound on the travelling wave speed. We apply this framework to several examples, including the porous-Fisher equation, and examine when nonlinear selection mechanisms dominate over the classical linear marginal stability criterion. Extending this approach, we then consider multi-species systems of reaction-diffusion equations and, reframed as Pontryagin-type optimality systems, we derive analogous bounds on the travelling wave speed using a variational framework under weak coupling. Finally, we employ numerical simulations to confirm the accuracy of the predicted wave speeds across a range of illustrative examples.

math.AP

Framing local structural identifiability and observability in terms of parameter-state symmetries

We introduce a subclass of Lie symmetries, called parameter-state symmetries, to analyse the local structural identifiability and observability of mechanistic models consisting of state-dependent ODEs with observed outputs. These symmetries act on parameters and states while preserving observed outputs at every time point. We prove that locally structurally identifiable parameter combinations and locally structurally observable states correspond to universal invariants of all parameter-state symmetries of a given model. We illustrate the framework on four previously studied mechanistic models, confirming known identifiability results and revealing novel insights into which states are observable, providing a unified symmetry-based approach for analysing structural properties of dynamical systems.

math.DS

Learning functional components of PDEs from data using neural networks

Partial differential equation (PDE) models frequently contain unknown functional terms that cannot be measured directly, limiting their predictive utility. While data-driven methods for estimating scalar PDE parameters are well established, the recovery of unknown functions remains comparatively underexplored. Here, we show that standard parameter estimation workflows can be extended to infer functional components of PDEs directly from data. Our approach embeds neural networks within the PDE framework, allowing unknown functions to be learned during training with high accuracy. Using nonlocal aggregation-diffusion equations as a case study, we infer interaction kernels and external potentials from steady-state observations. We systematically examine how reconstruction accuracy depends on factors such as the number and diversity of available solutions, sampling density, and measurement noise. The resulting framework retains the advantages of conventional PDE calibration approaches while extending them to functional inference: once trained, the PDE model can be used in the standard way to analyse system behaviour and generate predictions.

cs.LG

The spontaneous emergence of leaders and followers in a mathematical model of cranial neural crest cell migration

Many agent-based mathematical models of cranial neural crest cell (CNCC) migration impose a binary phenotypic partition of cells into either leaders or followers. In such models, the movement of leader cells at the front of collectives is guided by local chemoattractant gradients, while follower cells behind leaders move according to local cell-cell guidance cues. Although such model formulations have yielded many insights into the mechanisms underpinning CNCC migration, they rely on fixed phenotypic traits that are difficult to reconcile with evidence of phenotypic plasticity in vivo. A later agent-based model of CNCC migration aimed to address this limitation by allowing cells to adaptively combine chemotactic and cell-cell guidance cues during migration. In this model, cell behaviour adapts instantaneously in response to environmental cues, which precludes the identification of a persistent subset of cells as leader-like over biologically relevant timescales, as observed in vivo. Here, we build on previous leader-follower and adaptive phenotype models to develop a polarity-based agent-based model of CNCC migration, in which all cells evolve according to identical rules, interact via a pairwise interaction potential, and carry polarity vectors that evolve according to a dynamical system driven by time-averaged exposure to chemoattractant gradients. Numerical simulations of this model show that a leader-follower phenotypic partition emerges spontaneously from the underlying collective dynamics of the model. Furthermore, the model reproduces behaviour that is consistent with experimental observations of CNCC migration in the chick embryo. Thus, we provide an experimentally consistent, mechanistically-grounded mathematical model that captures the emergence of leader and follower cell phenotypes without their imposition a priori.

q-bio.CB

Survival and invasion dynamics in cell populations: an analytical framework for threshold behaviour in nonlinear age-structured models

Cell populations invade through a combination of proliferation and motility. Proliferation depends on the internal timing of cell division: how long cells take to complete the cell cycle. This timing varies substantially within (and across) cell types, creating age structure where cells at different times since their last division have different propensities to divide. Classical mathematical models of cell spreading treat division as memoryless and predict exponential cell-cycle-time distributions. Lineage tracing, by contrast, reveals peaked, gamma-like distributions that indicate a maturation delay leading to a fertility window. This gap motivates a modelling framework that incorporates age-dependent cell division rates while retaining analytical tractability. We address this through a moment-hierarchy framework that tracks time since cell division, with age resetting to zero at division. The framework yields explicit formulae for steady-state age distributions, cell-cycle-time distributions, and invasion speeds. For age-independent rates, we recover classical Fisher--KPP. Three fundamental principles emerge. First, age structure systematically reduces a population's carrying capacity and narrows the viable parameter range for positive steady states. Second, classical linear theory overestimates invasion speeds; the true minimal speed is slower when division is age-dependent. Third, the parameter condition for population survival is identical to the condition for a positive invasion speed.

q-bio.CB

Spatial correlations in SIS processes on random regular graphs

In network-based SIS models of infectious disease transmission, infection can only occur between directly connected individuals. This constraint naturally gives rise to spatial correlations between the states of neighboring nodes, as the infection status of connected individuals becomes interdependent. Although mean-field approximations and the standard pairwise model are commonly used to simplify disease forecasting on networks, they inadequately capture spatial correlations; mean-field frameworks assume that populations are well-mixed, while the pairwise model neglects correlations beyond nearest-neighbor connections, which leads to inaccurate predictions of infection numbers over time. As such, the development of approximations that account for higher order spatially correlated infections is of great interest, as they offer a compromise between accurate disease forecasting and analytic tractability. Here, we use existing corrections to mean-field theory on the regular lattice to construct a more general framework for equivalent corrections on random regular graph topologies. We derive and simulate a hierarchical system of ordinary differential equations for the time evolution of the spatial correlation function at various geodesic distances on random networks. Solving these equations allows us to predict the time-dependent global infection density, which agrees well with numerical simulations. Our results substantially improve on existing corrections to mean-field theory for infectious individuals in SIS processes and provide an in-depth characterization of how structural randomness in networks affects the dynamical trajectories of infectious diseases on networks.

cond-mat.stat-mech

An energy-based mathematical model of actin-driven protrusions in eukaryotic chemotaxis

In eukaryotic cell chemotaxis, cells extend and retract transient actin-driven protrusions at their membrane that facilitate both the detection of external chemical gradients and directional movement via the formation of focal adhesions with the extracellular matrix. Although extensive experimental work has detailed how cellular protrusions and morphology vary under different environmental conditions, the mechanistic principles linking protrusive activity to these factors remain poorly understood. Here, we model the extension of actin-based protrusions in chemotaxis as an optimisation problem, wherein cells balance the detection of chemical gradients with the energetic cost of protrusion formation. Our model, built on the assumption of energy minimisation, provides a framework that successfully reproduces experimentally observed patterns of protrusive activity across a range of biological systems and environmental conditions, suggesting that energetic efficiency may underpin the morphology and chemotactic behaviour of motile eukaryotic cells. Additionally, we leverage the model to generate novel predictions regarding cellular responses to other, experimentally untested environmental perturbations, providing testable hypotheses for future experimental work that may be used to validate and refine the model presented here.

q-bio.CB

Optimal experiment design for practical parameter identifiability and model discrimination

Mechanistic mathematical models of biological systems usually contain a number of unknown parameters whose values need to be estimated from available experimental data in order for the models to be validated and used to make quantitative predictions. This requires that the models are practically identifiable, that is, the values of the parameters can be confidently determined, given available data. A well-designed experiment can produce data that are much more informative for the purpose of inferring parameter values than a poorly designed experiment. It is, therefore, of great interest to optimally design experiments such that the resulting data maximise the practical identifiability of a chosen model. Experimental design is also useful for model discrimination, where we seek to distinguish between multiple distinct, competing models of the same biological system in order to determine which model better reveals insight into the underlying biological mechanisms. In many cases, an external stimulus can be used as a control input to probe the behaviour of the system. In this paper, we will explore techniques for optimally designing such a control for a given experiment, in order to maximise parameter identifiability and model discrimination, and demonstrate these techniques in the context of commonly applied ordinary differential equation models. We use a profile likelihood-based approach to assess parameter identifiability. We then show how the problem of optimal experimental design for model discrimination can be formulated as an optimal control problem, which can be solved efficiently by applying Pontryagin's Maximum Principle.

q-bio.QM

The influence of cell phenotype on collective cell invasion into the extracellular matrix

Understanding the interactions between cells and the extracellular matrix (ECM) during collective cell invasion is crucial for advancements in tissue engineering, cancer therapies, and regenerative medicine. This study focuses on the roles of contact guidance and ECM remodelling in directing cell behaviour, with a particular emphasis on exploring how differences in cell phenotype impact collective cell invasion. We present a computationally tractable two-dimensional hybrid model of collective cell migration within the ECM, where cells are modelled as individual entities and collagen fibres as a continuous tensorial field. Our model incorporates random motility, contact guidance, cell-cell adhesion, volume filling, and the dynamic remodelling of collagen fibres through cellular secretion and degradation. Through a comprehensive parameter sweep, we provide valuable insights into how differences in the cell phenotype, in terms of the ability of the cell to migrate, secrete, degrade, and respond to contact guidance cues from the ECM, impacts the characteristics of collective cell invasion.

q-bio.CB

A likelihood-based Bayesian inference framework for the calibration of and selection between stochastic velocity-jump models

Advances in experimental techniques allow the collection of high-resolution spatio-temporal data that track individual motile entities. These tracking data can be used to calibrate mathematical models describing the motility of individual entities. The challenges in calibrating models for single-agent motion derive from the intrinsic characteristics of experimental data, collected at discrete time steps and with measurement noise. We consider motion of individual agents that can be described by velocity-jump models in one spatial dimension. These agents transition between a network of \textit{n} states, in which each state is associated with a fixed velocity and fixed rates of switching to every other state. Exploiting approximate solutions to the resultant stochastic process, we develop a Bayesian inference framework to calibrate these models to discrete-time noisy data. We first demonstrate that the framework can be used to effectively recover the model parameters of data simulated from two-state and three-state models. Finally, we explore the question of model selection first using simulated data and then using experimental data tracking mRNA transport inside \textit{Drosophila} neurons. Overall, our results demonstrate that the framework is effective and efficient in calibrating and selecting between velocity-jump models and it can be applied to a range of motion processes.

stat.ME

A nonlocal-to-local approach to aggregation-diffusion equations

Over the past decades, nonlocal models have been widely used to describe aggregation phenomena in biology, physics, engineering, and the social sciences. These are often derived as mean-field limits of attraction-repulsion agent-based models, and consist of systems of nonlocal partial differential equations. Using differential adhesion between cells as a biological case study, we introduce a novel local model of aggregation-diffusion phenomena. This system of local aggregation-diffusion equations is fourth-order, resembling thin-film or Cahn-Hilliard type equations. In this framework, cell sorting phenomena are explained through relative surface tensions between distinct cell types. The local model emerges as a limiting case of short-range interactions, providing a significant simplification of earlier nonlocal models, while preserving the same phenomenology. This simplification makes the model easier to implement numerically and more amenable to calibration to quantitative data. Additionally, we discuss recent analytical results based on the gradient-flow structure of the model, along with open problems and future research directions.

q-bio.CB

Modelling collective cell migration in a data-rich age: challenges and opportunities for data-driven modelling

Mathematical modelling has a long history in the context of collective cell migration, with applications throughout development, disease and regenerative medicine. The aim of modelling in this context is to provide a framework in which to mathematically encode experimentally derived mechanistic hypotheses, and then to test and validate them to provide new insights and understanding. Traditionally, mathematical models have consisted of systems of partial differential equations that model the evolution of cell density over time, together with the dynamics of any associated biochemical signals or the underlying substrate. The various terms in the model are usually chosen to provide simplified, phenomenological descriptions of the underlying biology, and follow long-standing conventions in the field. However, with the recent development of a plethora of new experimental technologies that provide quantitative data on collective cell migration processes, we now have the opportunity to leverage statistical and machine learning tools to determine mathematical models directly from the data. This perspectives article aims to provide an overview of recently developed data-driven modelling approaches, outlining the main methodologies and the challenges involved in using them to interrogate real-world data relating to collective cell migration.

q-bio.QM

Optimal experimental design for parameter estimation in the presence of observation noise

Using mathematical models to assist in the interpretation of experiments is becoming increasingly important in research across applied mathematics, and in particular in biology and ecology. In this context, accurate parameter estimation is crucial; model parameters are used to both quantify observed behaviour, characterise behaviours that cannot be directly measured and make quantitative predictions. The extent to which parameter estimates are constrained by the quality and quantity of available data is known as parameter identifiability, and it is widely understood that for many dynamical models the uncertainty in parameter estimates can vary over orders of magnitude as the time points at which data are collected are varied. Here, we use both local sensitivity measures derived from the Fisher Information Matrix and global measures derived from Sobol' indices to explore how parameter uncertainty changes as the number of measurements, and their placement in time, are varied. We use these measures within an optimisation algorithm to determine the observation times that give rise to the lowest uncertainty in parameter estimates. Applying our framework to models in which the observation noise is both correlated and uncorrelated demonstrates that correlations in observation noise can significantly impact the optimal time points for observing a system, and highlights that proper consideration of observation noise should be a crucial part of the experimental design process.

math.ST