arXiv Science⌕ Search

arXiv · 2609.31935

Feature-Based Likelihood Ratios for Forensic Science: Combining Neural Networks with Bayesian Probability Calculus

Abstract

In forensic science, when crime-scene evidence (CSE) and suspect-related evidence (SRE) is present, it is customary to report on the value of this evidence in the form of a likelihood ratio (LR). The LR can be calculated as the probability of CSE given SRE divided by the probability of CSE given that it was generated by a randomly selected person from an alternative culprit population. In forensic science, this is known as a feature-based LR, and intuitively the LR contrasts "similarity" by "typicality". Since it is generally a problem for feature-based LRs to find appropriate models for the data, one either resorts to score-based LRs or to adjusting the feature-based output post-hoc to well-calibrated output. Either way, the above definition of the LR is broken and interpretation of the LR as similarity between CSE and SRE divided by typicality of CSE is destroyed. Here, we report on progress in obtaining instantly well-performing feature-based LRs using gradient descent in combination with Bayesian probability theory to train a two-level model, the main LR model in forensic science for describing distributions of continuous data. For a dataset of laser-ablation inductively-coupled-plasma mass-spectrometry measurements on glass fragments from forensic casework, we show that our best model on validation data yields much better calibrated feature-based LRs on the test set when compared to state-of-the-art feature-based LR systems trained on the same type of data, and that it improves a factor of 4.5 on average on $C_{\mathrm{llr}}$. For LRs interpretable in terms of "similarity" contrasting "typicality", this is a major advancement. However, a state-of-the-art LR system still performs a factor of 1.5 better on $C_{\mathrm{llr}}$ for this data. We also present future plans to close this remaining gap. In order to facilitate collaboration, we have put relevant code on GitHub.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Peter Vergeer. 2026-09-25. Feature-Based Likelihood Ratios for Forensic Science: Combining Neural Networks with Bayesian Probability Calculus. https://arxiv.org/abs/2609.31935

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Constrained convex clustering for interpretable spatial domain detection in spot-based spatial transcriptomics

Popular technologies for generating spatially resolved transcriptomic data measure gene expression at the resolution of a "spot", i.e., a small tissue region 55 microns in diameter. Each spot can contain many cells of different types. In typical analyses, researchers are interested in using these data to identify and profile discrete spatial domains in tissue. In this paper, we propose a new method, DUET, which simultaneously identifies discrete spatial domains and estimates each spot's expected cell-type proportion. This allows the identified spatial domains to be characterized in terms of the underlying expected cell-type proportions, which affords interpretability and biological insight. DUET utilizes a constrained version of model-based convex clustering, and as such, can accommodate Poisson, negative binomial, normal, and other types of expression data. Moreover, our convex clustering-type criterion allows for both the number of clusters and degree of spatial smoothness to be controlled by a single tuning parameter, which can be chosen in a data-driven fashion. Through simulation studies and a real data application, we show that DUET can achieve better clustering and deconvolution performance than some existing methods.

stat.AP↗

Extended State-dependent Hawkes Process for Limit Order Books: Mathematical Foundation and the Reproduction of Volatility Signature Plots

This paper proposes an Extended State-Dependent Hawkes Process (ExsdHawkes) to model the intricate dynamics of Limit Order Books (LOBs). Our theoretical contribution lies in relaxing traditional constraints by allowing for state disappearances---a phenomenon frequently observed in high-frequency trading. We mathematically prove, using Karush--Kuhn--Tucker (KKT) conditions, that the maximum likelihood estimation remains separable, justifying an efficient two-step procedure. In the empirical section, we apply our model to three months of high-frequency tick data of Mitsubishi UFJ Financial Group (8306). We demonstrate that ExsdHawkes successfully replicates the characteristic upward slope of the volatility signature plot by capturing the ``local super-criticality'' triggered during disequilibrium states. Crucially, we clarify that the transition out of equilibrium is deterministically triggered by Aggressive Market Orders (AMS/AMB), while Marketable Limit Orders (MLO) function as a critical liquidity-depletion catalyst within the expanded spread. Comparative analysis reveals that models lacking physical constraints (e.g., standard SD-Hawkes) suffer from explosive spectral radii and fail to maintain simulation stability. Our findings suggest that physical consistency is not merely a mathematical nicety, but a prerequisite for accurately modeling macro-level volatility. By enforcing the physical geometry to `pause' the residual accumulation during inadmissible periods, ExsdHawkes maintains statistical integrity where unconstrained models succumb to structural bias and simulation instability.

stat.AP↗

Artificial Intelligence for early detection of circulatory shock in ICU patients

Circulatory shock is one of the leading causes of mortality in intensive care units (ICUs), and its early detection is critical to enable timely treatment and improve clinical outcomes. This study aimed to develop and evaluate a two-stage cascade machine learning framework for the early detection and etiological classification of circulatory shock in critically ill patients. Using data from the Medical Information Mart for Intensive Care (MIMIC)-IV database, four patient groups were defined: septic shock, cardiogenic shock, hypovolemic shock, and a non?shock control group, comprising a total of 32,907 patients. Vital signs and laboratory data were collected during the first six hours after ICU admission. After data cleaning and missing-value imputation, the mean value of each variable was used for model development. Several machine learning algorithms were compared, including logistic regression, Random Forest, XGBoost, and multilayer perceptron (MLP) networks. Random Forest and XGBoost achieved the highest overall performance, with an AUROC of approximately 0.82-0.83 for shock detection, and a macro-averaged sensitivity of approximately 0.61 and precision of approximately 0.58 across all four classes. Classification performance was highest for the non?shock group, followed by septic and cardiogenic shock, while hypovolemic shock showed the lowest performance. These results indicate that machine learning models can identify early signs of hemodynamic deterioration associated with circulatory shock and may support clinical decision-making in the ICU. However, further improvements are needed for the classification of specific shock subtypes, particularly hypovolemic and cardiogenic shock, as well as for real-time clinical implementation.

stat.AP↗