arXiv ScienceSearch

arXiv · 2505.06620

Integrating Explainable AI in Medical Devices: Technical, Clinical and Regulatory Insights and Recommendations

Abstract

There is a growing demand for the use of Artificial Intelligence (AI) and Machine Learning (ML) in healthcare, particularly as clinical decision support systems to assist medical professionals. However, the complexity of many of these models, often referred to as black box models, raises concerns about their safe integration into clinical settings as it is difficult to understand how they arrived at their predictions. This paper discusses insights and recommendations derived from an expert working group convened by the UK Medicine and Healthcare products Regulatory Agency (MHRA). The group consisted of healthcare professionals, regulators, and data scientists, with a primary focus on evaluating the outputs from different AI algorithms in clinical decision-making contexts. Additionally, the group evaluated findings from a pilot study investigating clinicians' behaviour and interaction with AI methods during clinical diagnosis. Incorporating AI methods is crucial for ensuring the safety and trustworthiness of medical AI devices in clinical settings. Adequate training for stakeholders is essential to address potential issues, and further insights and recommendations for safely adopting AI systems in healthcare settings are provided.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Dima Alattal, Asal Khoshravan Azar, Puja Myles, Richard Branson, Hatim Abdulhussein, Allan Tucker. 2025-05-10. Integrating Explainable AI in Medical Devices: Technical, Clinical and Regulatory Insights and Recommendations. https://arxiv.org/abs/2505.06620

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

"A Party That Never Stops": Designing Digital Third Places for Young People's Friendships on Discord

Third places---informal settings where relationships grow through repeated casual contact---are increasingly inaccessible to young people. Prior work shows that virtual third places exist, but little is known about how persistent settings and control over participation work together as relationships form. We interviewed 25 active Discord users ages 15--24 who had reported forming at least one friendship on the platform. Their accounts highlighted two conditions that worked together. Persistent, socially legible servers let them return to the same community, recognize others, and build shared histories and norms. Controllable presence let them observe, participate, deepen a relationship, or withdraw while managing visibility, identity, and commitment. Graduated visibility was especially important for cautious entry and identity exploration. We translate these findings into design principles, showing how each can support some of Oldenburg's third-place functions while creating tensions with others. These findings suggest that persistent digital communities can complement increasingly scarce physical third places by giving young people recurring, low-pressure opportunities for social connection.

cs.HC

Demonstrably Informed Consent in Privacy Policy Flows: Evidence from a Randomized Experiment

Privacy policies govern how personal data is collected, used, and shared. Yet, in most privacy-policy consent flows, agreement is operationalized as a single click at the end of a long, opaque policy document. Recent privacy-law scholarship has argued for a standard of demonstrably informed consent. That is, the party drafting and designing privacy-policy consent mechanisms must generate reliable evidence that a person demonstrates comprehension of the consequential terms to which they agree. To this end, we study pedagogical friction as a design framing: minimal interventions embedded within a privacy-policy consent flow that aim to support demonstrated comprehension while keeping burden on the user low. In a randomized experiment, we tested pedagogical friction for demonstrably informed consent in the context of a privacy policy for an edtech app for young children. We recruited 293 parents of kids ages 3-8 to review the app's privacy policy under one of six conditions that varied presentation format and pacing, then complete a six-question comprehension quiz. Three conditions offered a second policy review and quiz retake for participants who did not pass this quiz on their first attempt. We find that the slide-based condition (G3) achieved the highest first-attempt threshold attainment (>=80%) (41.7%), followed by the paced, sectioned condition (G4) (30.6%). In the retake conditions, 64.9% of participants who completed a second attempt improved their score. Notably, in conditions that did not gate consent on demonstrated comprehension, 97.3% of participants who scored below the threshold still chose to consent, suggesting that ungated consent flows can record agreement without demonstrated comprehension. Our results suggest that pedagogical friction can strengthen the evidentiary basis of consent and clarify what it costs in time and burden.

cs.HC

Beyond the Townhall: Spatial Anchoring and LLM Agents for Scalable Participatory Urban Planning

LLMs are increasingly explored for public participation, but less is known about how they should be designed around the information citizens encounter beforehand. In an experiment (N=200) on participatory urban planning, participants learned about a sustainability project adapted from a real street redesign, through text and images or narrated 360-degree video with spatial audio and spatial anchoring; they subsequently engaged with factual and reflective LLM personas. The immersive condition improved recall and changed what participants said to the agents: they named beneficiaries more often, asked directionally more project-related questions, and provided more positive feedback. We contribute (1) a conceptualization of LLMs as dialogue partners complementing a shared information baseline, (2) causal evidence that this baseline shapes those conversations, (3) an account of how citizens use and judge factual versus reflective agent roles, and (4) a method for cities to test information packages and agent configurations before public release.

cs.HC