Back to all papers

Subcategorisation of Data for AI Models in Healthcare: A Case Study in Mammography.

July 28, 2026pubmed logopapers

Authors

Goldring JE,Cooke EA,van Engen R,Mackenzie A,Venton J,Roozemond C,Thomas SA,Smith NAS

Affiliations (5)

  • National Physical Laboratory, Teddington, Middlesex TW11 0LW, UK.
  • UCL Great Ormond Street Institute of Child Health, University College London, 30 Guilford St, London WC1N 1EH, UK.
  • Dutch Reference Centre for Screening (LRCB), Wijchenseweg 101, 6538 SW Nijmegen, The Netherlands.
  • Royal Surrey NHS Foundation Trust, Guildford GU2 7XX, UK.
  • TÜV SÜD UK, Chadwick House Warrington Rd, Birchwood Park, Risley, Warrington WA3 6AE, UK.

Abstract

<b>Background/Objectives:</b> Accurate data subcategorising is vital for reliability and traceability in the training and validation of all artificial intelligence (AI) models. <b>Methods:</b> In this paper we show the complexity of clinical and technical features likely to affect the appearance and interpretation of mammography images and in turn affect the output of AI software used to aid clinical decisions. <b>Results:</b> Using mammography as a case study, the equitability covers screened population characteristics (e.g., women's age and ethnicity) and image acquisition key factors (e.g., brand of system, exposure factors, image processing). We examine some studies and available datasets of mammography images, summarising the metadata available. <b>Conclusions:</b> We recommend that, where possible, AI models are trained and evaluated using data that includes subcategories based on these features, ensuring increased equitability in the data and coverage of image heterogeneities; or, where not possible, that the subcategories for which the AI model is valid are clearly defined. Such practices can easily be implemented in a wide range of AI applications but are illustrated here with mammography. Clinical Relevance: Reliable AI holds invaluable potential for both clinical efficiency and accuracy in diagnosis. With appropriately categorised training data, a reduction in subjective assessment can be achieved, leading to trustworthy and rapid assessment.

Topics

Journal Article

Ready to Sharpen Your Edge?

Subscribe to join 11k+ peers who rely on RadAI Slice. Get the essential weekly briefing that empowers you to navigate the future of radiology.

We respect your privacy. Unsubscribe at any time.