Performance of adult-trained artificial intelligence models in paediatric imaging—a scoping review

Эффективность моделей искусственного интеллекта, обученных на данных взрослых, при визуализации у детей: обзор скоупинга
Lene Bjerke Laborie, Jennifer Lee, Edward Antram, Regina Küfner Lein, Susan Cheng Shelmerdine
2026-02-12

Dice scoreadult-trained artificial intelligencepaediatric imagingperformance degradationpulmonary nodule detection
OBJECTIVES: This scoping review aims to evaluate the performance of artificial intelligence (AI) models designed for adults when applied to paediatric imaging datasets without additional adaptations, and to quantify performance degradation across different modalities, use-cases and age groups. MATERIALS AND METHODS: A literature search was conducted covering 10 years (1/01/2014-23/06/2025) using terms relating to "child", "adult", "artificial intelligence", "radiology" and "validation/performance". Two reviewers independently extracted data using standardised templates and conducted a narrative analysis. RESULTS: Of 5642 abstracts, 20 studies met the inclusion criteria. The studies evaluated AI tools across 16 paediatric dataset cohorts ranging from 30 to 7357 subjects. Three datasets were used more than once to evaluate different AI model performance metrics. The tools were applied to radiography (n = 7), CT (n = 7), MRI (n = 2), Dual-energy-x-ray-absorptiometry (DEXA) (n = 2) and ultrasound (n = 2) across different AI tasks: segmentation (n = 9), classification (n = 4), detection (n = 3), and mixed tasks (n = 4). Apart from two studies, all articles reported performance reduction when adult-trained AI tools were applied to paediatric populations. Cohort overlap represents the risk of duplication bias. Detection tasks showed the most severe deterioration, with sensitivity dropping from 68-100% in adults to 26-68% in children for pulmonary nodule detection. For segmentation tasks, Dice score reductions > 0.10 were noted across organs and imaging modalities. Children ≤ 2 years consistently showed the greatest performance deficits across all task types. CONCLUSION: AI tools intended for adult use do not perform to the same standard when used in a paediatric population without additional adaptation, particularly for children under 2 years. Careful model evaluation is required before clinical implementation. KEY POINTS: Question How do artificial intelligence-based radiology tools designed for adults perform when applied to paediatric imaging without additional adaptation? Findings Adult-trained AI models consistently demonstrated reduced performance in children, particularly in those under 2 years, with detection tasks showing the most severe deterioration. Clinical relevance Healthcare professionals should not assume that adult-trained radiology AI tools intended for adult use can be directly applied to the paediatric population without validation, additional training or fine-tuning, particularly for the youngest age groups.
1
A scoping review identified 20 studies evaluating adult-trained AI models on paediatric imaging datasets without additional adaptation.
2
Across radiography, CT, MRI, DEXA, and ultrasound, nearly all studies reported reduced performance in paediatric populations, except two.
3
Children aged 2 years or younger consistently exhibited the greatest performance deficits, highlighting the need for paediatric-specific validation before clinical implementation.
4
Detection performance deteriorated most severely, with pulmonary nodule sensitivity decreasing from 68–100% in adults to 26–68% in children.
5
Repeated use of three datasets across studies indicates a risk of cohort-overlap and duplication bias.
6
Segmentation Dice scores decreased by more than 0.10 across multiple organs and imaging modalities.

adult-trained artificial intelligence models applied to paediatric imaging datasets and populations

performance degradation and age-, modality-, and task-dependent generalization of adult-trained AI models in paediatric imaging without adaptation

Publication Details
Publication Date
2026-02-12
Journal
Publisher
ISSN
Cited by
6
Access Type
Author Information
Authors
Lene Bjerke Laborie
Jennifer Lee
Edward Antram
Regina Küfner Lein
Susan Cheng Shelmerdine
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat
Make a presentation
100%