Background and Objective : Generative AI is transforming medical imaging research through synthesis, enhancement, and reconstruction of clinical images. While these advances show promise in addressing data scarcity and supporting diagnostics, clinical adoption remains limited due to challenges in assessing the trustworthiness of generated content. This study aims to systematically evaluate the integration of uncertainty quantification (UQ) methods within generative models for medical imaging to enhance result reliability. Methods : A systematic review following PRISMA guidelines was conducted, analyzing studies from January 2018 to December 2025 that combined generative models with UQ techniques in medical imaging. The search strategy covered major medical and computer science databases, with studies evaluated against predefined inclusion criteria focusing on implementation methodology and performance metrics. Results : From the analysis of 41 eligible studies, 29 focused on radiology, 8 on microscopy, and 4 on optical coherence tomography. Across this heterogeneous body of evidence, integrating UQ was frequently associated with improved performance or with more informative reliability assessment, including reported gains in reconstruction quality, segmentation accuracy, and anomaly detection. Notably, 56% of studies (n=23) were published in 2025, indicating rapid field growth. Conclusions : UQ integration represents a crucial advancement toward trustworthy generative AI systems in medical imaging. Key priorities identified include standardizing uncertainty metrics, developing computationally efficient frameworks, and embedding uncertainty awareness within generation processes. These findings suggest that UQ methods can enhance the clinical reliability of generative AI applications in medical imaging.

Can we really trust generative models in healthcare? A systematic review of uncertainty quantification in generative AI for medical imaging / Shahini, A., Seoni, S., Manzin, A., Oria, M., Gudigar, A., Raghavendra, U., Sengur, A., Acharya, R., Salvi, M.. - In: COMPUTER METHODS AND PROGRAMS IN BIOMEDICINE. - ISSN 0169-2607. - 285:(2026). [10.1016/j.cmpb.2026.109560]

Can we really trust generative models in healthcare? A systematic review of uncertainty quantification in generative AI for medical imaging

Shahini, Alen;Seoni, Silvia;Oria, Martina;Salvi, Massimo
2026

Abstract

Background and Objective : Generative AI is transforming medical imaging research through synthesis, enhancement, and reconstruction of clinical images. While these advances show promise in addressing data scarcity and supporting diagnostics, clinical adoption remains limited due to challenges in assessing the trustworthiness of generated content. This study aims to systematically evaluate the integration of uncertainty quantification (UQ) methods within generative models for medical imaging to enhance result reliability. Methods : A systematic review following PRISMA guidelines was conducted, analyzing studies from January 2018 to December 2025 that combined generative models with UQ techniques in medical imaging. The search strategy covered major medical and computer science databases, with studies evaluated against predefined inclusion criteria focusing on implementation methodology and performance metrics. Results : From the analysis of 41 eligible studies, 29 focused on radiology, 8 on microscopy, and 4 on optical coherence tomography. Across this heterogeneous body of evidence, integrating UQ was frequently associated with improved performance or with more informative reliability assessment, including reported gains in reconstruction quality, segmentation accuracy, and anomaly detection. Notably, 56% of studies (n=23) were published in 2025, indicating rapid field growth. Conclusions : UQ integration represents a crucial advancement toward trustworthy generative AI systems in medical imaging. Key priorities identified include standardizing uncertainty metrics, developing computationally efficient frameworks, and embedding uncertainty awareness within generation processes. These findings suggest that UQ methods can enhance the clinical reliability of generative AI applications in medical imaging.
File in questo prodotto:
Non ci sono file associati a questo prodotto.
Pubblicazioni consigliate

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11583/3013678
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo