BACKGROUND
Advances in generative artificial intelligence (AI) have accelerated the development and application of synthetic medical imaging. Despite this rapid progress, the evaluation of synthetic medical images remains heterogeneous, with numerous metrics proposed to assess fidelity, realism, diversity, and clinical validity. Currently, no standardized framework exists to guide the selection, interpretation, or comparison of these metrics, limiting reproducibility and cross-study comparability. This systematic review aims to comprehensively summarize and categorize existing metrics used to assess these complementary dimensions of synthetic medical images.
METHODS
A systematic review was conducted in accordance with PRISMA guidelines. PubMed/MEDLINE, EMBASE, Scopus, and arXiv were searched for studies published between 2015 and April 30, 2025, supplemented by citation screening of included studies. Eligible studies were full-text articles that applied or proposed metrics to evaluate the fidelity, realism, diversity, and/or clinical validity in synthetic medical images.
RESULTS
A total of 47 studies were included. Evaluation practices were highly heterogeneous. Expert evaluation (n = 25, 53%) and reference-based evaluations were most common (n = 25, 53%), followed by no-reference metrics (n = 24, 51%), and task-based evaluations (n = 24, 51%). The most commonly used individual metrics were peak signal-to-noise ratio (PSNR) (n = 16, 34%), structural similarity index (SSIM) (n = 15, 32%), mean absolute error (MAE) (n = 12, 26%), and Fréchet Inception Distance (FID) (n = 12, 26%).
CONCLUSION
Evaluation strategies for synthetic medical imaging showed substantial variability and no single metric captured fidelity, realism, diversity, and clinical validity simultaneously. Metric choice is often dictated by data availability rather than clinical purpose. A task-specific, layered evaluation framework could improve comparability and facilitate clinical adoption.
D. D. de Wilde, Benjamin Schärli, Kym Ackermann et al.· European Journal of Radiolog...· 0 citations
Peer learning is a collaborative approach grounded in social constructivist theory, emphasizing that knowledge is co-created through interaction. In medical education, it promotes active engagement, critical thinking, and professional identity formation. Within the author’s research group, PhD students in the field of neurosurgery and medical students working on their undergraduate medical research projects, benefit from structured peer learning by sharing expertise, discussing challenges, and providing feedback. This aligns with communities of practice theory, where learning occurs through participation in a shared domain, and with Topping’s framework highlighting peer tutoring as a reciprocal process that enhances cognitive and social development. Recent studies demonstrate that peer assisted learning improves clinical knowledge, practical skills acquisition, and learner engagement compared with traditional teaching approaches. These effects appear particularly strong in clinical and applied contexts, with evidence suggesting retention beyond the immediate learning period and improved transfer to performance. Integrating peer learning as a deliberate strategy strengthens academic outcomes and fosters lifelong learning skills. The Kirkpatrick framework can be used to evaluate the design, implementation, and outcomes of peer learning (PL) interventions. In line with these traditions, recent work in Academic Medicine has clarified mechanisms through which longitudinal near-peer learning operates—structured preparation, dynamic role-shifting during joint encounters, and post-encounter debriefing within trusted relationships—offering concrete practices that can be translated into the proposed setting.
A. Elmi-Terander· Journal of Medical Education...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.