Aller au contenu principal
Accès ouvert déclaré 2026 preprint

Expert Discrimination of AI-Generated versus Authentic Radiologic Images: A Multimodal, Pre-Registered Visual Turing Test

0Citations signalées, ce qui n’est pas une note de qualité
15Institutions déclarées
3Pays d’affiliation déclarés

Rattachement africain : kr, ch, us. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

Abstract Background Frontier text-to-image models can synthesise radiologic images of high realism, raising the question of whether expert radiologists can serve as a provenance safeguard for the medical image record. Methods We conducted a prospective, pre-registered visual Turing test in which 60 invited Korean board-certified radiology faculty and trainees judged authentic (teaching-repository) and AI-generated radiologic images from a locked pool of 241 displayable cells (82 entities; nine subspecialties; six modalities; 60 readers x 60 trials = 3,600 reader–image observations) produced by two contemporary commercial generators. The primary endpoint was the confidence-weighted, reader-averaged multi-reader multi-case area under the curve for AI versus authentic images, conditional on the locked image pool; the key secondary endpoint was the Faculty-minus-Junior difference under a two one-sided tests equivalence framework. The pre-specified statistical analysis plan was registered on the Open Science Framework before data lock. Findings All 60 readers completed the test. The pooled confidence-weighted area under the curve was 0.71 (95% CI, 0.69 to 0.74), above the null value of 0.5 but within the pre-specified modest tier (0.60 to 0.75). The Faculty-minus-Junior contrast was 0.04 (95% CI, −0.02 to 0.10), including zero, and the two one-sided tests established equivalence within the +/−0.10 margin. No reader stratum and no pre-specified sensitivity analysis reached the deployable-classifier threshold (area under the curve >= 0.75). Interpretation In this single-country cohort, expert radiologists distinguished frontier-generated from authentic radiologic images only modestly, without a meaningful expertise gradient (equivalence within +/−0.10) and with no reader stratum reaching a standalone provenance safeguard. These findings support radiology AI-literacy training and pipeline-level provenance safeguards rather than reliance on reader judgment, and warrant retesting in an independent reader cohort. Funding This research was supported by a grant of the Korea Health Technology R&D Project through the Korea Health Industry Development Institute (KHIDI), funded by the Ministry of Health & Welfare, Republic of Korea (grant number: RS-2025-02213531).

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Expert Discrimination of AI-Generated versus Authentic Radiologic Images: A Multimodal, Pre-Registered Visual Turing Test
Date Crossref
06/07/2026
Éditeur
openRxiv
Type
posted-content

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Les institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Artificial Intelligence in Healthcare and EducationRadiology practices and educationAI in cancer detection

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.