GPT-4o matched the performance of experienced radiologists and surpassed residents in recommending follow-up imaging from routine radiology reports.
Key Details
- 1Study involved 100 CT/MRI oncologic cases across head and neck, liver, lung, and pancreas from two academic medical centers.
- 2GPT-4o, a radiology resident, and a board-certified radiologist each generated follow-up recommendations from report texts.
- 3Blinded senior radiologists rated recommendations for completeness, modality, timing, and overall quality on a five-point scale.
- 4GPT-4o achieved a median global quality score of 4 (same as board-certified, higher than resident), with 96% timing correctness and 92% completeness.
- 5GPT-4o showed the strongest performance in lung imaging (100% timing correctness).
- 6No significant differences found among readers for appropriateness of imaging modality.
Why It Matters

Source
AuntMinnie
Related News

Real-World Study: Radiology AI Best in Emergency and Inpatient Settings
A commercial AI tool for intracranial aneurysm detection outperformed in inpatient and emergency settings but yielded limited benefits for outpatients in a major health system study.

New Rubric Enhances Safety of AI-Generated Radiology Summaries
Researchers developed a five-factor rubric to assess the safety and quality of AI-generated, patient-friendly radiology report summaries.

Healthcare Leader Warns AI Will Dominate Diagnostic Radiology
A leading oncologist urges future radiologists to specialize in interventional procedures due to AI advances in image interpretation.