GPT-4o matched the performance of experienced radiologists and surpassed residents in recommending follow-up imaging from routine radiology reports.
Key Details
- 1Study involved 100 CT/MRI oncologic cases across head and neck, liver, lung, and pancreas from two academic medical centers.
- 2GPT-4o, a radiology resident, and a board-certified radiologist each generated follow-up recommendations from report texts.
- 3Blinded senior radiologists rated recommendations for completeness, modality, timing, and overall quality on a five-point scale.
- 4GPT-4o achieved a median global quality score of 4 (same as board-certified, higher than resident), with 96% timing correctness and 92% completeness.
- 5GPT-4o showed the strongest performance in lung imaging (100% timing correctness).
- 6No significant differences found among readers for appropriateness of imaging modality.
Why It Matters
The findings demonstrate GPT-4o's potential to standardize and improve the consistency of follow-up imaging recommendations, supporting its use as a decision support tool in radiology. This could reduce variability, enhance guideline adherence, and assist radiologists in routine clinical workflows.

Source
AuntMinnie
Related News

•Radiology Business
AI-Powered Tool Streamlines CT Scan Prioritization in Emergency Departments
An AI-based CT queue system significantly reduces wait times for ED patients by prioritizing scans likely to reveal critical findings.

•Radiology Business
AI Workflow Enables General Radiologists to Match Breast Specialists in Screening
AI-powered workflow helps generalist radiologists detect breast cancer at rates comparable to specialists.

•Radiology Business
LLMs Automate Radiology Report Quality Control, Study Finds
LLM-based systems can rapidly automate radiology report quality control, saving significant manual review time.