September 11, 2026 · Journal of the American College of Radiology : JACR · DOI: 10.1016/j.jacr.2026.09.006

Human Judgment and the Limits of Artificial Intelligence for Automated Rank Order Lists in Diagnostic Radiology Residency Selection

Listen to this summary

The authors aimed to evaluate the reliability of large language models (LLMs) in generating rank order lists (ROLs) for diagnostic radiology residency selection based solely on pre-interview application data compared to including human interview scores. The study found that LLMs showed low agreement with the final ROL when using application-only data, while including human input significantly improved alignment. Additionally, LLM-generated rankings exhibited systematic biases against certain applicant subgroups, indicating that AI should not replace human judgment in this context.

Cody H Savage, Rydhwana Hossain, James Mac Tonascia, Athanasios Pavlou, Charles S Resnik, Florence X Doo, Elana B Smith

This is one of 33,000+ journals available on OSLR. Try it free for 14 days.

Free 14-day trial. 33,000+ journals. Cancel anytime.

14-day free trial. No commitment.

“

"Oslr has become part of my weekly routine on my day off. The clinical relevance of the summaries is outstanding — I'd rate it 9/10. Being able to consume research hands-free is a huge advantage for busy physicians."

Dr. Jennifer Thompson

Dr. Jennifer Thompson

Portland, OR

Stay current without falling behind

33,000+ journals. 3-minute audio summaries. Free for 14 days.

Download on the App StoreGet it on Google Play