When AI Is Not Enough: Reducing Diagnostic Errors with Radiologist Oversight

J. Cai & Noa Zychlinski

Service Science2025https://doi.org/10.1287/serv.2024.0234article
AJG 1
Weight
0.44

What the paper says

Artificial intelligence (AI) is becoming increasingly prevalent, particularly in healthcare, in which it is shaping the future of decision-making processes. In radiology, AI has revolutionized diagnostics by enabling rapid analysis of patient imaging. However, the consequences of AI misdiagnoses can be significant. For example, an incorrect result can unnecessarily flag a healthy patient for treatment, whereas a missed detection may fail to identify a serious condition that requires immediate intervention. To mitigate such risks, most diagnostic systems combine AI analysis with radiologist review: AI first classifies cases, and then, radiologists review and confirm or modify the initial diagnosis of AI. Effective radiology scheduling must account for the likelihood and cost of false negatives and false positives as well as AI characteristics, such as sensitivity and specificity. To address the limitations of AI predictions, we develop a multiserver queuing model with separate queues for suspected positive and suspected negative cases. Using a fluid approximation, we derive an index-based policy, a modified version of the [Formula: see text] rule, to optimally schedule and allocate resources, taking into account AI characteristics and potential misclassifications. Our proposed policy naturally incorporates the anchoring effect, causing radiologists to devote more time to misclassified cases. As the anchoring effect is incorporated into the classes’ indexes, it may change the classes’ prioritization and significantly influence overall system performance. Furthermore, to prevent excessive waiting times, even for patients diagnosed as negative, we extend our model to incorporate diagnosis-based service-level requirements established by hospitals and regulators. Numerical results demonstrate the effectiveness and superiority of our policy compared with a widely used benchmark, underscoring its potential to improve diagnostic accuracy and efficiency. History: This paper has been accepted for the Service Science Special Issue on the Impact of AI on Service Design and Delivery. Funding: Partial financial support was received from Israel Science Foundation (ISF) [Grant 277/21] and The Bernard M. Gordon Center for Systems Engineering at the Technion. Supplemental Material: The online appendix is available at https://doi.org/10.1287/serv.2024.0234 .

3 citations

Open paper page →

Cite this paper

https://doi.org/https://doi.org/10.1287/serv.2024.0234

Or copy a formatted citation

@article{j.2025,
  title        = {{When AI Is Not Enough: Reducing Diagnostic Errors with Radiologist Oversight}},
  author       = {J. Cai & Noa Zychlinski},
  journal      = {Service Science},
  year         = {2025},
  doi          = {https://doi.org/https://doi.org/10.1287/serv.2024.0234},
}

Paste directly into BibTeX, Zotero, or your reference manager.

Flag this paper

When AI Is Not Enough: Reducing Diagnostic Errors with Radiologist Oversight

Flags are reviewed by the Arbiter methodology team within 5 business days.


Evidence weight

0.44

Balanced mode · F 0.40 / M 0.15 / V 0.05 / R 0.40

F · citation impact0.32 × 0.4 = 0.13
M · momentum0.57 × 0.15 = 0.09
V · venue signal0.50 × 0.05 = 0.03
R · text relevance †0.50 × 0.4 = 0.20

† Text relevance is estimated at 0.50 on the detail page — for your query’s actual relevance score, open this paper from a search result.