Impact of a deep learning assistant on the histopathologic classification of liver cancer

Kiani, Amirhossein; Uyumazturk, Bora; Rajpurkar, Pranav; Wang, Alex; Gao, Rebecca; Jones, Erik; Yu, Yifan; Langlotz, Curtis P.; Ball, Robyn L.; Montine, Thomas J.; Martin, Brock A.; Berry, Gerald J.; Ozawa, Michael G.; Hazard, Florette K.; Brown, Ryanne A.; Chen, Simon B.; Wood, Mona; Allard, Libby S.; Ylagan, Lourdes; Ng, Andrew Y.; Shen, Jeanne

Published in

Nature Research, npj Digital Medicine, 1(3), 2020

DOI: 10.1038/s41746-020-0232-8

Tools

Export citation

Search in Google Scholar

Impact of a deep learning assistant on the histopathologic classification of liver cancer

Journal article published in 2020 by Amirhossein Kiani

, Bora Uyumazturk, Pranav Rajpurkar, Alex Wang, Rebecca Gao, Erik Jones, Yifan Yu, Curtis P. Langlotz

, Robyn L. Ball

, Thomas J. Montine, Brock A. Martin

, Gerald J. Berry

, Michael G. Ozawa, Florette K. Hazard, Ryanne A. Brown and other authors.

This paper is made freely available by the publisher.

Full text: Download

Preprint: archiving allowed

Upload

Postprint: archiving forbidden

Published version: archiving allowed

Upload

Policy details

Data provided by

Abstract

AbstractArtificial intelligence (AI) algorithms continue to rival human performance on a variety of clinical tasks, while their actual impact on human diagnosticians, when incorporated into clinical workflows, remains relatively unexplored. In this study, we developed a deep learning-based assistant to help pathologists differentiate between two subtypes of primary liver cancer, hepatocellular carcinoma and cholangiocarcinoma, on hematoxylin and eosin-stained whole-slide images (WSI), and evaluated its effect on the diagnostic performance of 11 pathologists with varying levels of expertise. Our model achieved accuracies of 0.885 on a validation set of 26 WSI, and 0.842 on an independent test set of 80 WSI. Although use of the assistant did not change the mean accuracy of the 11 pathologists (p = 0.184, OR = 1.281), it significantly improved the accuracy (p = 0.045, OR = 1.499) of a subset of nine pathologists who fell within well-defined experience levels (GI subspecialists, non-GI subspecialists, and trainees). In the assisted state, model accuracy significantly impacted the diagnostic decisions of all 11 pathologists. As expected, when the model’s prediction was correct, assistance significantly improved accuracy (p = 0.000, OR = 4.289), whereas when the model’s prediction was incorrect, assistance significantly decreased accuracy (p = 0.000, OR = 0.253), with both effects holding across all pathologist experience levels and case difficulty levels. Our results highlight the challenges of translating AI models into the clinical setting, and emphasize the importance of taking into account potential unintended negative consequences of model assistance when designing and testing medical AI-assistance tools.

Published in

Links

Tools

Impact of a deep learning assistant on the histopathologic classification of liver cancer

Abstract