The impact of site-specific digital histology signatures on deep learning model accuracy and bias

被引:123
作者
Howard, Frederick M. [1 ]
Dolezal, James [1 ]
Kochanny, Sara [1 ]
Schulte, Jefree [2 ]
Chen, Heather [2 ]
Heij, Lara [3 ,4 ]
Huo, Dezheng [5 ,6 ]
Nanda, Rita [1 ,6 ]
Olopade, Olufunmilayo I. [1 ,6 ]
Kather, Jakob N. [7 ,8 ,9 ]
Cipriani, Nicole [2 ,6 ]
Grossman, Robert L. [1 ,6 ]
Pearson, Alexander T. [1 ,6 ]
机构
[1] Univ Chicago, Dept Med, Sect Hematol Oncol, 5841 S Maryland Ave, Chicago, IL 60637 USA
[2] Univ Chicago, Dept Pathol, 5841 S Maryland Ave, Chicago, IL 60637 USA
[3] Univ Hosp RWTH Aachen, Dept Surg & Transplantat, Aachen, Germany
[4] Univ Hosp RWTH Aachen, Inst Pathol, Aachen, Germany
[5] Univ Chicago, Dept Publ Hlth Sci, Chicago, IL 60637 USA
[6] Univ Chicago Comprehens Canc Ctr, Chicago, IL USA
[7] Univ Hosp RWTH Aachen, Dept Med 3, Aachen, Germany
[8] Univ Leeds, Leeds Inst Med Res St Jamess, Pathol & Data Analyt, Leeds, W Yorkshire, England
[9] Univ Heidelberg Hosp, Natl Ctr Tumor Dis, Med Oncol, Heidelberg, Germany
关键词
COMPREHENSIVE GENOMIC CHARACTERIZATION; OPERATING CHARACTERISTIC CURVES; BREAST-CANCER; MITOSIS DETECTION; HEALTH-CARE; HISTOPATHOLOGY; ANCESTRY; RESOURCE; BIOLOGY; AREAS;
D O I
10.1038/s41467-021-24698-1
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
The Cancer Genome Atlas (TCGA) is one of the largest biorepositories of digital histology. Deep learning (DL) models have been trained on TCGA to predict numerous features directly from histology, including survival, gene expression patterns, and driver mutations. However, we demonstrate that these features vary substantially across tissue submitting sites in TCGA for over 3,000 patients with six cancer subtypes. Additionally, we show that histologic image differences between submitting sites can easily be identified with DL. Site detection remains possible despite commonly used color normalization and augmentation methods, and we quantify the image characteristics constituting this site-specific digital histology signature. We demonstrate that these site-specific signatures lead to biased accuracy for prediction of features including survival, genomic mutations, and tumor stage. Furthermore, ethnicity can also be inferred from site-specific signatures, which must be accounted for to ensure equitable application of DL. These site-specific signatures can lead to overoptimistic estimates of model performance, and we propose a quadratic programming method that abrogates this bias by ensuring models are not trained and validated on samples from the same site. Deep learning models have been trained on The Cancer Genome Atlas to predict numerous features directly from histology, including survival, gene expression patterns, and driver mutations. Here, the authors demonstrate that site-specific histologic signatures can lead to biased estimates of accuracy for such models, and propose a method to minimize such bias.
引用
收藏
页数:13
相关论文
共 81 条
  • [1] AggNet: Deep Learning From Crowds for Mitosis Detection in Breast Cancer Histology Images
    Albarqouni, Shadi
    Baur, Christoph
    Achilles, Felix
    Belagiannis, Vasileios
    Demirci, Stefanie
    Navab, Nassir
    [J]. IEEE TRANSACTIONS ON MEDICAL IMAGING, 2016, 35 (05) : 1313 - 1321
  • [2] A High-Performance System for Robust Stain Normalization of Whole-Slide Images in Histopathology
    Anghel, Andreea
    Stanisavljevic, Milos
    Andani, Sonali
    Papandreou, Nikolaos
    Rueschoff, Jan Hendrick
    Wild, Peter
    Gabrani, Maria
    Pozidis, Haralampos
    [J]. FRONTIERS IN MEDICINE, 2019, 6
  • [3] [Anonymous], 2009, Proceedings of the Optical Tissue Image analysis in Microscopy, Histopathology and Endoscopy (MICCAI Workshop)
  • [4] [Anonymous], 2012, J. Signal Inf. Process., DOI [DOI 10.4236/JSIP.2012.32019, 10.4236/jsip.2012.32019]
  • [5] Identifying transcriptomic correlates of histology using deep learning
    Badea, Liviu
    Stanescu, Emil
    [J]. PLOS ONE, 2020, 15 (11):
  • [6] Missed opportunities: Racial disparities in adjuvant breast cancer treatment
    Bickell, NA
    Wang, JJ
    Oluwole, S
    Schrag, D
    Godfrey, H
    Hiotis, K
    Mendez, J
    Guth, AA
    [J]. JOURNAL OF CLINICAL ONCOLOGY, 2006, 24 (09) : 1357 - 1362
  • [7] Boyd L., 2004, Convex Optimization, DOI DOI 10.1017/CBO9780511804441
  • [8] Automated deep-learning system for Gleason grading of prostate cancer using biopsies: a diagnostic study
    Bulten, Wouter
    Pinckaers, Hans
    van Boven, Hester
    Vink, Robert
    de Bel, Thomas
    van Ginneken, Bram
    van der Laak, Jeroen
    Hulsbergen-van de Kaa, Christina
    Litjens, Geert
    [J]. LANCET ONCOLOGY, 2020, 21 (02) : 233 - 241
  • [9] Byfield P, 2019, PETER554 STAINTOOLS, DOI [10.5281/zenodo.3403170, DOI 10.5281/ZENODO.3403170]
  • [10] Cancer Genome Atlas Research Network, 2018, Nature, V559, pE12, DOI [10.1038/nature13385, 10.1038/s41586-018-0228-6]