Robustness of convolutional neural networks in recognition of pigmented skin lesions

Maron, Roman C. and Haggenmueller, Sarah and von Kalle, Christof and Utikal, Jochen S. and Meier, Friedegund and Gellrich, Frank F. and Hauschild, Axel and French, Lars E. and Schlaak, Max and Ghoreschi, Kamran and Kutzner, Heinz and Heppt, Markus V. and Haferkamp, Sebastian and Sondermann, Wiebke and Schadendorf, Dirk and Schilling, Bastian and Hekler, Achim and Krieghoff-Henning, Eva and Kather, Jakob N. and Froehling, Stefan and Lipka, Daniel B. and Brinker, Titus J. (2021) Robustness of convolutional neural networks in recognition of pigmented skin lesions. EUROPEAN JOURNAL OF CANCER, 145. pp. 81-91. ISSN 0959-8049, 1879-0852

Full text not available from this repository. (Request a copy)

Abstract

Background: A basic requirement for artificial intelligence (AI)-based image analysis systems, which are to be integrated into clinical practice, is a high robustness. Minor changes in how those images are acquired, for example, during routine skin cancer screening, should not change the diagnosis of such assistance systems. Objective: To quantify to what extent minor image perturbations affect the convolutional neural network (CNN)-mediated skin lesion classification and to evaluate three possible solutions for this problem (additional data augmentation, test-time augmentation, anti-aliasing). Methods: We trained three commonly used CNN architectures to differentiate between dermoscopic melanoma and nevus images. Subsequently, their performance and susceptibility to minor changes ('brittleness') was tested on two distinct test sets with multiple images per lesion. For the first set, image changes, such as rotations or zooms, were generated artificially. The second set contained natural changes that stemmed from multiple photographs taken of the same lesions. Results: All architectures exhibited brittleness on the artificial and natural test set. The three reviewed methods were able to decrease brittleness to varying degrees while still maintaining performance. The observed improvement was greater for the artificial than for the natural test set, where enhancements were minor. Conclusions: Minor image changes, relatively inconspicuous for humans, can have an effect on the robustness of CNNs differentiating skin lesions. By the methods tested here, this effect can be reduced, but not fully eliminated. Thus, further research to sustain the performance of AI classifiers is needed to facilitate the translation of such systems into the clinic. (C) 2020 The Author(s). Published by Elsevier Ltd.

Item Type: Article
Uncontrolled Keywords: IMAGE CLASSIFICATION; ARTIFICIAL-INTELLIGENCE; LEVEL CLASSIFICATION; DERMATOLOGISTS; CANCER; SUPERIOR; Artificial intelligence; Machine learning; Deep learning; Neural networks; Dermatology; Skin neoplasms; Melanoma; Nevus
Subjects: 600 Technology > 610 Medical sciences Medicine
Divisions: Medicine > Lehrstuhl für Dermatologie und Venerologie
Depositing User: Dr. Gernot Deinzer
Date Deposited: 27 Sep 2022 09:09
Last Modified: 27 Sep 2022 09:09
URI: https://pred.uni-regensburg.de/id/eprint/48043

Actions (login required)

View Item View Item