Limitations of Deep Learning Attention Mechanisms in Clinical Research: Empirical Case Study Based on the Korean Diabetic Disease Setting

Junetae Kim; Sangwon Lee; Eugene Hwang; Kwang Sun Ryu; Hanseok Jeong; Jae Wook Lee; Yul Hwangbo; Kui Son Choi; Hyo Soung Cha

doi:10.2196/18418

Limitations of Deep Learning Attention Mechanisms in Clinical Research: Empirical Case Study Based on the Korean Diabetic Disease Setting

J Med Internet Res. 2020 Dec 16;22(12):e18418. doi: 10.2196/18418.

Authors

Junetae Kim^#^{1

2

3}, Sangwon Lee^#², Eugene Hwang⁴, Kwang Sun Ryu², Hanseok Jeong¹, Jae Wook Lee⁵, Yul Hwangbo^{3

6}, Kui Son Choi^{1

2}, Hyo Soung Cha^{1

2}

Affiliations

¹ Graduate School of Cancer Science and Policy, National Cancer Center, Goyang-si, Gyeonggi-do, Republic of Korea.
² Cancer Data Center, National Cancer Control Institute, National Cancer Center, Goyang-si, Gyeonggi-do, Republic of Korea.
³ Healthcare AI Team, Healthcare Platform Center, National Cancer Center, Goyang-si, Gyeonggi-do, Republic of Korea.
⁴ School of Management Engineering, Korea Advanced Institute of Science and Technology, Seoul, Republic of Korea.
⁵ Division of Nephrology, Department of Internal Medicine, National Cancer Center, Goyang-si, Gyeonggi-do, Republic of Korea.
⁶ Division of Endocrinology, Department of Internal Medicine, National Cancer Center, Goyang-si, Gyeonggi-do, Republic of Korea.

^# Contributed equally.

PMID: 33325832
PMCID: PMC7773508
DOI: 10.2196/18418

Abstract

Background: Despite excellent prediction performance, noninterpretability has undermined the value of applying deep-learning algorithms in clinical practice. To overcome this limitation, attention mechanism has been introduced to clinical research as an explanatory modeling method. However, potential limitations of using this attractive method have not been clarified to clinical researchers. Furthermore, there has been a lack of introductory information explaining attention mechanisms to clinical researchers.

Objective: The aim of this study was to introduce the basic concepts and design approaches of attention mechanisms. In addition, we aimed to empirically assess the potential limitations of current attention mechanisms in terms of prediction and interpretability performance.

Methods: First, the basic concepts and several key considerations regarding attention mechanisms were identified. Second, four approaches to attention mechanisms were suggested according to a two-dimensional framework based on the degrees of freedom and uncertainty awareness. Third, the prediction performance, probability reliability, concentration of variable importance, consistency of attention results, and generalizability of attention results to conventional statistics were assessed in the diabetic classification modeling setting. Fourth, the potential limitations of attention mechanisms were considered.

Results: Prediction performance was very high for all models. Probability reliability was high in models with uncertainty awareness. Variable importance was concentrated in several variables when uncertainty awareness was not considered. The consistency of attention results was high when uncertainty awareness was considered. The generalizability of attention results to conventional statistics was poor regardless of the modeling approach.

Conclusions: The attention mechanism is an attractive technique with potential to be very promising in the future. However, it may not yet be desirable to rely on this method to assess variable importance in clinical settings. Therefore, along with theoretical studies enhancing attention mechanisms, more empirical studies investigating potential limitations should be encouraged.

Keywords: Bayesian deep learning; artificial intelligence; attention; deep learning; explainable artificial intelligence; health data; uncertainty awareness.

©Junetae Kim, Sangwon Lee, Eugene Hwang, Kwang Sun Ryu, Hanseok Jeong, Jae Wook Lee, Yul Hwangbo, Kui Son Choi, Hyo Soung Cha. Originally published in the Journal of Medical Internet Research (http://www.jmir.org), 16.12.2020.

Publication types

Research Support, Non-U.S. Gov't

MeSH terms

Algorithms
Deep Learning / standards*
Diabetes Mellitus / epidemiology*
Empirical Research
Humans
Reproducibility of Results
Republic of Korea
Research Design