Explaining multivariate molecular diagnostic tests via Shapley values.
Artificial intelligence
Explainability
Interpretability
Machine learning
Molecular diagnostic test
Shapley values
Journal
BMC medical informatics and decision making
ISSN: 1472-6947
Titre abrégé: BMC Med Inform Decis Mak
Pays: England
ID NLM: 101088682
Informations de publication
Date de publication:
08 07 2021
08 07 2021
Historique:
received:
29
03
2021
accepted:
29
06
2021
entrez:
9
7
2021
pubmed:
10
7
2021
medline:
10
8
2021
Statut:
epublish
Résumé
Machine learning (ML) can be an effective tool to extract information from attribute-rich molecular datasets for the generation of molecular diagnostic tests. However, the way in which the resulting scores or classifications are produced from the input data may not be transparent. Algorithmic explainability or interpretability has become a focus of ML research. Shapley values, first introduced in game theory, can provide explanations of the result generated from a specific set of input data by a complex ML algorithm. For a multivariate molecular diagnostic test in clinical use (the VeriStrat® test), we calculate and discuss the interpretation of exact Shapley values. We also employ some standard approximation techniques for Shapley value computation (local interpretable model-agnostic explanation (LIME) and Shapley Additive Explanations (SHAP) based methods) and compare the results with exact Shapley values. Exact Shapley values calculated for data collected from a cohort of 256 patients showed that the relative importance of attributes for test classification varied by sample. While all eight features used in the VeriStrat® test contributed equally to classification for some samples, other samples showed more complex patterns of attribute importance for classification generation. Exact Shapley values and Shapley-based interaction metrics were able to provide interpretable classification explanations at the sample or patient level, while patient subgroups could be defined by comparing Shapley value profiles between patients. LIME and SHAP approximation approaches, even those seeking to include correlations between attributes, produced results that were quantitatively and, in some cases qualitatively, different from the exact Shapley values. Shapley values can be used to determine the relative importance of input attributes to the result generated by a multivariate molecular diagnostic test for an individual sample or patient. Patient subgroups defined by Shapley value profiles may motivate translational research. However, correlations inherent in molecular data and the typically small ML training sets available for molecular diagnostic test development may cause some approximation methods to produce approximate Shapley values that differ both qualitatively and quantitatively from exact Shapley values. Hence, caution is advised when using approximate methods to evaluate Shapley explanations of the results of molecular diagnostic tests.
Sections du résumé
BACKGROUND
Machine learning (ML) can be an effective tool to extract information from attribute-rich molecular datasets for the generation of molecular diagnostic tests. However, the way in which the resulting scores or classifications are produced from the input data may not be transparent. Algorithmic explainability or interpretability has become a focus of ML research. Shapley values, first introduced in game theory, can provide explanations of the result generated from a specific set of input data by a complex ML algorithm.
METHODS
For a multivariate molecular diagnostic test in clinical use (the VeriStrat® test), we calculate and discuss the interpretation of exact Shapley values. We also employ some standard approximation techniques for Shapley value computation (local interpretable model-agnostic explanation (LIME) and Shapley Additive Explanations (SHAP) based methods) and compare the results with exact Shapley values.
RESULTS
Exact Shapley values calculated for data collected from a cohort of 256 patients showed that the relative importance of attributes for test classification varied by sample. While all eight features used in the VeriStrat® test contributed equally to classification for some samples, other samples showed more complex patterns of attribute importance for classification generation. Exact Shapley values and Shapley-based interaction metrics were able to provide interpretable classification explanations at the sample or patient level, while patient subgroups could be defined by comparing Shapley value profiles between patients. LIME and SHAP approximation approaches, even those seeking to include correlations between attributes, produced results that were quantitatively and, in some cases qualitatively, different from the exact Shapley values.
CONCLUSIONS
Shapley values can be used to determine the relative importance of input attributes to the result generated by a multivariate molecular diagnostic test for an individual sample or patient. Patient subgroups defined by Shapley value profiles may motivate translational research. However, correlations inherent in molecular data and the typically small ML training sets available for molecular diagnostic test development may cause some approximation methods to produce approximate Shapley values that differ both qualitatively and quantitatively from exact Shapley values. Hence, caution is advised when using approximate methods to evaluate Shapley explanations of the results of molecular diagnostic tests.
Identifiants
pubmed: 34238309
doi: 10.1186/s12911-021-01569-9
pii: 10.1186/s12911-021-01569-9
pmc: PMC8265031
doi:
Types de publication
Journal Article
Langues
eng
Sous-ensembles de citation
IM
Pagination
211Références
J Med Internet Res. 2020 Nov 6;22(11):e24018
pubmed: 33027032
FEBS Lett. 2003 Feb 27;537(1-3):166-70
pubmed: 12606051
Cancer Med. 2014 Jun;3(3):572-9
pubmed: 24574334
Curr Med Res Opin. 2020 Sep;36(9):1497-1505
pubmed: 32615813
J Natl Cancer Inst. 2007 Jun 6;99(11):838-46
pubmed: 17551144
BMC Cancer. 2018 Mar 20;18(1):310
pubmed: 29558888
Cancers (Basel). 2020 Dec 28;13(1):
pubmed: 33379188
Science. 2019 Oct 25;366(6464):447-453
pubmed: 31649194
Lancet Oncol. 2014 Jun;15(7):713-21
pubmed: 24831979
Cancers (Basel). 2020 Sep 04;12(9):
pubmed: 32899818
Sci Rep. 2021 Feb 25;11(1):4565
pubmed: 33633172
J Proteomics. 2012 Dec 5;76 Spec No.:91-101
pubmed: 22771314
BMC Med Inform Decis Mak. 2020 Nov 30;20(1):310
pubmed: 33256715
Cancer Immunol Res. 2018 Jan;6(1):79-86
pubmed: 29208646
BMC Bioinformatics. 2019 Jun 13;20(1):325
pubmed: 31196002
Clin Cancer Res. 2020 Oct 1;26(19):5188-5197
pubmed: 32631957
Cancer Epidemiol Biomarkers Prev. 2010 Feb;19(2):358-65
pubmed: 20086114
PLoS One. 2020 Nov 12;15(11):e0239172
pubmed: 33180787