Human Evaluation of Models Built for Interpretability

Isaac Lage; Emily Chen; Jeffrey He; Menaka Narayanan; Been Kim; Samuel J Gershman; Finale Doshi-Velez

Human Evaluation of Models Built for Interpretability

Proc AAAI Conf Hum Comput Crowdsourc. 2019;7(1):59-67. Epub 2019 Oct 28.

Authors

Isaac Lage¹, Emily Chen¹, Jeffrey He¹, Menaka Narayanan¹, Been Kim², Samuel J Gershman¹, Finale Doshi-Velez¹

Affiliations

¹ Harvard University.
² Google.

PMID: 33623933
PMCID: PMC7899148

Abstract

Recent years have seen a boom in interest in interpretable machine learning systems built on models that can be understood, at least to some degree, by domain experts. However, exactly what kinds of models are truly human-interpretable remains poorly understood. This work advances our understanding of precisely which factors make models interpretable in the context of decision sets, a specific class of logic-based model. We conduct carefully controlled human-subject experiments in two domains across three tasks based on human-simulatability through which we identify specific types of complexity that affect performance more heavily than others-trends that are consistent across tasks and domains. These results can inform the choice of regularizers during optimization to learn more interpretable models, and their consistency suggests that there may exist common design principles for interpretable machine learning systems.

Grants and funding

T32 LM012411/LM/NLM NIH HHS/United States