vix.ing · top · new · best · stats · spec

Towards Benchmarking the Utility of Explanations for Model Debugging

2021/05/10 by Idahl, Maximilian, Lyu, Lijun, Gadiraju, Ujwal +1
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG)

paper · doi:10.48550/arxiv.2105.04505

Abstract

Post-hoc explanation methods are an important class of approaches that help understand the rationale underlying a trained model's decision. But how useful are they for an end-user towards accomplishing a given task? In this vision paper, we argue the need for a benchmark to facilitate evaluations of the utility of post-hoc explanation methods. As a first step to this end, we enumerate desirable properties that such a benchmark should possess for the task of debugging text classifiers. Additionally, we highlight that such a benchmark facilitates not only assessing the effectiveness of explanations but also their efficiency.

Related