Right for the Wrong Scientific Reasons: Revising Deep Networks by Interacting with their Explanations
Schramowski, Patrick, Stammer, Wolfgang, Teso, Stefano, Brugger, Anna, Shao, Xiaoting, Luigs, Hans-Georg, Mahlein, Anne-Katrin, Kersting, Kristian
–arXiv.org Artificial Intelligence
Right for the Wrong Scientific Reasons: Revising Deep Networks by Interacting with their Explanations Patrick Schramowski, Wolfgang Stammer, Stefano Teso, Anna Brugger, Franziska Herbert, Xiaoting Shao, Hans-Georg Luigs Anne-Katrin Mahlein & Kristian Kersting Abstract Deep neural networks have shown excellent performances in many real-world applications. Unfortunately, they may show "Clever Hans"-like behavior--making use of confounding factors within datasets--to achieve high performance. In this work we introduce the novel learning setting of explanatory interactive learning (XIL) and illustrate its benefits on a plant phenotyping research task. XIL adds the scientist into the training loop such that she interactively revises the original model via providing feedback on its explanations. Our experimental results demonstrate that XIL can help avoiding Clever Hans moments in machine learning and encourages (or discourages, if appropriate) trust into the underlying model. Imagine a plant phenotyping team attempting to characterize crop resistance to plant pathogens. The plant physiologist records a larger amount of hyperspectral imaging data. Impressed by the results of deep learning in other scientific areas, she wants to establish similar results for phenotyping. Consequently, she asks a machine learning expert to apply deep learning to analyze the data. Luckily, the resulting predictive accuracy is very high. The plant physiologist, however, remains skeptical. The results are "too good, to be true". Checking the decision process of the deep model using explainable artificial intelligence (AI), the machine learning expert is flabbergasted to find that the learned deep model uses clues within the data that do not relate to the biological problem at hand, so-called confounding factors. The physiologist loses trust in AI and turns away from it, proclaiming it to be useless. Indeed, the seminal paper of Lapuschkin et al. [3] helps in "unmasking Clever Hans predictors and assessing what machines really learn". However, rather than proclaiming, as the plant physiologist might, that the machines have learned the right predictions for wrong reasons and can therefore not be trusted, we here showcase that interactions between the learning system and the human user can correct the model towards making the right predictions for the right reasons. This may also increase the trust in machine learning models.
arXiv.org Artificial Intelligence
Jan-31-2020
- Country:
- North America > United States (0.04)
- Europe
- Netherlands > North Holland
- Haarlem (0.04)
- Germany
- Lower Saxony > Gottingen (0.04)
- North Rhine-Westphalia > Cologne Region
- Hesse > Darmstadt Region
- Darmstadt (0.04)
- Belgium > Flanders
- Flemish Brabant > Leuven (0.04)
- Netherlands > North Holland
- Asia
- Middle East > Jordan (0.04)
- Japan > Honshū
- Kantō > Tokyo Metropolis Prefecture > Tokyo (0.14)
- Genre:
- Research Report
- New Finding (1.00)
- Experimental Study (1.00)
- Research Report
- Industry:
- Health & Medicine (1.00)
- Education > Educational Setting (0.49)
- Technology: