ML Interpretability: Simple Isn’t Easy
Tim Räz Note: University of Bern, Institute of Philosophy, Länggassstrasse 49a, 3012 Bern, Switzerland. E-mail:
Abstract
The interpretability of ML models is important, but it is not clear what it amounts to. So far, most philosophers have discussed the lack of interpretability of black-box models such as neural networks, and methods such as explainable AI that aim to make these models more transparent. The goal of this paper is to clarify the nature of interpretability by focussing on the other end of the “interpretability spectrum”. The reasons why some models, linear models and decision trees, are highly interpretable will be examined, and also how more general models, MARS and GAM, retain some degree of interpretability. I find that while there is heterogeneity in how we gain interpretability, what interpretability is in particular cases can be explicated in a clear manner.
原文 arXiv:2211.13617;中英对照 + 大白话阅读 https://aha.fim.ai/paper/2211.13617v1