Towards interpretable-by-design deep learning algorithms
Angelov, Plamen, Kangin, Dmitry, Zhang, Ziyang
–arXiv.org Artificial Intelligence
The proposed framework named IDEAL (Interpretable-by-design DEep learning ALgorithms) recasts the standard supervised classification problem into a function of similarity to a set of prototypes derived from the training data, while taking advantage of existing latent spaces of large neural networks forming so-called Foundation Models (FM). This addresses the issue of explainability (stage B) while retaining the benefits from the tremendous achievements offered by DL models (e.g., visual transformers, ViT) pre-trained on huge data sets such as IG-3.6B + ImageNet-1K or LVD-142M (stage A). We show that one can turn such DL models into conceptually simpler, explainable-through-prototypes ones. The key findings can be summarized as follows: (1) the proposed models are interpretable through prototypes, mitigating the issue of confounded interpretations, (2) the proposed IDEAL framework circumvents the issue of catastrophic forgetting allowing efficient class-incremental learning, and (3) the proposed IDEAL approach demonstrates that ViT architectures narrow the gap between finetuned and non-finetuned models allowing for transfer learning in a fraction of time \textbf{without} finetuning of the feature space on a target dataset with iterative supervised methods.
arXiv.org Artificial Intelligence
Nov-19-2023
- Country:
- North America
- Greenland (0.04)
- United States
- Maine (0.04)
- District of Columbia > Washington (0.04)
- Pennsylvania > Allegheny County
- Pittsburgh (0.04)
- California > Alameda County
- Oakland (0.04)
- Canada > Ontario
- Toronto (0.14)
- Europe > Netherlands
- North Holland > Amsterdam (0.04)
- Asia > Middle East
- Israel > Tel Aviv District > Tel Aviv (0.04)
- North America
- Genre:
- Research Report
- Promising Solution (0.46)
- New Finding (0.46)
- Research Report
- Industry:
- Education (1.00)
- Health & Medicine (0.67)
- Technology: