Skip to main navigation Skip to search Skip to main content

A Comprehensive Survey on Self-Interpretable Neural Networks

  • Yang Ji
  • , Ying Sun
  • , Yuting Zhang
  • , Zhigaoyuan Wang
  • , Yuanxin Zhuang
  • , Zheng Gong
  • , Dazhong Shen
  • , Chuan Qin
  • , Hengshu Zhu
  • , Hui Xiong

Research output: Contribution to journalReview articlepeer-review

Abstract

Neural networks have achieved remarkable success across various fields. However, the lack of interpretability limits their practical use, particularly in critical decision-making scenarios. Posthoc interpretability, which provides explanations for pretrained models, is often at risk of fidelity and robustness. This has inspired a rising interest in self-interpretable neural networks (SINNs), which inherently reveal the prediction rationale through model structures. Despite this progress, existing research remains fragmented, relying on intuitive designs tailored to specific tasks. To bridge these efforts and foster a unified framework, we first collect and review existing works on SINNs and provide a structured summary of their methodologies from five key perspectives: attribution-based, function-based, concept-based, prototype-based, and rule-based self-interpretation. We also present concrete, visualized examples of model explanations and discuss their applicability across diverse scenarios, including image, text, graph data, and deep reinforcement learning (DRL). Additionally, we summarize existing evaluation metrics for self-interpretation and identify open challenges in this field, offering insights for future research.

Original languageEnglish (US)
Pages (from-to)783-813
Number of pages31
JournalProceedings of the IEEE
Volume113
Issue number8
DOIs
StatePublished - 2025

All Science Journal Classification (ASJC) codes

  • General Computer Science
  • Electrical and Electronic Engineering

Fingerprint

Dive into the research topics of 'A Comprehensive Survey on Self-Interpretable Neural Networks'. Together they form a unique fingerprint.

Cite this