Less is More: Discovering Concise Network Explanations

Kondapaneni, Neehar; Marks, Markus; MacAodha, Oisin; Perona, Pietro

Computer Science > Computer Vision and Pattern Recognition

arXiv:2405.15243 (cs)

[Submitted on 24 May 2024 (v1), last revised 14 Jun 2024 (this version, v2)]

Title:Less is More: Discovering Concise Network Explanations

Authors:Neehar Kondapaneni, Markus Marks, Oisin MacAodha, Pietro Perona

View PDF HTML (experimental)

Abstract:We introduce Discovering Conceptual Network Explanations (DCNE), a new approach for generating human-comprehensible visual explanations to enhance the interpretability of deep neural image classifiers. Our method automatically finds visual explanations that are critical for discriminating between classes. This is achieved by simultaneously optimizing three criteria: the explanations should be few, diverse, and human-interpretable. Our approach builds on the recently introduced Concept Relevance Propagation (CRP) explainability method. While CRP is effective at describing individual neuronal activations, it generates too many concepts, which impacts human comprehension. Instead, DCNE selects the few most important explanations. We introduce a new evaluation dataset centered on the challenging task of classifying birds, enabling us to compare the alignment of DCNE's explanations to those of human expert-defined ones. Compared to existing eXplainable Artificial Intelligence (XAI) methods, DCNE has a desirable trade-off between conciseness and completeness when summarizing network explanations. It produces 1/30 of CRP's explanations while only resulting in a slight reduction in explanation quality. DCNE represents a step forward in making neural network decisions accessible and interpretable to humans, providing a valuable tool for both researchers and practitioners in XAI and model alignment.

Comments:	9 pages, 5 figures; ICLR Re-Align Workshop 2024; Project Page: this https URL Github: this https URL
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2405.15243 [cs.CV]
	(or arXiv:2405.15243v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2405.15243

Submission history

From: Neehar Kondapaneni [view email]
[v1] Fri, 24 May 2024 06:10:23 UTC (34,782 KB)
[v2] Fri, 14 Jun 2024 03:11:43 UTC (34,782 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Less is More: Discovering Concise Network Explanations

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Less is More: Discovering Concise Network Explanations

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators