Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Learning local discrete features in explainable-by-design convolutional neural networks

Oct 31, 2024

Pantelis I. Kaplanoglou, Konstantinos Diamantaras

Figure 1 for Learning local discrete features in explainable-by-design convolutional neural networks

Figure 2 for Learning local discrete features in explainable-by-design convolutional neural networks

Figure 3 for Learning local discrete features in explainable-by-design convolutional neural networks

Figure 4 for Learning local discrete features in explainable-by-design convolutional neural networks

Share this with someone who'll enjoy it:

Abstract:Our proposed framework attempts to break the trade-off between performance and explainability by introducing an explainable-by-design convolutional neural network (CNN) based on the lateral inhibition mechanism. The ExplaiNet model consists of the predictor, that is a high-accuracy CNN with residual or dense skip connections, and the explainer probabilistic graph that expresses the spatial interactions of the network neurons. The value on each graph node is a local discrete feature (LDF) vector, a patch descriptor that represents the indices of antagonistic neurons ordered by the strength of their activations, which are learned with gradient descent. Using LDFs as sequences we can increase the conciseness of explanations by repurposing EXTREME, an EM-based sequence motif discovery method that is typically used in molecular biology. Having a discrete feature motif matrix for each one of intermediate image representations, instead of a continuous activation tensor, allows us to leverage the inherent explainability of Bayesian networks. By collecting observations and directly calculating probabilities, we can explain causal relationships between motifs of adjacent levels and attribute the model's output to global motifs. Moreover, experiments on various tiny image benchmark datasets confirm that our predictor ensures the same level of performance as the baseline architecture for a given count of parameters and/or layers. Our novel method shows promise to exceed this performance while providing an additional stream of explanations. In the solved MNIST classification task, it reaches a comparable to the state-of-the-art performance for single models, using standard training setup and 0.75 million parameters.

View paper on

Share this with someone who'll enjoy it:

Title:Learning local discrete features in explainable-by-design convolutional neural networks

Paper and Code