Pith. sign in

Interpreting CNNs via Decision Trees

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

This paper aims to quantitatively explain rationales of each prediction that is made by a pre-trained convolutional neural network (CNN). We propose to learn a decision tree, which clarifies the specific reason for each prediction made by the CNN at the semantic level. I.e., the decision tree decomposes feature representations in high conv-layers of the CNN into elementary concepts of object parts. In this way, the decision tree tells people which object parts activate which filters for the prediction and how much they contribute to the prediction score. Such semantic and quantitative explanations for CNN predictions have specific values beyond the traditional pixel-level analysis of CNNs. More specifically, our method mines all potential decision modes of the CNN, where each mode represents a common case of how the CNN uses object parts for prediction. The decision tree organizes all potential decision modes in a coarse-to-fine manner to explain CNN predictions at different fine-grained levels. Experiments have demonstrated the effectiveness of the proposed method.

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2019 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Facial age estimation by deep residual decision making

cs.CV · 2019-08-28 · conditional · novelty 3.0

Incorporating residual learning into a deep neural decision forest yields age-estimation accuracy comparable to the prior deep regression forest with about 4x less compute, plus gradient-based routing saliency maps.

citing papers explorer

Showing 1 of 1 citing paper.

  • Facial age estimation by deep residual decision making cs.CV · 2019-08-28 · conditional · none · ref 36 · internal anchor

    Incorporating residual learning into a deep neural decision forest yields age-estimation accuracy comparable to the prior deep regression forest with about 4x less compute, plus gradient-based routing saliency maps.