Pith. sign in

REVIEW

Improving Interpretability via Explicit Word Interaction Graph Layer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2302.02016 v1 pith:X7RDEWQD submitted 2023-02-03 cs.CL cs.AI

classification cs.CLcs.AI
keywords layerinterpretabilitymodelswordgraphimprovinginteractionneural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent NLP literature has seen growing interest in improving model interpretability. Along this direction, we propose a trainable neural network layer that learns a global interaction graph between words and then selects more informative words using the learned word interactions. Our layer, we call WIGRAPH, can plug into any neural network-based NLP text classifiers right after its word embedding layer. Across multiple SOTA NLP models and various NLP datasets, we demonstrate that adding the WIGRAPH layer substantially improves NLP models' interpretability and enhances models' prediction performance at the same time.

Discussion (0). Continue with ORCID to comment.

Pith tools