Pith. sign in

REVIEW 1 cited by

Rational Recurrences

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1808.09357 v1 pith:73R6SQ7Z submitted 2018-08-28 cs.CL

classification cs.CL
keywords neuralmodelswfsasrationalrecurrencesclassicalconnectionfinite
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Despite the tremendous empirical success of neural models in natural language processing, many of them lack the strong intuitions that accompany classical machine learning approaches. Recently, connections have been shown between convolutional neural networks (CNNs) and weighted finite state automata (WFSAs), leading to new interpretations and insights. In this work, we show that some recurrent neural networks also share this connection to WFSAs. We characterize this connection formally, defining rational recurrences to be recurrent hidden state update functions that can be written as the Forward calculation of a finite set of WFSAs. We show that several recent neural models use rational recurrences. Our analysis provides a fresh view of these models and facilitates devising new neural architectures that draw inspiration from WFSAs. We present one such model, which performs better than two recent baselines on language modeling and text classification. Our results demonstrate that transferring intuitions from classical models like WFSAs can be an effective approach to designing and understanding neural models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Aligning Brain Activity with Advanced Transformer Models: Exploring the Role of Punctuation in Semantic Processing

    cs.CL 2025-01 conditional novelty 4.0 of 10

    RoBERTa and DistilBERT align slightly better with fMRI brain responses than BERT, and removing punctuation yields a small improvement in BERT's later-layer alignment.

Pith tools