Pith. sign in

REVIEW 1 cited by

Performance of Three Slim Variants of The Long Short-Term Memory (LSTM) Layer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1901.00525 v1 pith:LNVIE7RR submitted 2019-01-02 cs.NE cs.AIcs.LG

classification cs.NEcs.AIcs.LG
keywords lstmlayerslimneuralperformancearchitecturelayerslong
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

The Long Short-Term Memory (LSTM) layer is an important advancement in the field of neural networks and machine learning, allowing for effective training and impressive inference performance. LSTM-based neural networks have been successfully employed in various applications such as speech processing and language translation. The LSTM layer can be simplified by removing certain components, potentially speeding up training and runtime with limited change in performance. In particular, the recently introduced variants, called SLIM LSTMs, have shown success in initial experiments to support this view. Here, we perform computational analysis of the validation accuracy of a convolutional plus recurrent neural network architecture using comparatively the standard LSTM and three SLIM LSTM layers. We have found that some realizations of the SLIM LSTM layers can potentially perform as well as the standard LSTM layer for our considered architecture.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Restricted Recurrent Neural Networks

    cs.CL 2019-08 conditional novelty 5.0 of 10

    Partial sharing of input and hidden-state weights in RNN, LSTM, and GRU yields about 50% parameter reduction with roughly unchanged perplexity on two language-modeling benchmarks.

Pith tools