Pith. sign in

REVIEW 8 cited by

Understanding LSTM -- a tutorial into Long Short-Term Memory Recurrent Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1909.09586 v1 pith:C5L6G4HG submitted 2019-09-12 cs.NE cs.CLcs.LG

classification cs.NEcs.CLcs.LG
keywords understandingwelllongmemorynetworksneuralpublicationsrecurrent
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Long Short-Term Memory Recurrent Neural Networks (LSTM-RNN) are one of the most powerful dynamic classifiers publicly known. The network itself and the related learning algorithms are reasonably well documented to get an idea how it works. This paper will shed more light into understanding how LSTM-RNNs evolved and why they work impressively well, focusing on the early, ground-breaking publications. We significantly improved documentation and fixed a number of errors and inconsistencies that accumulated in previous publications. To support understanding we as well revised and unified the notation used.

Discussion (0). Sign in to comment.

Forward citations

Cited by 8 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Urdu Katib Handwritten Dataset: A Historical Document Dataset for Offline Urdu Handwritten Text Recognition with CRNN-Based Baseline Evaluation

    cs.CV 2026-06 unverdicted novelty 7.0 of 10

    Presents UKHD, the first historical offline Urdu handwritten text lines dataset from Katib materials, and benchmarks CRNN-based models with CNN-BGRU-CTC showing lowest CER and WER.

  2. BEAM: Bi-level Memory-adaptive Algorithmic Evolution for LLM-Powered Heuristic Design

    cs.AI 2026-04 unverdicted novelty 7.0 of 10

    BEAM reformulates LLM-based heuristic design as bi-level optimization using GA for structures, MCTS for placeholders, and adaptive memory to outperform prior single-layer methods on CVRP and MIS tasks.

  3. Differential-UMamba: Rethinking Tumor Segmentation Under Limited Data Scenarios

    cs.CV 2025-07 unverdicted novelty 6.0 of 10

    Diff-UMamba combines UNet with Mamba and adds signal differencing for noise reduction, yielding 1-3% segmentation gains on public medical datasets and 4-5% on a small internal lung cancer dataset under limited data co...

  4. Risk Averse Alert Prioritization for IDS Using Subnormal Gaussian Fuzzy Models

    cs.CR 2026-05 unverdicted novelty 5.0 of 10

    A subnormal Gaussian fuzzy number model for risk-averse IDS alert prioritization that shows improved robustness over baselines on CIC-IDS2017 and NSL-KDD under detector degradation.

  5. Neural Inhibition Improves Dynamic Routing and Mixture of Experts

    cs.LG 2025-07 conditional novelty 5.0 of 10

    Neural inhibition gating on MoE router inputs improves a synthetic digit/squares benchmark by about four points over plain MoE, but the language-model evidence is unreliable.

  6. Automating the Deep Space Network Data Systems; A Case Study in Adaptive Anomaly Detection through Agentic AI

    cs.LG 2025-08 reject novelty 4.0 of 10

    An internship report integrates reconstruction-based deep learning, Q-learning, and a Mistral LLM into an agentic workflow for DSN anomaly detection, without reporting any performance metrics.

  7. Leveraging Large Language Models for Sentiment Analysis: Multi-Modal Analysis of Decentraland's MANA Token

    cs.CL 2026-04 unverdicted novelty 3.0 of 10

    Incorporating BERT-derived Discord sentiment into an LSTM improves MANA token return forecasts over a historical-price baseline.

  8. Exploitation of Hidden Context in Dynamic Movement Forecasting: A Neural Network Journey from Recurrent to Graph Neural Networks and General Purpose Transformers

    cs.LG 2026-05 unverdicted novelty 2.0 of 10

    Empirical comparison of LSTM, GNN, and Transformer architectures for NBA trajectory forecasting finds hybrid LSTM with contextual information yields lowest FDE of 1.51m over horizons up to 2s.

Pith tools