REVIEW 8 cited by
Understanding LSTM -- a tutorial into Long Short-Term Memory Recurrent Neural Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Long Short-Term Memory Recurrent Neural Networks (LSTM-RNN) are one of the most powerful dynamic classifiers publicly known. The network itself and the related learning algorithms are reasonably well documented to get an idea how it works. This paper will shed more light into understanding how LSTM-RNNs evolved and why they work impressively well, focusing on the early, ground-breaking publications. We significantly improved documentation and fixed a number of errors and inconsistencies that accumulated in previous publications. To support understanding we as well revised and unified the notation used.
Forward citations
Cited by 8 Pith papers
-
Urdu Katib Handwritten Dataset: A Historical Document Dataset for Offline Urdu Handwritten Text Recognition with CRNN-Based Baseline Evaluation
Presents UKHD, the first historical offline Urdu handwritten text lines dataset from Katib materials, and benchmarks CRNN-based models with CNN-BGRU-CTC showing lowest CER and WER.
-
BEAM: Bi-level Memory-adaptive Algorithmic Evolution for LLM-Powered Heuristic Design
BEAM reformulates LLM-based heuristic design as bi-level optimization using GA for structures, MCTS for placeholders, and adaptive memory to outperform prior single-layer methods on CVRP and MIS tasks.
-
Differential-UMamba: Rethinking Tumor Segmentation Under Limited Data Scenarios
Diff-UMamba combines UNet with Mamba and adds signal differencing for noise reduction, yielding 1-3% segmentation gains on public medical datasets and 4-5% on a small internal lung cancer dataset under limited data co...
-
Risk Averse Alert Prioritization for IDS Using Subnormal Gaussian Fuzzy Models
A subnormal Gaussian fuzzy number model for risk-averse IDS alert prioritization that shows improved robustness over baselines on CIC-IDS2017 and NSL-KDD under detector degradation.
-
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
Neural inhibition gating on MoE router inputs improves a synthetic digit/squares benchmark by about four points over plain MoE, but the language-model evidence is unreliable.
-
Automating the Deep Space Network Data Systems; A Case Study in Adaptive Anomaly Detection through Agentic AI
An internship report integrates reconstruction-based deep learning, Q-learning, and a Mistral LLM into an agentic workflow for DSN anomaly detection, without reporting any performance metrics.
-
Leveraging Large Language Models for Sentiment Analysis: Multi-Modal Analysis of Decentraland's MANA Token
Incorporating BERT-derived Discord sentiment into an LSTM improves MANA token return forecasts over a historical-price baseline.
-
Exploitation of Hidden Context in Dynamic Movement Forecasting: A Neural Network Journey from Recurrent to Graph Neural Networks and General Purpose Transformers
Empirical comparison of LSTM, GNN, and Transformer architectures for NBA trajectory forecasting finds hybrid LSTM with contextual information yields lowest FDE of 1.51m over horizons up to 2s.
Discussion (0). Sign in to comment.