Who Cited It

Long Short-Term Memory

1997 · Neural Computation · 101,359 citations · 89 from inside this corpus

Sepp Hochreiter, Jürgen Schmidhuber

The source holds an abstract for this work, but its best open-access copy is under no open licence, which does not permit us to republish the text. Read it at the source below.

Long Short-Term Memory (1997)Long Short-Term MemoryLearning long-term dependencies with gradient descent is difficult (1994)Learning long-term dependenci…Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Transla… (2014)Learning Phrase Representatio…Deep learning in neural networks: An overview (2014)Deep learning in neural netwo…A survey on deep learning in medical image analysis (2017)A survey on deep learning in …Sequence to Sequence Learning with Neural Networks (2014)Sequence to Sequence Learning…A Comprehensive Survey on Graph Neural Networks (2020)A Comprehensive Survey on Gra…Speech recognition with deep recurrent neural networks (2013)Speech recognition with deep …Explaining and Harnessing Adversarial Examples (2014)Explaining and Harnessing Adv…Graph neural networks: A review of methods and applications (2020)Graph neural networks: A revi…Fundamentals of Recurrent Neural Network (RNN) and Long Short-Term Memory (LSTM) network (2020)Fundamentals of Recurrent Neu…A Review of Recurrent Neural Networks: LSTM Cells and Network Architectures (2019)A Review of Recurrent Neural …Framewise phoneme classification with bidirectional LSTM and other neural network archite… (2005)Framewise phoneme classificat…Connectionist temporal classification (2006)Connectionist temporal classi…Learning to Forget: Continual Prediction with LSTM (2000)Learning to Forget: Continual…Deep learning and process understanding for data-driven Earth system science (2019)Deep learning and process und…Hierarchical Attention Networks for Document Classification (2016)Hierarchical Attention Networ…Neural Architectures for Named Entity Recognition (2016)Neural Architectures for Name…BNAI, NO-TOKEN, and MIND-UNITY: Pillars of a Systemic Revolution in Artificial Intelligen… (2022)BNAI, NO-TOKEN, and MIND-UNIT…Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Langu… (2022)Pre-train, Prompt, and Predic…On the difficulty of training Recurrent Neural Networks (2012)On the difficulty of training…Deep Convolutional Neural Networks for Image Classification: A Comprehensive Review (2017)Deep Convolutional Neural Net…On the importance of initialization and momentum in deep learning (2013)On the importance of initiali…Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Transla… (2014)Learning Phrase Representatio…Sequence to Sequence Learning with Neural Networks (2014)Sequence to Sequence Learning…Grandmaster level in StarCraft II using multi-agent reinforcement learning (2019)Grandmaster level in StarCraf…A Comprehensive Survey on Graph Neural Networks (2020)A Comprehensive Survey on Gra…Bidirectional LSTM-CRF Models for Sequence Tagging (2015)Bidirectional LSTM-CRF Models…Character-level Convolutional Networks for Text Classification (2015)Character-level Convolutional…Transformer-XL: Attentive Language Models beyond a Fixed-Length Context (2019)Transformer-XL: Attentive Lan…Deep learning for healthcare: review, opportunities and challenges (2017)Deep learning for healthcare:…Reservoir computing approaches to recurrent neural network training (2009)Reservoir computing approache…Long short-term memory recurrent neural network architectures for large scale acoustic mo… (2014)Long short-term memory recurr…Recent Trends in Deep Learning Based Natural Language Processing [Review Article] (2018)Recent Trends in Deep Learnin…Deep learning and its applications to machine health monitoring (2018)Deep learning and its applica…Generalizing from a Few Examples (2020)Generalizing from a Few Examp…Natural TTS Synthesis by Conditioning Wavenet on MEL Spectrogram Predictions (2018)Natural TTS Synthesis by Cond…
36 of 37 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

What cites it, inside the corpus

PaperYearCited
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Transla…201425,048
Deep learning in neural networks: An overview201418,236
A survey on deep learning in medical image analysis201715,110
Sequence to Sequence Learning with Neural Networks201413,351
A Comprehensive Survey on Graph Neural Networks20209,855
Speech recognition with deep recurrent neural networks20138,916
Explaining and Harnessing Adversarial Examples20148,146
Graph neural networks: A review of methods and applications20205,808
Fundamentals of Recurrent Neural Network (RNN) and Long Short-Term Memory (LSTM) network20205,639
A Review of Recurrent Neural Networks: LSTM Cells and Network Architectures20195,628
Framewise phoneme classification with bidirectional LSTM and other neural network archite…20055,619
Connectionist temporal classification20065,582
Learning to Forget: Continual Prediction with LSTM20005,538
Deep learning and process understanding for data-driven Earth system science20195,475
Hierarchical Attention Networks for Document Classification20164,871
Neural Architectures for Named Entity Recognition20164,482
BNAI, NO-TOKEN, and MIND-UNITY: Pillars of a Systemic Revolution in Artificial Intelligen…20224,315
Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Langu…20223,836
On the difficulty of training Recurrent Neural Networks20123,801
Deep Convolutional Neural Networks for Image Classification: A Comprehensive Review20173,570
On the importance of initialization and momentum in deep learning20133,523
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Transla…20143,518
Sequence to Sequence Learning with Neural Networks20143,514
Grandmaster level in StarCraft II using multi-agent reinforcement learning20193,508
A Comprehensive Survey on Graph Neural Networks20203,307
Bidirectional LSTM-CRF Models for Sequence Tagging20153,287
Character-level Convolutional Networks for Text Classification20153,280
Transformer-XL: Attentive Language Models beyond a Fixed-Length Context20193,202
Deep learning for healthcare: review, opportunities and challenges20173,115
Reservoir computing approaches to recurrent neural network training20092,992
Long short-term memory recurrent neural network architectures for large scale acoustic mo…20142,986
Recent Trends in Deep Learning Based Natural Language Processing [Review Article]20182,907
Deep learning and its applications to machine health monitoring20182,803
Generalizing from a Few Examples20202,761
Natural TTS Synthesis by Conditioning Wavenet on MEL Spectrogram Predictions20182,708

Links

DOI · OpenAlex record

Topics

Neural Networks and ApplicationsComputer Science
Domain Adaptation and Few-Shot LearningComputer Science

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports2 author record(s) attached.
  • supports34 reference(s) recorded.
  • supportsThe DOI's year agrees with the publication year.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:40+00:00.

sha256 7e3d99a592f7f61f…