Performance of Three Slim Variants of The Long Short-Term Memory (LSTM) Layer

January 02, 2019 · Declared Dead · 🏛 Midwest Symposium on Circuits and Systems

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Daniel Kent, Fathi M. Salem arXiv ID 1901.00525 Category cs.NE: Neural & Evolutionary Cross-listed cs.AI, cs.LG Citations 27 Venue Midwest Symposium on Circuits and Systems Last Checked 3 months ago

Abstract

The Long Short-Term Memory (LSTM) layer is an important advancement in the field of neural networks and machine learning, allowing for effective training and impressive inference performance. LSTM-based neural networks have been successfully employed in various applications such as speech processing and language translation. The LSTM layer can be simplified by removing certain components, potentially speeding up training and runtime with limited change in performance. In particular, the recently introduced variants, called SLIM LSTMs, have shown success in initial experiments to support this view. Here, we perform computational analysis of the validation accuracy of a convolutional plus recurrent neural network architecture using comparatively the standard LSTM and three SLIM LSTM layers. We have found that some realizations of the SLIM LSTM layers can potentially perform as well as the standard LSTM layer for our considered architecture.

📄 View on arXiv 🌐 View on ar5iv 📑 PDF 🎉 Report Code Found

Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

📜 Similar Papers

In the same crypt — Neural & Evolutionary

🔮 🔮 The Ethereal

LSTM: A Search Space Odyssey

Klaus Greff, Rupesh Kumar Srivastava, ... (+3 more)

cs.NE 🏛 IEEE TNNLS 📚 6.0K cites 11 years ago

R.I.P. 👻 Ghosted

Deep Learning using Rectified Linear Units (ReLU)

Abien Fred Agarap

cs.NE 🏛 arXiv 📚 3.8K cites 8 years ago

R.I.P. 👻 Ghosted

Generative Adversarial Text to Image Synthesis

Scott Reed, Zeynep Akata, ... (+4 more)

cs.NE 🏛 ICML 📚 3.4K cites 10 years ago

R.I.P. 👻 Ghosted

Regularized Evolution for Image Classifier Architecture Search

Esteban Real, Alok Aggarwal, ... (+2 more)

cs.NE 🏛 AAAI 📚 3.2K cites 8 years ago

R.I.P. 👻 Ghosted

Temporal Ensembling for Semi-Supervised Learning

Samuli Laine, Timo Aila

cs.NE 🏛 ICLR 📚 2.8K cites 9 years ago

🌅 🌅 Old Age

Learning Structured Sparsity in Deep Neural Networks

Wei Wen, Chunpeng Wu, ... (+3 more)

cs.NE 🏛 NeurIPS 📚 2.5K cites 9 years ago

Died the same way — 👻 Ghosted

R.I.P. 👻 Ghosted

Federated Learning: Strategies for Improving Communication Efficiency

Jakub Konečný, H. Brendan McMahan, ... (+4 more)

cs.LG 🏛 arXiv 📚 5.2K cites 9 years ago

R.I.P. 👻 Ghosted

In-Datacenter Performance Analysis of a Tensor Processing Unit

Norman P. Jouppi, Cliff Young, ... (+73 more)

cs.AR 🏛 ISCA 📚 5.1K cites 9 years ago

R.I.P. 👻 Ghosted

Deep Convolutional Neural Networks for Computer-Aided Detection: CNN Architectures, Dataset Characteristics and Transfer Learning

Hoo-Chang Shin, Holger R. Roth, ... (+7 more)

cs.CV 🏛 IEEE TMI 📚 4.9K cites 10 years ago

R.I.P. 👻 Ghosted

Explanation in Artificial Intelligence: Insights from the Social Sciences

Tim Miller

cs.AI 🏛 AI 📚 4.9K cites 9 years ago