🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Training

Were RNNs All We Needed?

Free while signed in. Answers cite the passages they came from.

First page
Were RNNs All We Needed?
The curator’s take

revisits RNNs and shows that by removing the hidden states from input, forget, and update gates RNNs can be efficiently trained in parallel; this is possible because with this change architectures like LSTMs and GRUs no longer require backpropagate through time (BPTT); they introduce minLSTMs and minGRUs that are 175x faster for a 512 sequence length.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack