2). Were RNNs All We Needed? – revisits RNNs and shows that by removing the hidden states from input, forget, and update gates RNNs can be efficiently trained in parallel.
RNNs Revisited: Efficient Parallel Training Without Hidden States
By
–

By
–

2). Were RNNs All We Needed? – revisits RNNs and shows that by removing the hidden states from input, forget, and update gates RNNs can be efficiently trained in parallel.