> a randomly-initialized LSTM [12] with a learned linear output layer can predict time series where traditional RNNs fail
Yeah, though I think we could have referenced a few more works in the reservoir computing area. Will keep that in mind for the next draft.
just starting to read this. It reminds me of meiosis networks + stochastic delta rule where the weights were drawn from a random distribution (normal I think).