Fantastic article! How does this compare to ANN for the use case?
https://erikbern.com/2015/10/01/nearest-neighbors-and-vector...
https://erikbern.com/2015/10/01/nearest-neighbors-and-vector...
In fact the described forest of trees schema can probably be interpreted as an LSH.
Disclaimer: I haven't touched this stuff for more than 10 years. Don't know what's the state of the art now.
There is another good and more technical explanation (using bands) in chapter 2 of mining massive datasets by Leskovec, Rajaraman and Ullman.
I have dim memories of using random k basis vectors to convert high dimensionality feature vectors to k dimensions, and doing m times to generate multiple projections as part of a an LSH schema. Min-hashing might have been involved.