narengogi.github.io
some fun writing CUDA accelerated code to cluster the data and then find similarity scores
Reply on news.ycombinator.com