I am so curious about your implementation, for instance, what sort of preprocessing did you have to carry out? I had written a script sometime back to analyze Paul Graham's essays (link: https://github.com/futureUnsure/pg-essay-lda), and had to remove date and times because they appeared a lot and distorted the top topics. I'm wondering if you had to do something similar for text that described code?
Also, did you write an LDA library yourself or did you leverage an existing library?
I apologize in advance if my questions sound naive/stupid, am just a noob...