I have to fully agree. The language-processing part is really simple (partly due to the really cool markovify library [1])
You may also like my older project [2] even though it is partly in German. I'm using Markov Chains for the titles and some custom regex-based language processing for the descriptions.