HN Hiring
hnhiring.me
hnhiring.me
As an aside, I've been meaning for ages to do analysis of hiring trends based on the data. If anybody is interested in this, I have the past year of posts available in JSON available at http://hnhiring.me/data/comments-{thread-id}.json. The thread ids for fulltime/freelancer for each month are here: http://hnhiring.me/data/threads.json.
Basically my question is why couldn't you build the whole scraper with just flat files and no database/key-value store at all?
e.g., I'm currently looking to contract somebody for some devops work here at Mozilla Research, and going through the set of listings that have 'devops' in them and actually finding a link to a resume / github account / anything more substantial than a twitter feed is pretty maddening.
https://news.ycombinator.com/item?id=7685170
is coming up on the 1st of June.
I modified the template:
https://gist.github.com/cjbarber/189c84750cd42309201d
Thoughts?
In the link above you can get a copy of the threads/comments data up until May 2014 for the HN Hiring threads (I didn't do anything for the Freelancers ones).
My initial goal was to do the same (do an analysis of the data) so I also have an additional schema created that would allow us to associate the variety of programming languages, tools, locations, and companies together in a structured way.
I have an idea of how I want to create a frontend that users could contribute to (I'd like to gamify it a bit so contributors can get some props on the site for helping get the data analyzed, since that is a pretty manual process), but haven't had the time to work on it.