We should mutualize scraping efforts, creating a sort of Wikipedia of scraped data. I bet a ton of people and cool applications would benefit from it.
The most important part is being able to consistently scrape every day or so for a long time. That isn't easy.