Show HN: Scrapple – A framework for creating semi-automatic web extractors
github.com
github.com
I built this along with a friend, as part of a project. I would love suggestions and feedback !
* Clues such as how the hell "key-value based configuration files" have anything to do with web scraping (without making people view the source)
* Clues such as why anyone should bother using this (or even if it's worth considering for their tasks)
* Clues as to what form this "framework" takes.
* Clues as to what functionality it provides
* Clues as to what i have to do to make it scrape something...
In short. Please watch this 5 minute talk, contemplate the implications, and update the readme: https://youtu.be/23xzRCoDZf4 You do have a readme, but you may as well not, because it doesn't really tell me much.
It would be awesome if you could create a video tutorial that goes through the process of setting up a project and getting the first tiny bit of data out of it. I'd be more than willing to help you with this if you can show me how to get started. I am sure I can figure this out somehow but time is a valuable asset.
Scrapers are in huge demand as you can see on projects like Scrapy, Pyspider and the like so it would be a pity if this one goes down unnoticed.
I have put a tutorial in the project documentation [http://scrapple.readthedocs.org/en/latest/#experimentation-r...]. I understand how a video tutorial would be more helpful.
It would be great if you could give some advice/suggestions on how to do the video tutorial.
I focused more on writing the documentation for the package, and ended up putting too little time on the repository readme. I see the need for it to be more elaborate. I'll update that.