Show HN: Stowbots – Sort image downloads automatically using deep learning
stowbots.com
stowbots.com
I assume you need a fairly extensive training dataset to train a net, and that would limit your ability to provide a net trained for arbitrary tags.
For that matter, I have multiple empty requests bots (apparently). How do you delete them? I could probably edit stowbotlist.json but I can't imagine that's the intended interface.
Do you specify tags you want to filter by? For example, if I wanted to do (random suggestion) SMT IC package identification, how would I do that?
------
From the sound of it, right now, it sounds like you review the requested categories, and manually/semi-manually generate each specific network?
------
I'm basically in a position where it'd be much more valuable to me to have a thing I can point at a directory and have it tell me what it found then something I have to say "sort into these categories".
In general, there's very few cases where I have a dataset where I want to filter by a specific content, and don't already have that metadata. Content discovery would be much more useful.
I have a mirror of *booru image galleries (basically, anime/manga images, but REALLY well tagged: upwards of 20+ tags per image) that I've been occationally meaning to see if I can turn into a google-deep-dream-like nightmare to make the worst hentai possible (it started as the idea for a terrible idea hackathon[1]).
I have a fair bit of experience writing web-scrapers [2], but no idea what I'm doing with neural net stuff. I keep meaning to spend time to figure out how to do anything, but other projects (and work) keep me distracted.
Anyways, if you have any use for 5496863 images (probably mostly hentai) with 196852748 tags, hit me up.
Or you could just run the scraper [3] yourself, but I hope you have ~5+ TB of free disk space.
1: http://www.stupidhackathon.com/ 2: https://github.com/fake-name 3: https://github.com/fake-name/DanbooruScraper
I've also chatted with Gwern a bit about the project. We had similar interests.
My scraper wound up being very similar to what he implemented, though it was already up and running when he made his first posts (I started my project about 6 months before Gwern).