Show HN: Deep learning visual search and data analytics
github.com
github.com
1. https://projects.csail.mit.edu/soundnet/
2. https://github.com/aalireza/SimpleAudioIndexer (using PocketSphinx, NOT Watson)
- Find similar looking frames?
Question:
- Does it perform object detection on the frame? Similar to the video demo on Clarifai - https://clarifai.com/demo ?
To answer second question we also detect objects (VOC, YOLO 9000, Faces etc.), detected objects are also indexed and retrieved when performing visual search. Further you can perform clustering on these set of "indexing" vectors for things such as fast retrieval and quick labeling/annotations. We use Flickr LOPQ to implement ANN but like all other things you can use custom algorithm. I am working on adding indexing over any set of annotations/detections/frames.
You can find more information about the design goals and vision behind the project in presentation at https://deepvideoanalytics.com/
To summarize yes we provide two indexers out of the box a general purpose inception v3 and a facenet. We plan to add more indexers soon, e.g. trained on Open Images or other domain specific dataset.
Sharing visual data opens up a whole new set of opportunities for both businesses and researchers.