Announcing AudioSet: A Dataset for Audio Event Research
research.googleblog.com
research.googleblog.com
I also like that they even leveraged stuff like "x for y hours" style videos on youtube. (Eg: Dog barking to 10 hours).
I remember Andrew Ng mention something about being able synthesize even more trainable data by using combination of data as noise. (Eg: Women talking + dog barking (as introduced noise)).
Good job google, keep it up. I wonder when I will see other major players publish large data sets? Maybe Microsoft is next?
[0] https://research.google.com/pubs/pub41880.html [1] https://www.microsoft.com/en-us/research/project/mslr/ [2] https://www.microsoft.com/en-us/research/project/microsoft-3...
Looks like the... tag names? and example urls have been released, but the videos and sound are under their respective licenses -- ie, mostly the standard Youtube license.
This is neat. Can a ML model developed on this dataset be used for commercial purposes? I guess, at minimum, the paper and tag list are provided as help for those corporations that would wish to build/use a private dataset for similar purposes?