AudioSet by Google - A massive dataset of manually annotated audio events

Osaurus

•9yr ago

Replies

Best

Osaurus

Hunter

This is remarkable and will do much to advance the state of voice-based computing. Thanks, Uncle Google! AudioSet consists of an expanding ontology of 632 audio event classes and a collection of 2,084,320 human-labeled 10-second sound clips drawn from YouTube videos. The ontology is specified as a hierarchical graph of event categories, covering a wide range of human and animal sounds, musical instruments and genres, and common everyday environmental sounds. By releasing AudioSet, we hope to provide a common, realistic-scale evaluation task for audio event detection, as well as a starting point for a comprehensive vocabulary of sound events.

Report

9yr ago

Thank you so much for sharing this amazing audio dataset! My team is developing a machine learning model that can understand what a video about. This dataset will help us move faster to our goal.

Report

9yr ago

I don't consider this a product, it's a dataset, why is it here?

Report

9yr ago