audioset
Datasets
All datasets matching “audioset”AudioSet
Dataset Card for AudioSet
Dataset Summary
AudioSet is a dataset of 10-second clips from YouTube, annotated into one or more sound categories, following the AudioSet ontology.
Supported Tasks and Leaderboards
audio-classification: Classify audio clips into categories. The leaderboard is available here
Languages
The class labels in the dataset are in English.
Dataset Structure
Data Instances
Example… See the full description on the dataset page: https://huggingface.co/datasets/agkphysics/AudioSet.audioset
AudioSet
AudioSet[1] consists of an expanding ontology of 527 audio event classes and a collection of 2M human-labelled 10-second sound clips drawn from YouTube.
Some clips are missing on YouTube, so the number of files downloaded is different from time to time.
This repository contains 20550 / 22160 of the balanced train set, 1913637 / 2041789 of the unbalanced train set (separated into 41 parts), and 18887 / 20371 of the evaluation set.
The pre-process script can be found at… See the full description on the dataset page: https://huggingface.co/datasets/yangwang825/audioset.AudioSet-Strongaudioset_cla_label_des_naiveaudioset_strongaudioset-16khz-wds
AudioSet
AudioSet[1] is a large-scale dataset comprising approximately 2 million 10-second YouTube audio clips, categorised into 527 sound classes.
We have pre-processed all audio files to a 16 kHz sampling rate and stored them in the WebDataset format for efficient large-scale training and retrieval.
Download
We recommend using the following commands to download the confit/audioset-16khz-wds dataset from HuggingFace.
The dataset is available in two versions:… See the full description on the dataset page: https://huggingface.co/datasets/confit/audioset-16khz-wds.
