CoolFace
20 results

audioset

agkphysics /AudioSet Dataset Card for AudioSet Dataset Summary AudioSet is a dataset of 10-second clips from YouTube, annotated into one or more sound categories, following the AudioSet ontology. Supported Tasks and Leaderboards audio-classification: Classify audio clips into categories. The leaderboard is available here Languages The class labels in the dataset are in English. Dataset Structure Data Instances Example… See the full description on the dataset page: https://huggingface.co/datasets/agkphysics/AudioSet.audioaudio-classification1M<n<10M109 likes59k downloads11mo agoHugging Faceyangwang825 /audioset AudioSet AudioSet[1] consists of an expanding ontology of 527 audio event classes and a collection of 2M human-labelled 10-second sound clips drawn from YouTube. Some clips are missing on YouTube, so the number of files downloaded is different from time to time. This repository contains 20550 / 22160 of the balanced train set, 1913637 / 2041789 of the unbalanced train set (separated into 41 parts), and 18887 / 20371 of the evaluation set. The pre-process script can be found at… See the full description on the dataset page: https://huggingface.co/datasets/yangwang825/audioset.textaudio-classification1M<n<10M5 likes11k downloads3y agoHugging Faceenyoukai /AudioSet-Strongaudio10K<n<100K4 likes2.5k downloads10mo agoHugging FaceMYJOKERML /audioset_cla_label_des_naiveaudio100K<n<1M0 likes2.2k downloads1y agoHugging FaceCLAPv2 /audioset_strongaudio100K<n<1M1 likes2.1k downloads2y agoHugging Faceconfit /audioset-16khz-wdsgated AudioSet AudioSet[1] is a large-scale dataset comprising approximately 2 million 10-second YouTube audio clips, categorised into 527 sound classes. We have pre-processed all audio files to a 16 kHz sampling rate and stored them in the WebDataset format for efficient large-scale training and retrieval. Download We recommend using the following commands to download the confit/audioset-16khz-wds dataset from HuggingFace. The dataset is available in two versions:… See the full description on the dataset page: https://huggingface.co/datasets/confit/audioset-16khz-wds.audioaudio-classification1M<n<10M8 likes2k downloads2y agoHugging Face