CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01hf-internal-testing /librispeech_asr_dummyaudion<1K11 likes106k downloads2y agoHugging Face02openslr /librispeech_asr Dataset Card for librispeech_asr Dataset Summary LibriSpeech is a corpus of approximately 1000 hours of 16kHz read English speech, prepared by Vassil Panayotov with the assistance of Daniel Povey. The data is derived from read audiobooks from the LibriVox project, and has been carefully segmented and aligned. Supported Tasks and Leaderboards automatic-speech-recognition, audio-speaker-identification: The dataset can be used to train a model for Automatic… See the full description on the dataset page: https://huggingface.co/datasets/openslr/librispeech_asr.audioautomatic-speech-recognition100K<n<1M245 likes54k downloads1y agoHugging Face03facebook /multilingual_librispeech Dataset Card for MultiLingual LibriSpeech Dataset Summary This is a streamable version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish.… See the full description on the dataset page: https://huggingface.co/datasets/facebook/multilingual_librispeech.audioautomatic-speech-recognition1M<n<10M190 likes36k downloads2y agoHugging Face04distil-whisper /librispeech_long Dataset Card for "librispeech_long" More Information needed audion<1K4 likes11k downloads3y agoHugging Face05WillHeld /test_librispeech_parquetaudion<1K0 likes5.6k downloads3y agoHugging Face06distil-whisper /librispeech_asr-noise Dataset Card for "librispeech_asr-noise" More Information needed audio100K<n<1M2 likes4.5k downloads3y agoHugging Face07zihan-audio /librispeech_asr Dataset Card for librispeech_asr Dataset Summary LibriSpeech is a corpus of approximately 1000 hours of 16kHz read English speech, prepared by Vassil Panayotov with the assistance of Daniel Povey. The data is derived from read audiobooks from the LibriVox project, and has been carefully segmented and aligned. Supported Tasks and Leaderboards automatic-speech-recognition, audio-speaker-identification: The dataset can be used to train a model for… See the full description on the dataset page: https://huggingface.co/datasets/zihan-audio/librispeech_asr.audioautomatic-speech-recognition100K<n<1M0 likes4.3k downloads26d agoHugging Face08hf-internal-testing /librispeech_asr_demoaudion<1K3 likes3.2k downloads1y agoHugging Face09gilkeyio /librispeech-alignments Dataset Card for Librispeech Alignments Librispeech with alignments generated by the Montreal Forced Aligner. The original alignments in TextGrid format can be found here Dataset Details Dataset Description Librispeech is a corpus of read English speech, designed for training and evaluating automatic speech recognition (ASR) systems. The dataset contains 1000 hours of 16kHz read English speech derived from audiobooks. The Montreal Forced Aligner (MFA) was used… See the full description on the dataset page: https://huggingface.co/datasets/gilkeyio/librispeech-alignments.audioautomatic-speech-recognition100K<n<1M21 likes2.4k downloads3y agoHugging Face10lmms-lab-audio /librispeechaudio10K<n<100K5 likes2.2k downloads2y agoHugging Face11Splend1dchan /librispeech_asr_arrowaudio10K<n<100K0 likes1.9k downloads3y agoHugging Face12fixie-ai /librispeech_asraudio100K<n<1M5 likes1.8k downloads2y agoHugging Face13sanchit-gandhi /librispeech-data Dataset Card for "librispeech-data" More Information needed audio100K<n<1M2 likes1.1k downloads3y agoHugging Face14Codec-SUPERB /librispeech_synth Dataset Card for "librispeech_synth" More Information needed audio1M<n<10M1 likes932 downloads3y agoHugging Face15istupakov /russian_librispeech Russian LibriSpeech (RuLS) Identifier: SLR96 from openslr.org Summary: This dataset is based on LibriVox audiobooks Category: Speech License: The dataset is Public Domain in the USA. About this resource: Russian LibriSpeech (RuLS) dataset is based on LibriVox's public domain audio books (see BOOKS.TXT for the list of included books) and contains about 98 hours of audio data. audioautomatic-speech-recognition10K<n<100K6 likes858 downloads1y agoHugging Face16distil-whisper /librispeech_asr-prompted Dataset Card for "librispeech_asr-prompted" More Information needed audio100K<n<1M0 likes610 downloads3y agoHugging Face17arkubeth /librispeechaudio1K<n<10K0 likes585 downloads3y agoHugging Face18cmu-mlsp /hubert_layer9-librispeech-asr100h Dataset Card for "hubert_layer9-librispeech-asr100h" More Information needed audio10K<n<100K0 likes547 downloads3y agoHugging Face19flozi00 /multilingual-librispeech-german-labeledaudio100K<n<1M1 likes527 downloads2y agoHugging Face20davidggphy /librispeech-arpabet-processed LibriSpeech ARPAbet Processed Dataset Pre-processed dataset for training ARPAbet phoneme recognition models using CTC loss. Dataset Description This dataset is derived from LibriSpeech (train-clean-100 split) with the following preprocessing: Audio: Resampled to 16kHz, normalized using Wav2Vec2 feature extractor Labels: Text transcriptions converted to ARPAbet phoneme sequences using CMU Pronouncing Dictionary Filtering: Samples with out-of-vocabulary words (not in CMU… See the full description on the dataset page: https://huggingface.co/datasets/davidggphy/librispeech-arpabet-processed.audioautomatic-speech-recognition10K<n<100K0 likes516 downloads8mo agoHugging Face21TwinkStart /librispeech This dataset only contains test data, which is integrated into UltraEval-Audio(https://github.com/OpenBMB/UltraEval-Audio) framework. python audio_evals/main.py --dataset librispeech-test-clean --model gpt4o_audio python audio_evals/main.py --dataset librispeech-dev-clean --model gpt4o_audio python audio_evals/main.py --dataset librispeech-test-other --model gpt4o_audio python audio_evals/main.py --dataset librispeech-dev-other --model gpt4o_audio 🚀超凡体验,尽在UltraEval-Audio🚀… See the full description on the dataset page: https://huggingface.co/datasets/TwinkStart/librispeech.audio10K<n<100K0 likes482 downloads2y agoHugging Face220x3 /librispeech_asr Dataset Card for librispeech_asr Dataset Summary LibriSpeech is a corpus of approximately 1000 hours of 16kHz read English speech, prepared by Vassil Panayotov with the assistance of Daniel Povey. The data is derived from read audiobooks from the LibriVox project, and has been carefully segmented and aligned. Supported Tasks and Leaderboards automatic-speech-recognition, audio-speaker-identification: The dataset can be used to train a model for Automatic… See the full description on the dataset page: https://huggingface.co/datasets/0x3/librispeech_asr.audioautomatic-speech-recognition100K<n<1M0 likes481 downloads4mo agoHugging Face23CodecSR /librispeech_asr_test_48k_synthaudio100K<n<1M0 likes478 downloads3y agoHugging Face24dys-asr /librispeech-sr16000audio100K<n<1M0 likes477 downloads7mo agoHugging Face25sovitrath /librispeech_asr Dataset Card for librispeech_asr Dataset Summary LibriSpeech is a corpus of approximately 1000 hours of 16kHz read English speech, prepared by Vassil Panayotov with the assistance of Daniel Povey. The data is derived from read audiobooks from the LibriVox project, and has been carefully segmented and aligned. Supported Tasks and Leaderboards automatic-speech-recognition, audio-speaker-identification: The dataset can be used to train a model for Automatic… See the full description on the dataset page: https://huggingface.co/datasets/sovitrath/librispeech_asr.audioautomatic-speech-recognition100K<n<1M0 likes461 downloads5mo agoHugging Face26DTU54DL /librispeech-augmentated-train-prepared Dataset Card for "librispeech-augmentated-train-prepared" More Information needed audio1K<n<10K0 likes460 downloads4y agoHugging Face27WillHeld /librispeech_parquetaudio100K<n<1M0 likes453 downloads3y agoHugging Face28cmu-mlsp /wavlm-large_layer21-librispeech-asr100h Dataset Card for "wavlm-large_layer21-librispeech-asr100h" More Information needed audio10K<n<100K0 likes451 downloads3y agoHugging Face29CodecSR /librispeech_asr_test_synthaudio100K<n<1M0 likes428 downloads3y agoHugging Face30mythicinfinity /librispeech-pc-44khz-opus LibriSpeech-PC 44kHz Opus Summary This dataset is a high-quality audio replacement variant of Librispeech PC. It preserves the row identity and text fields while replacing audio content from the source audio with the highest available quality (usually mp3 128kpbs) which is then encoded as Opus (64 kbps). Sampling rate is increased from 16khz up to 48khz depending the on source audio. LibriSpeech-PC is a merge of openslr/librispeech_asr audio metadata with SLR145… See the full description on the dataset page: https://huggingface.co/datasets/mythicinfinity/librispeech-pc-44khz-opus.audioautomatic-speech-recognition100K<n<1M5 likes413 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.