CoolFace
27 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Menlo /instruction-speech-encodec-v1 Dataset Card for "Instruction Speech" The largest open-source English speech instruction to text answer dataset Dataset Overview This dataset contains nearly 450,000 English speech instruction to text answer samples, using: A subset of OpenHermes 2.5 with user's prompt length less than 64. Audio generation using WhisperSpeech. Tokenized using Encodec. Usage from datasets import load_dataset, Audio # Load Instruction Speech dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Menlo/instruction-speech-encodec-v1.audio100K<n<1M18 likes1.4k downloads2y agoHugging Face02Menlo /instruction-speech-encodec-v1.5 Dataset Card for "Instruction Speech" The largest open-source English speech instruction to text answer dataset Dataset Overview This dataset contains over 332,000 English speech instruction to text answer samples, using: A subset of jan-hq/prompt-voice-v1.5. Audio generation using WhisperSpeech. Tokenized using Encodec. Usage from datasets import load_dataset, Audio # Load Instruction Speech dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Menlo/instruction-speech-encodec-v1.5.audio100K<n<1M7 likes350 downloads2y agoHugging Face03cmu-mlsp /encodec_24khz-opt-125m-pretrained-ft-librispeech_asr Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr" More Information needed audio10K<n<100K1 likes288 downloads3y agoHugging Face04MrJackTung /cs-envi-dual-encoder-60audion<1K0 likes237 downloads4mo agoHugging Face05theodorr /librispeech_asr_encodecaudio100K<n<1M0 likes234 downloads2y agoHugging Face06cmu-mlsp /encodec_24khz-librispeech_asr-train.clean.100-features Dataset Card for "encodec_24khz-librispeech_asr-train.clean.100-features" More Information needed audio10K<n<100K0 likes173 downloads3y agoHugging Face07danjacobellis /MUSDB_stems_encodec_12kbpsaudion<1K0 likes163 downloads2y agoHugging Face08cmu-mlsp /encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-train.clean.100-features Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-train.clean.100-features" More Information needed audio10K<n<100K0 likes143 downloads3y agoHugging Face09danjacobellis /MUSDB_encodec_12kbpsaudion<1K0 likes138 downloads2y agoHugging Face10alperiox /encoder-training-dataaudio1K<n<10K0 likes115 downloads3mo agoHugging Face11cmu-mlsp /encodec_24khz-opt-125m-pretrained-ft-librispeech_asr_dummy-validation-features Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr_dummy-validation-features" More Information needed audion<1K1 likes65 downloads3y agoHugging Face12roshansh /encodec_24khz-librispeech_asr100h Dataset Card for "encodec_24khz-librispeech_asr100h" More Information needed audio10K<n<100K0 likes62 downloads3y agoHugging Face13naiveen /ljspeech-encodec Dataset Card for "ljspeech-encodec" More Information needed audio10K<n<100K1 likes46 downloads1y agoHugging Face14cmu-mlsp /encodec_24khz-b24.0-librispeech_asr-features Dataset Card for "encodec_24khz-b24.0-librispeech_asr-features" More Information needed audio1K<n<10K0 likes27 downloads3y agoHugging Face15JST-SUPERB /MUSAN-music_encodec_24k_6bps Dataset Card for "MUSAN-music_unit" More Information needed audio1K<n<10K0 likes27 downloads2y agoHugging Face16JST-SUPERB /MUSAN-speech_encodec_24k_6bps Dataset Card for "MUSAN-speech_unit" More Information needed audio1K<n<10K0 likes25 downloads2y agoHugging Face17cmu-mlsp /encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-validation.clean-features Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-validation.clean-features" More Information needed audio1K<n<10K0 likes21 downloads3y agoHugging Face18cmu-mlsp /encodec_24khz-librispeech_asr-validation.clean-features Dataset Card for "encodec_24khz-librispeech_asr-validation.clean-features" More Information needed audio1K<n<10K0 likes16 downloads3y agoHugging Face19cmu-mlsp /encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-test.clean-features Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-test.clean-features" More Information needed audio1K<n<10K0 likes14 downloads3y agoHugging Face20cmu-mlsp /encodec_24khz-librispeech_asr-test.clean-features Dataset Card for "encodec_24khz-librispeech_asr-test.clean-features" More Information needed audio1K<n<10K0 likes13 downloads3y agoHugging Face21polinaeterna /test_encode_exampleaudion<1K0 likes11 downloads5y agoHugging Face22theodorr /mls_10k_eng_encodecaudio1M<n<10M0 likes7 downloads2y agoHugging Face23linhqyy /result_with_finetuned_taggenv2_20epoch_encoder_embeddings Dataset Card for "result_with_finetuned_taggenv2_20epoch_encoder_embeddings" More Information needed audio1K<n<10K0 likes6 downloads3y agoHugging Face24linhqyy /result_with_finetuned_taggenv2_9epoch_encoder_embeddings Dataset Card for "result_with_finetuned_taggenv2_9epoch_encoder_embeddings" More Information needed audio1K<n<10K0 likes5 downloads3y agoHugging Face25linhqyy /result_with_finetuned_taggenv2_10epoch_encoder_embeddings_decoder_roberta Dataset Card for "result_with_finetuned_taggenv2_10epoch_encoder_embeddings_decoder_roberta" More Information needed audio1K<n<10K0 likes5 downloads3y agoHugging Face26instinct-org /default_voices_chunked_encodedgated default_voices_chunked_encoded This is a gated Uzbek encoded derived speech artifact from instinct-org. This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows. Language Primary language: uz (Uzbek) Intended Use speech-to-text training and evaluation Internal dataset curation, quality checks, and model evaluation Research or commercial use only after access approval and license review… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_encoded.audioautomatic-speech-recognition0 likes5 downloads4mo agoHugging Face27instinct-org /zy_chuncked_encodedgated zy_chuncked_encoded This is a gated Uzbek encoded derived speech artifact from instinct-org. This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows. Language Primary language: uz (Uzbek) Intended Use speech-to-text training and evaluation Internal dataset curation, quality checks, and model evaluation Research or commercial use only after access approval and license review Data… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/zy_chuncked_encoded.textautomatic-speech-recognition100K<n<1M0 likes1 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.