datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
instruction-speech-encodec-v1
Dataset Card for "Instruction Speech"
The largest open-source English speech instruction to text answer dataset
Dataset Overview
This dataset contains nearly 450,000 English speech instruction to text answer samples, using:
A subset of OpenHermes 2.5 with user's prompt length less than 64.
Audio generation using WhisperSpeech.
Tokenized using Encodec.
Usage
from datasets import load_dataset, Audio
# Load Instruction Speech dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Menlo/instruction-speech-encodec-v1.instruction-speech-encodec-v1.5
Dataset Card for "Instruction Speech"
The largest open-source English speech instruction to text answer dataset
Dataset Overview
This dataset contains over 332,000 English speech instruction to text answer samples, using:
A subset of jan-hq/prompt-voice-v1.5.
Audio generation using WhisperSpeech.
Tokenized using Encodec.
Usage
from datasets import load_dataset, Audio
# Load Instruction Speech dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Menlo/instruction-speech-encodec-v1.5.encodec_24khz-opt-125m-pretrained-ft-librispeech_asr
Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr"
More Information needed
cs-envi-dual-encoder-60librispeech_asr_encodecencodec_24khz-librispeech_asr-train.clean.100-features
Dataset Card for "encodec_24khz-librispeech_asr-train.clean.100-features"
More Information needed
MUSDB_stems_encodec_12kbpsencodec_24khz-opt-125m-pretrained-ft-librispeech_asr-train.clean.100-features
Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-train.clean.100-features"
More Information needed
MUSDB_encodec_12kbpsencoder-training-dataencodec_24khz-opt-125m-pretrained-ft-librispeech_asr_dummy-validation-features
Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr_dummy-validation-features"
More Information needed
encodec_24khz-librispeech_asr100h
Dataset Card for "encodec_24khz-librispeech_asr100h"
More Information needed
ljspeech-encodec
Dataset Card for "ljspeech-encodec"
More Information needed
encodec_24khz-b24.0-librispeech_asr-features
Dataset Card for "encodec_24khz-b24.0-librispeech_asr-features"
More Information needed
MUSAN-music_encodec_24k_6bps
Dataset Card for "MUSAN-music_unit"
More Information needed
MUSAN-speech_encodec_24k_6bps
Dataset Card for "MUSAN-speech_unit"
More Information needed
encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-validation.clean-features
Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-validation.clean-features"
More Information needed
encodec_24khz-librispeech_asr-validation.clean-features
Dataset Card for "encodec_24khz-librispeech_asr-validation.clean-features"
More Information needed
encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-test.clean-features
Dataset Card for "encodec_24khz-opt-125m-pretrained-ft-librispeech_asr-test.clean-features"
More Information needed
encodec_24khz-librispeech_asr-test.clean-features
Dataset Card for "encodec_24khz-librispeech_asr-test.clean-features"
More Information needed
test_encode_examplemls_10k_eng_encodecresult_with_finetuned_taggenv2_20epoch_encoder_embeddings
Dataset Card for "result_with_finetuned_taggenv2_20epoch_encoder_embeddings"
More Information needed
result_with_finetuned_taggenv2_9epoch_encoder_embeddings
Dataset Card for "result_with_finetuned_taggenv2_9epoch_encoder_embeddings"
More Information needed
result_with_finetuned_taggenv2_10epoch_encoder_embeddings_decoder_roberta
Dataset Card for "result_with_finetuned_taggenv2_10epoch_encoder_embeddings_decoder_roberta"
More Information needed
default_voices_chunked_encoded
default_voices_chunked_encoded
This is a gated Uzbek encoded derived speech artifact from instinct-org.
This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows.
Language
Primary language: uz (Uzbek)
Intended Use
speech-to-text training and evaluation
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_encoded.zy_chuncked_encoded
zy_chuncked_encoded
This is a gated Uzbek encoded derived speech artifact from instinct-org.
This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows.
Language
Primary language: uz (Uzbek)
Intended Use
speech-to-text training and evaluation
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/zy_chuncked_encoded.
