CoolFace
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01maikezu /f-actor-behavior-sd-nanocodec F-Actor Nano-Codec Dataset This repository contains the data accompanying the paper F-Actor: Controllable Conversational Behaviour in Full-Duplex Models. The data consists of the Behavior-SD dataset, encoded using nvidia/nemo-nano-codec-22khz-0.6kbps-12.5fps, and augmented with a different narrative. About our work: Spoken conversational systems require more than accurate speech generation to have human-like conversations: to feel natural and engaging, they must produce… See the full description on the dataset page: https://huggingface.co/datasets/maikezu/f-actor-behavior-sd-nanocodec.tabular100K<n<1M1 likes377 downloads8mo agoHugging Face02Rcarvalo /kanitts2-fr-nanocodectabular100K<n<1M0 likes215 downloads6mo agoHugging Face03nineninesix /emolia_filtered_nano_codec_21_dataset Emolia · Filtered · NanoCodec (FSQ) Tokens A cleaned, pre-tokenized version of laion/Emolia prepared for text-to-speech (TTS) training. The pipeline is two steps: Quality filtering with the open-source audio_filter tool — this removes the dirtiest recordings (noise, clipping, band-limiting, robotic artifacts, overlapping speakers), which matters a lot for TTS quality. Discrete audio tokenization with NVIDIA nvidia/nemo-nano-codec-22khz-1.89kbps-21.5fps (an FSQ neural audio… See the full description on the dataset page: https://huggingface.co/datasets/nineninesix/emolia_filtered_nano_codec_21_dataset.texttext-to-speech1M<n<10M1 likes206 downloads3mo agoHugging Face04nineninesix /elise-en-nano-codec-dataset Elise EN Nano-Codec Dataset This dataset is built upon the Elise dataset and re-encoded using NVIDIA’s NeMo Audio Codec into nano audio tokens. It is designed for fine-tuning multimodal LLMs and speech systems (TTS/ASR) that rely on codec-based audio token representations. Dataset Structure text: transcription of the utterance. speaker: speaker identifier (string). nano_layer_1 … nano_layer_4: tokenized audio representations from the NVIDIA NeMo Nano Codec… See the full description on the dataset page: https://huggingface.co/datasets/nineninesix/elise-en-nano-codec-dataset.textfeature-extraction1K<n<10K0 likes20 downloads1y agoHugging Face05ankitdhiman /indicvoice-hi-nanocodec-tokenstext100K<n<1M0 likes19 downloads1y agoHugging Face06matthewspring /puck-gemini-flash-en-nano-codec-datasettext10K<n<100K0 likes19 downloads10mo agoHugging Face07ankitdhiman /indicvoices-nanocodec-tokensThis dataset contains 200K samples of text and corresponding audio tokenized using nvidia/nemo-nano-codec-22khz-1.89kbps-21.5fps First 100K samples are in Hindi, sampled from the hindi split of indicvoices dataset, and next 100K samples are in english with Indian accent, sampled from skbose/indian-english-nptel-v0 dataset. The dataset can be useful for training TTS models. text100K<n<1M1 likes18 downloads1y agoHugging Face08jsbeaudry /nvidia-nemo-nano-codec-bibletextn<1K0 likes18 downloads1y agoHugging Face09Praha-Labs /rasa-malayalam-nano-codectext10K<n<100K0 likes17 downloads1y agoHugging Face10Praha-Labs /exp-nano-codectext1K<n<10K0 likes15 downloads11mo agoHugging Face11mahwizzzz /ur_nano_codec Urdu Nano Codec TTS Dataset This dataset contains Urdu sentences along with their Nano Codec tokenized representations (nano_layer_1 to nano_layer_4) generated using NVIDIA NeMo Nano Codec model (22kHz, 0.6 kbps, 12.5 fps). It can be used for: TTS model training Audio reconstruction from tokens Low-resource Urdu speech research Dataset Structure text: Original Urdu sentences nano_layer_1 … nano_layer_4: Tokenized representations (quantized codebooks) encoded_len:… See the full description on the dataset page: https://huggingface.co/datasets/mahwizzzz/ur_nano_codec.text10K<n<100K0 likes14 downloads11mo agoHugging Face12nineninesix /jinsaryko-tifa-en-nano-codec-dataset Tifa EN Nano-Codec Dataset This dataset is built upon the Tifa dataset and re-encoded using NVIDIA’s NeMo Audio Codec into nano audio tokens. It is designed for fine-tuning multimodal LLMs and speech systems (TTS/ASR) that rely on codec-based audio token representations. Dataset Structure text: transcription of the utterance. speaker: speaker identifier (string). nano_layer_1 … nano_layer_4: tokenized audio representations from the NVIDIA NeMo Nano Codec… See the full description on the dataset page: https://huggingface.co/datasets/nineninesix/jinsaryko-tifa-en-nano-codec-dataset.textfeature-extraction1K<n<10K0 likes13 downloads1y agoHugging Face13nineninesix /optimus-prime-en-nano-codec-datasettext1K<n<10K0 likes13 downloads1y agoHugging Face14nineninesix /expresso-conversational-en-nano-codec-dataset Expresso Conversational EN Nano-Codec Dataset This dataset is built upon the Expresso conversational dataset and re-encoded using NVIDIA’s NeMo Audio Codec into nano audio tokens. It is designed for fine-tuning multimodal LLMs and speech systems (TTS/ASR) that rely on codec-based audio token representations. Dataset Structure text: transcription of the utterance. speaker: speaker identifier (string). nano_layer_1 … nano_layer_4: tokenized audio representations… See the full description on the dataset page: https://huggingface.co/datasets/nineninesix/expresso-conversational-en-nano-codec-dataset.textfeature-extraction10K<n<100K0 likes13 downloads1y agoHugging Face15Praha-Labs /IndexTTS-nano-codectext10K<n<100K0 likes13 downloads1y agoHugging Face16Praha-Labs /imasc_slr_Malayalam-nano-codectext10K<n<100K0 likes13 downloads1y agoHugging Face17nineninesix /kore-gemini-flash-en-nano-codec-datasettext10K<n<100K0 likes9 downloads1y agoHugging Face18Mo-alaa /magic-data-nano-codec-datasettextn<1K0 likes9 downloads11mo agoHugging Face19Anilosan15 /Nano_Codec_Nisan_Kumrutext1K<n<10K2 likes9 downloads11mo agoHugging Face20Praha-Labs /IndicVoices-r-ML-nano-codectext10K<n<100K0 likes8 downloads1y agoHugging Face21Praha-Labs /SPRINGLab-IndicTTS_Malayalam-nano-codectext10K<n<100K0 likes7 downloads1y agoHugging Face22nineninesix /puck-gemini-flash-en-nano-codec-datasettext10K<n<100K0 likes6 downloads1y agoHugging Face23nineninesix /kuroyukihime-ja-nano-codec-datasettext1K<n<10K0 likes6 downloads10mo agoHugging Face24TalhaAhmed /urdu-tts-nano-codectext10K<n<100K1 likes4 downloads10mo agoHugging Face25chumdz97 /thien-tts-nanocodectext10K<n<100K0 likes4 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.