CoolFace
Datasetpublic

nineninesix/expresso-conversational-en-nano-codec-dataset

Expresso Conversational EN Nano-Codec Dataset This dataset is built upon the Expresso conversational dataset and re-encoded using NVIDIA’s NeMo Audio Codec into nano audio tokens. It is designed for fine-tuning multimodal LLMs and speech systems (TTS/ASR) that rely on codec-based audio token representations. Dataset Structure text: transcription of the utterance. speaker: speaker identifier (string). nano_layer_1 … nano_layer_4: tokenized audio… See the full description on the dataset page: https://huggingface.co/datasets/nineninesix/expresso-conversational-en-nano-codec-dataset.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
0likes16downloads
6 commits on main
c3e77e51y ago

Update README.md

Simonlob
266d2b21y ago

Update README.md

Simonlob
439b0ab1y ago

Update README.md

Simonlob
1b31b461y ago

Update README.md

Simonlob
f82a1801y ago

Upload dataset

Simonlob
ed99f011y ago

initial commit

Simonlob