CoolFace
3 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Speech-data /Cebuano-Speech-Dataset 🎧 Cebuano Speech Dataset The Cebuano Speech Dataset is a high-quality speech audio dataset designed to deliver structured and diverse audio data for AI-powered voice applications. It includes 108 hours of audio data distributed across 807 files, provided in MP3 and WAV formats, with a total size of 135 MB. This well-organized audio dataset ensures balanced voice data, with 49% female and 51% male speakers, and a broad age range from 18 to 50+ years. The dataset language is Cebuano… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Cebuano-Speech-Dataset.audioautomatic-speech-recognitionn<1K1 likes28 downloads6mo agoHugging Face02jojo-ai-mst /Roleplay-Cebuano RolePlay-Cebuano Roleplay-Cebuano Dataset is a dataset for roleplaying in the Amharic language for the Large Language Model. The base dataset is the GPTeacher role play dataset by teknium 1, which can be found under this link, released under MIT License. The dataset is then translated into respective languages. The translation process is powered by Google Translate, using cloud translation API. For more information and other language datasets for roleplay, see this github repo. For… See the full description on the dataset page: https://huggingface.co/datasets/jojo-ai-mst/Roleplay-Cebuano.texttext-generation1K<n<10K0 likes26 downloads2y agoHugging Face03roncc13 /CebuanoAnnotatedFakeandLegitNewstext1K<n<10K0 likes17 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.