CoolFace
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01google /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and… See the full description on the dataset page: https://huggingface.co/datasets/google/WaxalNLP.audioautomatic-speech-recognition1M<n<10M286 likes13k downloads24d agoHugging Face02KTH /waxholmThe Waxholm corpus was collected in 1993 - 1994 at the department of Speech, Hearing and Music (TMH), KTH.automatic-speech-recognition0 likes6.1k downloads2y agoHugging Face03fiifinketia /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language… See the full description on the dataset page: https://huggingface.co/datasets/fiifinketia/WaxalNLP.audioautomatic-speech-recognition1M<n<10M1 likes1.9k downloads6mo agoHugging Face04adab-tech /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and… See the full description on the dataset page: https://huggingface.co/datasets/adab-tech/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes1.8k downloads3mo agoHugging Face05anyantudre /waxal-pseudo WAXAL Pseudo-Labels (3-model agreement) PRIVATE working artifact for the Google WAXAL ASR Challenge — not for redistribution. High-confidence pseudo-labels for the WAXAL unlabeled split, produced by a 3-model agreement cascade: a clip is kept only when the fine-tuned champion (w2v-BERT-2.0 CTC) and XLS-R-300m agree (CER ≤ 0.12), and the fine-tuned Omnilingual-ASR-300M independently confirms the champion transcript (CER ≤ 0.22). omni is architecturally diverse (different… See the full description on the dataset page: https://huggingface.co/datasets/anyantudre/waxal-pseudo.audioautomatic-speech-recognition10K<n<100K1 likes848 downloads2mo agoHugging Face06jessteru /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and… See the full description on the dataset page: https://huggingface.co/datasets/jessteru/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes799 downloads3mo agoHugging Face07youvoi /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and… See the full description on the dataset page: https://huggingface.co/datasets/youvoi/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes631 downloads3mo agoHugging Face08claudefitz /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language… See the full description on the dataset page: https://huggingface.co/datasets/claudefitz/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes613 downloads7mo agoHugging Face09novelwolde36 /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language… See the full description on the dataset page: https://huggingface.co/datasets/novelwolde36/WaxalNLP.audioautomatic-speech-recognition1M<n<10M1 likes413 downloads7mo agoHugging Face10BolajiJsAI /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language… See the full description on the dataset page: https://huggingface.co/datasets/BolajiJsAI/WaxalNLP.audioautomatic-speech-recognition1M<n<10M1 likes273 downloads7mo agoHugging Face11anyantudre /waxal-linsna WAXAL Phase-2 Lingala / Shona training corpus (derived) This corpus was built for the Google WAXAL ASR Challenge on Zindi. This repository is the exact training corpus behind our Google WAXAL ASR Challenge (Phase 2) submission: the TSV manifests plus the derived 16 kHz mono FLAC audio that our training configs read. It exists so that the whole recipe can be rebuilt and audited from one place. The manifests below, together with the google/WaxalNLP train and validation splits read… See the full description on the dataset page: https://huggingface.co/datasets/anyantudre/waxal-linsna.audioautomatic-speech-recognition1K<n<10K0 likes273 downloads2mo agoHugging Face12Phsntom /WaxalNLP Waxal Datasets Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language technology for these underserved languages, and to serve as a repository for digital preservation. The Waxal datasets are collections acquired through partnerships with Makerere… See the full description on the dataset page: https://huggingface.co/datasets/Phsntom/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes264 downloads8mo agoHugging Face13matrixdose /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and… See the full description on the dataset page: https://huggingface.co/datasets/matrixdose/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes259 downloads4mo agoHugging Face140xzanuee /WaxalNLP Waxal Datasets Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language technology for these underserved languages, and to serve as a repository for digital preservation. The Waxal datasets are collections acquired through partnerships with… See the full description on the dataset page: https://huggingface.co/datasets/0xzanuee/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes255 downloads8mo agoHugging Face15manassehzw /sna-waxal-annotated-unlabeled Shona WAXAL annotated-unlabeled checkpoint This is a self-contained operational checkpoint for pseudo-labeling Shona ASR data. It contains 90,253 conservatively segmented FLAC clips (441.585 hours), but intentionally contains no transcripts. Fields transcription is intentionally empty. speaker_id is an approximate source-blind EOM cluster or unknown; speaker_clip_count is zero for unknown assignments. gender is always unknown; available classifiers were not… See the full description on the dataset page: https://huggingface.co/datasets/manassehzw/sna-waxal-annotated-unlabeled.audioautomatic-speech-recognition10K<n<100K0 likes242 downloads2mo agoHugging Face16Ephraimmm /WaxalNLPr Waxal NLP Datasets Overview This repository hosts a large multilingual speech corpus for 27 African languages, split into two task collections: Automatic Speech Recognition (ASR) — natural speech paired with human transcriptions — and Text-to-Speech (TTS) — single-speaker studio recordings paired with the scripted text that was read aloud. The dataset card and data itself indicate the underlying recordings were gathered through partnerships with Makerere… See the full description on the dataset page: https://huggingface.co/datasets/Ephraimmm/WaxalNLPr.audioautomatic-speech-recognition1M<n<10M0 likes222 downloads3mo agoHugging Face17manassehzw /shona-waxal-pseudo-labeled Shona WAXAL pseudo-labelled speech This release contains 90,253 Shona speech clips, totalling 441.585 hours. Each clip keeps its original FLAC audio and a Sunbird Whisper pseudo-transcript. These are model outputs, not human reference transcriptions. What this release contains The source is the unlabeled Shona ASR split from WAXAL NLP, preserved in the operational checkpoint manassehzw/sna-waxal-annotated-unlabeled. The source checkpoint has no transcripts. This… See the full description on the dataset page: https://huggingface.co/datasets/manassehzw/shona-waxal-pseudo-labeled.audioautomatic-speech-recognition10K<n<100K3 likes190 downloads25d agoHugging Face18vnahata /waxal-audio-text-retrieval WAXAL speech–text retrieval (MTEB) Multilingual speech↔text retrieval over 16 Sub-Saharan African languages, derived from WAXAL (Google and partners). Most of these languages have no presence in mteb's existing multilingual audio tasks, which skew European and South/East Asian. Prepared as WaxalA2TRetrieval and WaxalT2ARetrieval. Contents One config per language, each with id, audio, text, speaker_id, gender, language. 1,722 utterances total. code language… See the full description on the dataset page: https://huggingface.co/datasets/vnahata/waxal-audio-text-retrieval.audioautomatic-speech-recognition1K<n<10K0 likes133 downloads27d agoHugging Face19bdallhrajh371 /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language… See the full description on the dataset page: https://huggingface.co/datasets/bdallhrajh371/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes103 downloads6mo agoHugging Face20ngoloan /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and language… See the full description on the dataset page: https://huggingface.co/datasets/ngoloan/WaxalNLP.audioautomatic-speech-recognition1M<n<10M0 likes80 downloads7mo agoHugging Face21djelia /bambara-tts-waxal bambara-tts-waxal Bambara studio speech from the WAXAL corpus — 1,926 recordings, 16 hours, 8 speakers, 44.1 kHz mono. Load from datasets import load_dataset ds = load_dataset("djelia/bambara-tts-waxal", "google_waxal", split="train") Splits: train, validation, test. Fields Field Description audio 44.1 kHz mono text Transcript speaker_id Speaker identifier (8 distinct) gender Speaker gender locale Locale code id Record… See the full description on the dataset page: https://huggingface.co/datasets/djelia/bambara-tts-waxal.audiotext-to-speech1K<n<10K0 likes69 downloads2mo agoHugging Face22isaacoluwafemiog /WaxalNLP Waxal Datasets The WAXAL dataset is a large-scale multilingual speech corpus for African languages, introduced in the paper WAXAL: A Large-Scale Multilingual African Language Speech Corpus. Dataset Description The Waxal project provides datasets for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) for African languages. The goal of this dataset's creation and release is to facilitate research that improves the accuracy and fluency of speech and… See the full description on the dataset page: https://huggingface.co/datasets/isaacoluwafemiog/WaxalNLP.audioautomatic-speech-recognition0 likes53 downloads2mo agoHugging Face23teckedd /serendepify-gsl-asr-ak-waxal-gnlp-whisper-small-replay-fullft-v0.1 serendepify-gsl-asr-ak-waxal-gnlp-whisper-small-replay-fullft-v0.1 This is the first pushed Ghanaian Speech Lab ASR pipeline artifact. It is a review artifact for the v0.1 Akan ASR pass, not a trained model checkpoint. Expected future model repo: teckedd/serendepify-gsl-asr-ak-waxal-gnlp-whisper-small-replay-fullft-v0.1 What This Artifact Contains data/manifest.jsonl: harmonized Waxal + GhanaNLP manifest references. reports/sanitize.json: sanitization report and… See the full description on the dataset page: https://huggingface.co/datasets/teckedd/serendepify-gsl-asr-ak-waxal-gnlp-whisper-small-replay-fullft-v0.1.automatic-speech-recognition0 likes24 downloads3mo agoHugging Face24b1n1yam /waxal-orm-tts-merged Waxal Oromo TTS Merged This dataset merges the human-labeled Oromo ASR split from google/WaxalNLP with the autolabeled Oromo split from israel/waxal-autolabled. For TTS use, the leading [ORM] language tag has been removed from autolabeled transcriptions. Rows include both text and transcription with the same cleaned value. Target repo: b1n1yam/waxal-orm-tts-merged audiotext-to-speech100K<n<1M1 likes14 downloads4mo agoHugging Face25manassehzw /sna-waxal-unlabeled-tar WAXAL Shona unlabeled operational TAR dataset WebDataset packaging of the Shona unlabeled split from google/WaxalNLP. config: sna_asr split: unlabeled pinned upstream revision: e0a62aaebc61bd5bb8cac17a08d1b42c65551dd2 samples: 85,384 audio: source-encoded bytes preserved without transcoding The pinned upstream Parquet files remain the recoverable source. This repo is an operational derivative optimized for sequential streaming. Each example is a matching audio and .json pair.… See the full description on the dataset page: https://huggingface.co/datasets/manassehzw/sna-waxal-unlabeled-tar.automatic-speech-recognition0 likes9 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.