djelia/bambara-audio-b
bambara-audio-b Bambara speech derived from scripture recordings, published in four processing stages: raw segments, a length-filtered version, a speaker-diarized long-form cut, and a CTC forced-alignment cut. 30.55 GB of Parquet. Access is gated with manual approval — request it on the dataset page and authenticate (hf auth login or HF_TOKEN) before loading. Load from datasets import load_dataset ds = load_dataset("djelia/bambara-audio-b", "short-filtered"… See the full description on the dataset page: https://huggingface.co/datasets/djelia/bambara-audio-b.
16
No card is published for this repository, or it could not be fetched from Hugging Face right now.
