buaiir
Datasets
All datasets matching “buaiir”buaiir_spectrabuaiir_voice_jap
BUAIIR Japadhola Voice (BUAIR/buaiir_voice_jap)
Structured student read-speech in Japadhola (Adhola, ISO 639-3: adh) from Busitema University
Phase-2 batches (v, e, v2, e2). Maintained by BUAIIR.
Separate from Papoli community speech at BUAIR/popolivoice.
from datasets import load_dataset, Audio
ds = load_dataset("BUAIR/buaiir_voice_jap", split="train")
ds = ds.cast_column("audio", Audio(sampling_rate=16_000))
License: CC BY 4.0Updated: 2026-08-13
buaiir_voice_jap
BUAIIR Japadhola Voice (Bateesa/buaiir_voice_jap)
Separate dataset — structured student read-speech in Japadhola (Adhola, ISO 639-3: adh)
from Busitema University cohorts (Phase-2 batches v, e, v2, e2).
This repo does not include Papoli community recordings; those are published separately at
Bateesa/popolivoice.
Summary
Property
Value
Recordings
10,332 utterances
Duration
~27.6 hours
Language
Japadhola (adh)
Collection
Structured read speech… See the full description on the dataset page: https://huggingface.co/datasets/Bateesa/buaiir_voice_jap.
