CoolFace
Datasetpublic

kreasof-ai/bigc-bem-eng

Dataset Details This is dataset of speech translation task for Bemba-to-English Language. This dataset is acquired from (Big-C)[https://github.com/csikasote/bigc] github repository. Big-C is a large conversations dataset between Bemba Speakers based on Image [1]. This dataset provide data for speech translation. Preprocessing Steps Some preprocessing was done in this dataset. Drop some unused columns other than audio_id, sentence, translation, and speaker_id.… See the full description on the dataset page: https://huggingface.co/datasets/kreasof-ai/bigc-bem-eng.

sourceHugging Faceupdated 1y agoView on Hugging Face
1likes214downloads
3 commits on main
a0ee81b1y ago

Update README.md

cobrayyxx
e182c1b2y ago

Upload dataset

cobrayyxx
122c0e42y ago

initial commit

cobrayyxx