kreasof-ai/bigc-bem-eng
Dataset Details This is dataset of speech translation task for Bemba-to-English Language. This dataset is acquired from (Big-C)[https://github.com/csikasote/bigc] github repository. Big-C is a large conversations dataset between Bemba Speakers based on Image [1]. This dataset provide data for speech translation. Preprocessing Steps Some preprocessing was done in this dataset. Drop some unused columns other than audio_id, sentence, translation, and speaker_id.… See the full description on the dataset page: https://huggingface.co/datasets/kreasof-ai/bigc-bem-eng.
1214
Update README.md
Upload dataset
initial commit
