ghomala
Datasets
All datasets matching “ghomala”ghomala-spoken-bible
Ghomálá' Spoken New Testament — aligned audio + trilingual text
Part of the Lingo / NativeAI language-preservation project. This is
~20 hours of spoken Ghomálá' (Ghomala, ISO bbj; a Grassfields Bantu language of
West Cameroon) — recorded readings of the New Testament — aligned chapter-by-chapter
with parallel text in Ghomálá', French, and English.
Spoken-language data is exactly what oral-first Cameroonian languages lack, which makes
this a rare resource for building ASR, TTS… See the full description on the dataset page: https://huggingface.co/datasets/flagship-ai/ghomala-spoken-bible.mlx-french-ghomala-bandjounFrench to ghomala translation dataset optimized for MLX tuning
french-ghomala-bandjoun
Dataset Card for Dataset Name
Dataset containing translation of a given set of French words and expressions to Ghomala, the native language of Bandjoun, a village part of Cameroun (Central Africa). Its aim is to be used to tune LLMs in recognising ghomala and support translation from/to that language.
Dataset extracted from Dictionnaire Ghomálá’-Français edited by Erika EICHHOLZER with Prof. Dr. Engelbert DOMCHE-TEKO, Dr. Gabriel MBA and P. Gabriel NISSIM, Version 2.0, Décembre… See the full description on the dataset page: https://huggingface.co/datasets/stfotso/french-ghomala-bandjoun.french-ghomala-bandjoun-llm-readyFrench to ghomala translation dataset optimized for IBM Granite LLM tuning
english_ghomala
