datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wikipedia-2023-11-kikongo-lingala-cohere-multilingual-v3Code-170k-lingala
Dataset Description
Code-170k-lingala is a groundbreaking dataset containing 176,999 programming conversations, originally sourced from glaiveai/glaive-code-assistant-v2 and translated into Lingala, making coding education accessible to Lingala speakers.
🌟 Key Features
176,999 high-quality conversations about programming and coding
Pure Lingala language - democratizing coding education
Multi-turn dialogues covering various programming concepts
Diverse topics: algorithms… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/Code-170k-lingala.Code-170k-lingala
Dataset Description
Code-170k-lingala is a groundbreaking dataset containing 176,999 programming conversations, originally sourced from glaiveai/glaive-code-assistant-v2 and translated into Lingala, making coding education accessible to Lingala speakers.
🌟 Key Features
176,999 high-quality conversations about programming and coding
Pure Lingala language - democratizing coding education
Multi-turn dialogues covering various programming concepts
Diverse topics: algorithms… See the full description on the dataset page: https://huggingface.co/datasets/tokossapp/Code-170k-lingala.
