Henok/aya_amharic_dataset
AYA Amharic Dataset This is the Amharic-only extract from the Aya Dataset, a multilingual instruction fine-tuning dataset. By @henok Dataset Summary The Aya Dataset is a multilingual instruction fine-tuning dataset curated by an open-science community via Aya Annotation Platform from Cohere For AI. The dataset contains a total of 204k human-annotated prompt-completion pairs along with the demographics data of the annotators. This dataset can be used to train… See the full description on the dataset page: https://huggingface.co/datasets/Henok/aya_amharic_dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face