CoolFace
Datasetpublic

Henok/aya_amharic_dataset

AYA Amharic Dataset This is the Amharic-only extract from the Aya Dataset, a multilingual instruction fine-tuning dataset. By @henok Dataset Summary The Aya Dataset is a multilingual instruction fine-tuning dataset curated by an open-science community via Aya Annotation Platform from Cohere For AI. The dataset contains a total of 204k human-annotated prompt-completion pairs along with the demographics data of the annotators. This dataset can be used to train… See the full description on the dataset page: https://huggingface.co/datasets/Henok/aya_amharic_dataset.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes26downloads

Henok/aya_amharic_dataset · main · files are served by the source, never re-hosted here