datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
linguistic-similarityadaption-language-linguistics-qa
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-language_linguistics_qa
This dataset consists of instruction and response pairs covering a broad range of topics within language and linguistics. Entries address fundamental concepts such as grammar syntax, vocabulary, pronunciation, and writing systems, alongside applied disciplines like computational linguistics, translation, and localization. Additional content explores language… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/adaption-language-linguistics-qa.Yoruba-linguistics-dataset
Yoruba Qwen Fine-Tuned Model
Overview
This model is a LoRA fine-tuned version of Qwen2.5-0.5B-Instruct developed for Yoruba language instruction-following tasks.
The project explores the use of parameter-efficient fine-tuning for low-resource African language NLP, with a particular focus on Yoruba.
Dataset
The dataset was adapted using Adaption Lab and contains:
Total records: 3,598
Training samples: 3,238
Validation samples: 360
Language: Yoruba… See the full description on the dataset page: https://huggingface.co/datasets/Kolawolettyy12/Yoruba-linguistics-dataset.
