datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Nigeria_Machinery_Dataset
Nigeria Machinery Usage and Failures Dataset
A structured numeric dataset covering machinery usage rates, equipment failures,
capacity utilization, maintenance costs, and operational downtime across Nigeria's
industrial manufacturing and oil & gas sectors, 2006–2025. It ships
with a companion chain-of-thought reasoning layer derived directly from the
records, for fine-tuning and evaluating LLMs on domain-grounded numeric tasks.
This dataset addresses a real gap: machine-level… See the full description on the dataset page: https://huggingface.co/datasets/gospelgit/Nigeria_Machinery_Dataset.African-Languages_Sentiments
African Languages Sentiment Dataset (Hausa, Yorùbá, Swahili)
A stitched multi-source sentiment classification dataset combining three
independently collected sentiment corpora for Hausa, Yorùbá, and Swahili,
built for the Adaption Labs AutoScientist Challenge
(Language category).
Companion model: fine-tuned weights trained on the adapted version of this dataset via
AutoScientist are released separately at… See the full description on the dataset page: https://huggingface.co/datasets/gospelgit/African-Languages_Sentiments.
