datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MobileBench
MobileBench: The On-Device LLM Benchmark
A standardized evaluation benchmark designed specifically for mobile and edge-deployed language models.
Why MobileBench?
Existing benchmarks (MMLU, HumanEval, GSM8K) test what large models can do on servers. MobileBench tests what small models can do on phones — the tasks users actually perform:
Summarization — The #1 on-device task (messages, emails, notifications)
Classification — Spam detection, sentiment, intent… See the full description on the dataset page: https://huggingface.co/datasets/dispatchAI/MobileBench.gulf-climate-dataset
Gulf Climate & Environmental Dataset
Curated environmental data for the Arabian Peninsula — built for on-device climate
modeling with a regional focus no San Francisco lab would build.
Contents
File
Description
gulf_monthly_temperatures.csv
Monthly average temperatures for 11 Gulf cities
gulf_climate_data.json
Full structured data (temps, dust storms, solar, Q&A)
climate_instructions.jsonl
Instruction-tuning pairs for climate Q&A (Arabic + English)… See the full description on the dataset page: https://huggingface.co/datasets/dispatchAI/gulf-climate-dataset.
