datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
asteria-bhojpuri-assamese-civic-qa
Asteria — Bhojpuri & Assamese Civic Q&A Dataset
A dataset of government scheme Q&A pairs in Bhojpuri and Assamese — two low-resource Indian languages.
Dataset Description
This dataset was collected by Asteria, an AI Agent built for the AI Agents Hackathon 2026. The agent helps rural Indian citizens access government welfare schemes by conversing in their native language.
Supported Languages
Bhojpuri (bho) — spoken by 50+ million people in Bihar, UP… See the full description on the dataset page: https://huggingface.co/datasets/Afuu-coder/asteria-bhojpuri-assamese-civic-qa.xnli2.0_train_bhojpuri
