datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
customer-service-robot-support
This Dialogue
Comprised of fictitious examples of dialogues between a customer encountering problems with a robotic arm and a technical support agent. Check out the example below:
"id": 1,
"description": "Robotic arm calibration issue",
"dialogue": "Customer: My robotic arm seems to be misaligned. It's not picking objects accurately. What can I do? Agent: It appears that the arm may need recalibration. Please follow the instructions in the user manual to reset the calibration… See the full description on the dataset page: https://huggingface.co/datasets/FunDialogues/customer-service-robot-support.arabic-rag-support-25K
Arabic RAG customer-support scenarios (27,927 rows)
Synthetic Modern Standard Arabic customer-support scenarios for training small
RAG answerers, distilled from unsloth/gemma-4-31B-it-NVFP4 on a local vLLM.
Built as the training set for oddadmix/Nawah-50M-RAG-Support.
Each row: a customer question + the knowledge-base chunks of one fictional
company (products, prices, policies, FAQ entries) + the ideal grounded agent
answer. One generation request invents one company KB and 4 QA… See the full description on the dataset page: https://huggingface.co/datasets/oddadmix/arabic-rag-support-25K.NAITS_LFQA_with_supports_v2
NAITS_LFQA_with_supports_v2
Dataset Description
NAITS_LFQA_with_supports_v2 is a bilingual Arabic-English dataset designed for Long-Form Question Answering (LFQA) and Retrieval-Augmented Generation (RAG) research.
The dataset contains 102 manually curated samples. Each sample consists of:
A question in Arabic and English.
A long-form answer in Arabic and English.
Supporting passages in Arabic and English from which the answer can be derived.
A reference field… See the full description on the dataset page: https://huggingface.co/datasets/fahdsoliman/NAITS_LFQA_with_supports_v2.
