fatymahaly/urdu_rag_dataset.csv
Dataset Card for Urdu RAG Knowledge Base Dataset Overview This dataset is designed specifically to bootstrap and evaluate Retrieval-Augmented Generation (RAG) applications, search systems, and semantic retrieval pipelines using the Urdu language. It contains 185 clean, structured, and informative text chunks covering a wide array of domains. Language: Urdu (ur) Script: Nastaliq / Arabic script (Unicode UTF-8) Total Rows: 185 chunks Format: CSV (id, title… See the full description on the dataset page: https://huggingface.co/datasets/fatymahaly/urdu_rag_dataset.csv.
This repository belongs to fatymahaly on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
