CoolFace
Datasetpublic

Tim2190/kaz-rag-search-benchmark

Kaz-RAG-Search-Benchmark Evidence-based benchmark for Kazakh information retrieval — the independent proof base for the Kazakh Stemmer. Corpus: 8,370 passages from Kazakh Wikipedia Queries: 300 queries × 3 categories (natural / inflected / vocabulary-gap) Format: BEIR-compatible — three subsets: corpus, queries, qrels Browse the data: use the subset switcher at the top of the Data Studio viewer to move between corpus (Kazakh passages), queries (the 300 questions), and qrels… See the full description on the dataset page: https://huggingface.co/datasets/Tim2190/kaz-rag-search-benchmark.

sourceHugging Facecc-by-sa-4.0updated 3mo agoView on Hugging Face
0likes23downloads
9 commits on main
4d39ef33mo ago

Update README.md

Tim2190
2dc62e63mo ago

Update README.md

Tim2190
b2b3a934mo ago

Update README.md

Tim2190
ea741e04mo ago

Update README.md

Tim2190
a3749ae4mo ago

Update README.md

Tim2190
bfe9eba4mo ago

Update README.md

Tim2190
144d8974mo ago

Upload folder using huggingface_hub

Tim2190
16c3a734mo ago

Upload folder using huggingface_hub

Tim2190
e1cf91b4mo ago

initial commit

Tim2190