CoolFace
Datasetpublic

Tim2190/kaz-rag-search-benchmark

Kaz-RAG-Search-Benchmark Evidence-based benchmark for Kazakh information retrieval — the independent proof base for the Kazakh Stemmer. Corpus: 8,370 passages from Kazakh Wikipedia Queries: 300 queries × 3 categories (natural / inflected / vocabulary-gap) Format: BEIR-compatible — three subsets: corpus, queries, qrels Browse the data: use the subset switcher at the top of the Data Studio viewer to move between corpus (Kazakh passages), queries (the 300 questions), and qrels… See the full description on the dataset page: https://huggingface.co/datasets/Tim2190/kaz-rag-search-benchmark.

sourceHugging Facecc-by-sa-4.0updated 3mo agoView on Hugging Face
0likes23downloads

Tim2190/kaz-rag-search-benchmark · main · files are served by the source, never re-hosted here