MCINext/persian-web-document-retrieval
Dataset Summary Persian Web Document Retrieval is a Persian (Farsi) dataset designed for the Retrieval task. It is a component of the FaMTEB (Farsi Massive Text Embedding Benchmark). This dataset consists of real-world queries collected from the Zarrebin search engine and web documents labeled by humans for relevance. It is curated to evaluate model performance in web search scenarios. Language(s): Persian (Farsi) Task(s): Retrieval (Web Search) Source: Collected from… See the full description on the dataset page: https://huggingface.co/datasets/MCINext/persian-web-document-retrieval.
This repository belongs to MCINext on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
