CoolFace
Datasetpublic

MCINext/cqadupstack-tex-fa

Dataset Summary CQADupstack-tex-Fa is a Persian (Farsi) dataset curated for the Retrieval task, specifically targeting duplicate question detection in community question-answering (CQA) forums. This dataset is a translation of the "TeX - LaTeX" StackExchange subforum from the English CQADupstack collection and is part of the FaMTEB benchmark under the BEIR-Fa suite. Language(s): Persian (Farsi) Task(s): Retrieval (Duplicate Question Retrieval) Source: Translated from… See the full description on the dataset page: https://huggingface.co/datasets/MCINext/cqadupstack-tex-fa.

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes21downloads
corpus.jsonl4 linesDownload Raw Back to root
1version https://git-lfs.github.com/spec/v12oid sha256:a59a1799d68169175b0da097cfcbfd555c545d28a4a769714b4760bd763468483size 2220618694