CoolFace
7 results

cqadupstack_stats

mteb /cqadupstack-stats CQADupstackStatsRetrieval An MTEB dataset Massive Text Embedding Benchmark CQADupStack: A Benchmark Data Set for Community Question-Answering Research Task category t2t Domains Written, Academic, Non-fiction Referencehttp://nlp.cis.unimelb.edu.au/resources/cqadupstack/ How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task = mteb.get_tasks(["CQADupstackStatsRetrieval"]) evaluator =… See the full description on the dataset page: https://huggingface.co/datasets/mteb/cqadupstack-stats.texttext-retrieval10K<n<100K0 likes1.1k downloads1y agoHugging Facemteb /CQADupstack-Stats-PL CQADupstack-Stats-PL An MTEB dataset Massive Text Embedding Benchmark CQADupStack: A Stack Exchange Question Duplicate Pairs Dataset Task category t2t Domains Written, Academic, Non-fiction Reference https://huggingface.co/datasets/clarin-knext/cqadupstack-stats-pl How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task = mteb.get_tasks(["CQADupstack-Stats-PL"]) evaluator =… See the full description on the dataset page: https://huggingface.co/datasets/mteb/CQADupstack-Stats-PL.texttext-retrieval10K<n<100K0 likes48 downloads1y agoHugging FaceGreenNode /cqadupstack-stats-vn How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task = mteb.get_tasks(["CQADupstackStats-VN"]) evaluator = mteb.MTEB(task) model = mteb.get_model(YOUR_MODEL) evaluator.run(model) To learn more about how to run models on mteb task check out the GitHub repitory. Citation If you use this dataset, please cite the dataset as well as mteb, as this dataset likely includes additional processing as a… See the full description on the dataset page: https://huggingface.co/datasets/GreenNode/cqadupstack-stats-vn.texttext-retrieval10K<n<100K0 likes40 downloads1y agoHugging FaceHyukkyu /beir-cqadupstack-stats CQADupstackStatsRetrieval — BEIR, unified schema A normalised copy of the dataset behind the mteb task CQADupstackStatsRetrieval, one of the tasks of the BEIR benchmark as mteb defines it (a member of the aggregate task CQADupstackRetrieval). Same queries, documents and relevance judgements as the benchmark evaluates — reshaped into one strict schema shared by every dataset in this collection. Source mteb/cqadupstack-stats @ 65ac3a16b8e9 (the revision pinned in… See the full description on the dataset page: https://huggingface.co/datasets/Hyukkyu/beir-cqadupstack-stats.texttext-retrieval10K<n<100K0 likes38 downloads16d agoHugging FaceMCINext /cqadupstack-stats-fa Dataset Summary CQADupstack-stats-Fa is a Persian (Farsi) dataset developed for the Retrieval task, focusing on duplicate question detection in community question-answering (CQA) platforms. This dataset is a translated version of the "Cross Validated" (Stats) subforum from the English CQADupstack collection and is part of the FaMTEB benchmark under the BEIR-Fa suite. Language(s): Persian (Farsi) Task(s): Retrieval (Duplicate Question Retrieval) Source: Translated from English… See the full description on the dataset page: https://huggingface.co/datasets/MCINext/cqadupstack-stats-fa.text10K<n<100K0 likes20 downloads1y agoHugging Faceclarin-knext /cqadupstack-stats-plPart of BEIR-PL: Zero Shot Information Retrieval Benchmark for the Polish Language. Link to arxiv: https://arxiv.org/pdf/2305.19840.pdf Contact: konrad.wojtasik@pwr.edu.pl 0 likes19 downloads2y agoHugging Face