CoolFace
Datasetpublic

tuskanny/fiqa_lateon

FiQA-2018, LateOn Token-level (late-interaction) embeddings of the BEIR FiQA-2018 corpus and queries, encoded with LateOn, in the TACHIOM multivector format. Source BEIR FiQA-2018, test split. Corpus, queries and qrels were read from the official BEIR files via ir_datasets (beir/fiqa/test); PyLate only did the encoding 57,638 documents, 648 queries, 1,706 qrels Text given to the encoder for each document: the passage text (FiQA documents have no title). The text… See the full description on the dataset page: https://huggingface.co/datasets/tuskanny/fiqa_lateon.

sourceHugging Faceupdated 3d agoView on Hugging Face
0likes41downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
tuskanny/fiqa_lateon · CoolFace