CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01vidore /vidore_v3_computer_scienceViDoRe V3 : Computer Science This dataset, Computer Science, is a corpus of textbooks from the openstacks website, intended for long-document understanding tasks. It is one of the 10 corpora comprising the ViDoRe v3 Benchmark. About ViDoRe v3 ViDoRe V3 is our latest benchmark for RAG evaluation on visually-rich documents from real-world applications. It features 10 datasets with, in total, 26,000 pages and 3099 queries, translated into 6 languages. Each query comes with… See the full description on the dataset page: https://huggingface.co/datasets/vidore/vidore_v3_computer_science.documentvisual-document-retrieval1K<n<10K6 likes2.2k downloads8mo agoHugging Face02WenxingZhu /vidore_v3_computer_science_embeddingNOTE ViDoRe V3: Computer Science dataset ColQwen2 Embeddings This dataset contains pre-computed embeddings for the ViDoRe V3 : Computer Science dataset using the ColQwen2 model. ViDoRe V3 : Computer Science This dataset, Computer Science, is a corpus of textbooks from the openstacks website, intended for long-document understanding tasks. It is one of the 10 corpora comprising the ViDoRe v3 Benchmark. About ViDoRe v3 ViDoRe V3 is our latest benchmark for RAG evaluation on… See the full description on the dataset page: https://huggingface.co/datasets/WenxingZhu/vidore_v3_computer_science_embedding.document1K<n<10K0 likes56 downloads11mo agoHugging Face03gabrieljimenez /epfl-computer-science-mcqatabularn<1K0 likes40 downloads1y agoHugging Face04Jerichog0731 /vidore_v3_computer_scienceViDoRe V3 : Computer Science This dataset, Computer Science, is a corpus of textbooks from the openstacks website, intended for long-document understanding tasks. It is one of the 10 corpora comprising the ViDoRe v3 Benchmark. About ViDoRe v3 ViDoRe V3 is our latest benchmark for RAG evaluation on visually-rich documents from real-world applications. It features 10 datasets with, in total, 26,000 pages and 3099 queries, translated into 6 languages. Each query comes with… See the full description on the dataset page: https://huggingface.co/datasets/Jerichog0731/vidore_v3_computer_science.documentvisual-document-retrieval1K<n<10K0 likes23 downloads3mo agoHugging Face05robro612 /vidore3_computerscience_neomme_260m_li vidore3_computerscience_neomme_260m_li Multi-vector (late-interaction) embeddings of ViDoRe computerscience (vidore/computerscience), encoded with Hcompany/NeoMME-260M-Retriever-ST-late at revision 023be2a8ab9d797f5aa76f5bf8b5dde78d819659. Source data: Hugging Face dataset vidore/vidore_v3_computer_science at revision d5cc75883d92e294f0c0fc2662551c9708a06ebc, configs corpus / queries / qrels, split test, loaded with datasets. Document, query and qrel ids are the source's own ids… See the full description on the dataset page: https://huggingface.co/datasets/robro612/vidore3_computerscience_neomme_260m_li.tabular1K<n<10K0 likes12 downloads21h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.