datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
perspective-information-retrieval-allsidesperspective-information-retrieval-perspectrumperspective-information-retrieval-agnewsperspective-information-retrieval-ambigqaperspective-information-retrieval-storyQuick Actions
perspective-information-retrieval-exfeverVisual_information_retrieval
GDZ Scientific Document Retrieval Benchmark
A needle‑in‑a‑haystack benchmark for scientific document retrieval, built from historical volumes of the Göttinger Digitalisierungszentrum (GDZ). This dataset explicitly adapts the IRPAPERS methodology onto a real‑world, multilingual corpus to evaluate both text-based and visual document retrieval models.
Dataset Structure
The dataset is divided into two operational configurations:
1. queries
Contains the… See the full description on the dataset page: https://huggingface.co/datasets/Trungdaik/Visual_information_retrieval.Advanced-Information-Retrieval2AIRTC-Amharic-Adhoc-Information-Retrieval-Test-Collection
Original Dataset and Paper
Original dataset: https://www.irit.fr/AmharicResources/airtc-the-amharic-adhoc-information-retrieval-test-collection/
Evaluation is highly important for designing, developing, and maintaining information retrieval (IR) systems. The IR community has developed shared tasks where evaluation framework, evaluation measures and test collections have been developed for different languages. Although Amharic is the official language of Ethiopia currently having an… See the full description on the dataset page: https://huggingface.co/datasets/rasyosef/2AIRTC-Amharic-Adhoc-Information-Retrieval-Test-Collection.Information_Retrievalphysbert_information_retrievalFinsights-Grey-RAG-Effective-Information-Retrieval-logsFinsights-Grey-RAG-Effective-Information-Retrieval-logsFinsightsGrey-RAG-InformationRetrieval-log
