datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
researchy_questions
Introduction
Researchy Questions is a set of about 100k Bing queries that users spent the most effort on. After a labor-intensive filtering funnel from billions of queries, these "needles in the haystack" are non-factoid, multi-perspective questions that probably require a lot of sub-questions and research in order to answer adequetly. These questions are shown to be harder than other open domain QA datasets like Natural Questions.
The train dataset has about 90k samples.… See the full description on the dataset page: https://huggingface.co/datasets/corbyrosset/researchy_questions.forecastbench-single_question
ForecastBench Single Questions
This dataset contains single-ID forecasting questions derived from the ForecastBench project. It includes two configurations:
forecastbench_single_questions_2024-12-08: Contains 429 forecasting questions with resolved real-world outcomes.
forecastbench_single_questions_human_2024-07-21: Contains 473 questions with resolved real-world outcomes, augmented with human forecast probabilities from public and superforecaster groups.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Duruo/forecastbench-single_question.scientific-question-outcomes
Scientific Question Outcomes
980 astronomy research questions, frozen at five historical cutoffs, each
labelled with what the following five years of literature actually did with
it.
Systems that propose research questions are usually evaluated by asking a
person or a model how good the questions sound. This dataset supplies the
alternative: questions frozen using only pre-cutoff literature, and outcome
labels drawn from the literature published afterwards. It is, to our… See the full description on the dataset page: https://huggingface.co/datasets/huiluckylucky/scientific-question-outcomes.stackoverflow_DL-related_questionseu-tenders-with-questions-for-agentic-checklist-filling
eu-tenders-with-questions-for-agentic-checklist-filling
Dataset Description
This dataset contains questions and answers for evaluating Retrieval-Augmented Generation (RAG) systems in the context of generative agentic checklist-filling. The dataset is designed to benchmark various RAG architectures (Hybrid RAG, Graph RAG, Multi-Hop/Agentic RAG) on document analysis tasks.
Dataset Summary
Total Questions: 97
Document Families: 7
Languages: EN
Domain: Procurement… See the full description on the dataset page: https://huggingface.co/datasets/tmskss/eu-tenders-with-questions-for-agentic-checklist-filling.vietquill-qcpg-100k-synthesis-questionmedical_questions_pairs_koOriginal Data: curaihealth/medical_questions_pairs.
Translated into Korean by "solar-1-mini-translate-enko".
turkish-question-departmenthindi-driving-test-questionsEntity-deduction-arena-20-questions
twentyquestions
The twentyquestios dataset provides
20 Questions style questions
and answers collected from real games of twentyquestions between people on
Mechanical Turk.
In this directory, you'll find the following files:
twentyquestions-train.jsonl
twentyquestions-dev.jsonl
twentyquestions-test.jsonl
twentyquestions-all.jsonl
The files represent the train, dev, and test splits as well as all the data
together before being split. Train, dev, and test were split by dividing up… See the full description on the dataset page: https://huggingface.co/datasets/jtv199/Entity-deduction-arena-20-questions.fincen_all_questions_5versions
About
These question-answer pairs are created using published pdf documents at fincen.gov.
Each question has 5 paraphased versions differentiated by column "question_version" (the first versions (No. 4) are at the end of the datafile).
The data is used to fine-tune Gemma-2b and Gemma-7b listed here
shijunju/gemma_7b_finRisk_r10_4VersionQ
shijunju/gemma_7b_finRisk_r6_4VersionQ
shijunju/gemma_7b_finRisk_r6_3VersionQ
shijunju/gemma_2b_finRisk
Number of rows: 14,550
Author: Shijun… See the full description on the dataset page: https://huggingface.co/datasets/shijunju/fincen_all_questions_5versions.vqa_training_questionsquestion_answerrkf-questionswildchat-oracle-questions-mini-gemini3sinhala-alevel-physics-questions
Dataset Details
This dataset contains 20 physics questions and answers focused on Sinhala language.
washnorm2021_test_questions
Dataset Card for "WASHNORM 2021 Test Questions and Answers"
This dataset contains 90 Question and Answer pairs, 2 extra reference answers for each question. It was created from the WASHNORM 2021 Report's Executive Summary which can be found on UNICEF Nigeria's Website.
Dataset Description
Notebook: Contains code where majority of data extraction and generation was carried out.
Repository: Contains code for the WASH Services Chatbot that this dataset was generated to test.… See the full description on the dataset page: https://huggingface.co/datasets/rnabage/washnorm2021_test_questions.ctet-hindi-questionswildchat-oracle-questions-1kfeedback_qesconv_badareas_questions_reflectionsValidation Results:
PAIR-Reflections:
• "Precision": 0.8955223880597015
• "Recall": 0.2553191489361703
• "Accuracy": 0.77
• "F1-score": 0.3973509933774836
Rule-based HasQuestions:
• "Precision": 1.0
• "Recall": 0.9629629629629629
• "Accuracy": 0.74
• "F1-score": 0.9811320754716981
ONC_Question_Similarity_And_Matching
