CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01google /boolq Dataset Card for Boolq Dataset Summary BoolQ is a question answering dataset for yes/no questions containing 15942 examples. These questions are naturally occurring ---they are generated in unprompted and unconstrained settings. Each example is a triplet of (question, passage, answer), with the title of the page as optional additional context. The text-pair classification setup is similar to existing natural language inference tasks. Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/google/boolq.texttext-classification10K<n<100K109 likes211k downloads3y agoHugging Face02lighteval /boolq_helmtext10K<n<100K2 likes3.5k downloads1y agoHugging Face03automated-research-group /llama2_7b_chat-boolq-results Dataset Card for "llama2_7b_chat-boolq-results" More Information needed text100K<n<1M1 likes2k downloads3y agoHugging Face04hassansh /boolq_n_shottext10K<n<100K2 likes451 downloads3y agoHugging Face05sbhargav /boolq_nqtext10M<n<100M0 likes209 downloads1y agoHugging Face06fixie-ai /boolq-audio Dataset Card for Dataset Name This is a derivative of https://huggingface.co/datasets/google/boolq, but with an audio version of the questions as an additional feature. The audio was generated by running the existing question values through the Azure TTS generator with a 16KHz sample rate. Dataset Details Dataset Description Curated by: Fixie.ai Language(s) (NLP): English License: Creative Commons Share-Alike 3.0 license. Uses Training and… See the full description on the dataset page: https://huggingface.co/datasets/fixie-ai/boolq-audio.audiotext-classification10K<n<100K7 likes166 downloads2y agoHugging Face07stjokerli /TextToText_boolqtabular10K<n<100K0 likes157 downloads5y agoHugging Face08tytodd /boolq-qwen3-vl-32btext1K<n<10K0 likes138 downloads8mo agoHugging Face09ai4bharat /boolq-translatedtext10K<n<100K0 likes136 downloads3y agoHugging Face10manu /french_boolq Dataset Card for "test_fboolq" More Information needed textn<1K2 likes120 downloads3y agoHugging Face11Lots-of-LoRAs /task380_boolq_yes_no_question Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task380_boolq_yes_no_question Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks}… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task380_boolq_yes_no_question.texttext-generation1K<n<10K0 likes104 downloads2y agoHugging Face12sarvamai /boolq-indic Indic BoolQ Dataset A multilingual version of the BoolQ (Boolean Questions) dataset, translated from English into 10 Indian languages. It is a question-answering dataset for yes/no questions containing ~12k naturally occurring questions. Languages Covered The dataset includes translations in the following languages: Bengali (bn) Gujarati (gu) Hindi (hi) Kannada (kn) Marathi (mr) Malayalam (ml) Oriya (or) Punjabi (pa) Tamil (ta) Telugu (te) Dataset Format Each… See the full description on the dataset page: https://huggingface.co/datasets/sarvamai/boolq-indic.textquestion-answering100K<n<1M0 likes102 downloads2y agoHugging Face13learning-machine-inc /boolq-cot-opus5text1K<n<10K0 likes97 downloads5d agoHugging Face14KETI-NLP /kor_boolq Dataset Card for "kor_boolq" More Information needed Source Data Citation Information @inproceedings{clark2019boolq, title = {BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions}, author = {Clark, Christopher and Lee, Kenton and Chang, Ming-Wei, and Kwiatkowski, Tom and Collins, Michael, and Toutanova, Kristina}, booktitle = {NAACL}, year = {2019}, } text10K<n<100K1 likes84 downloads3y agoHugging Face15marcov /super_glue_boolq_promptsourcetext100K<n<1M1 likes78 downloads2y agoHugging Face16sbhargav /boolq_msmarcotext1M<n<10M0 likes78 downloads1y agoHugging Face17automated-research-group /boolq Dataset Card for "boolq" More Information needed text1K<n<10K1 likes67 downloads3y agoHugging Face18wojemann /snli_boolq Dataset Card for "snli_boolq" More Information needed text100K<n<1M1 likes58 downloads3y agoHugging Face19TTimur /boolq_kg BoolQ (Kyrgyz) This dataset is the Kyrgyz-translated version of the BoolQ benchmark, a reading comprehension task requiring a yes/no answer. 🏔️ Part of the KyrgyzLLM-Bench This dataset is a component of the KyrgyzLLM-Bench, a comprehensive suite for evaluating LLMs in Kyrgyz. Main Paper: Bridging the Gap in Less-Resourced Languages: Building a Benchmark for Kyrgyz Language Models Hugging Face Hub: https://huggingface.co/TTimur GitHub Project:… See the full description on the dataset page: https://huggingface.co/datasets/TTimur/boolq_kg.text10K<n<100K0 likes58 downloads11mo agoHugging Face20hishab /boolq_bn Dataset Summary BoolQ Bangla (BN) is a question-answering dataset for yes/no questions, generated using GPT-4. The dataset contains 15,942 examples, with each entry consisting of a triplet: (question, passage, answer). The questions are naturally occurring, generated from unprompted and unconstrained settings. Input passages were sourced from Bangla Wikipedia, Banglapedia, and News Articles, and GPT-4 was used to generate corresponding yes/no questions with answers. The dataset was… See the full description on the dataset page: https://huggingface.co/datasets/hishab/boolq_bn.textquestion-answering1K<n<10K1 likes57 downloads1y agoHugging Face21quarter100 /boolq_logtextn<1K0 likes55 downloads5y agoHugging Face22wojemann /snli_boolq_train Dataset Card for "snli_boolq_train" This dataset contains validation and training data from both boolq and snli snli_boolq_train is a mixed dataset containing both boolq (https://huggingface.co/datasets/boolq) and the "entail" and "contradict" samples from snli (https://huggingface.co/datasets/snli). The selection of this data was to increase the finetuning dataset amount for large language models with uninitialized weights. neutral statements were removed to… See the full description on the dataset page: https://huggingface.co/datasets/wojemann/snli_boolq_train.text100K<n<1M1 likes53 downloads3y agoHugging Face23jasonkrone /boolq_with_dev_hpotabular10K<n<100K0 likes50 downloads2y agoHugging Face24wojemann /tars_boolq Dataset Card for "tars_boolq" More Information needed text10K<n<100K0 likes46 downloads3y agoHugging Face25yjoonjang /boolq_ragsoluted_qwentext1K<n<10K0 likes45 downloads1y agoHugging Face26Lots-of-LoRAs /task381_boolq_question_generation Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task381_boolq_question_generation Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks}… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task381_boolq_question_generation.texttext-generation1K<n<10K0 likes37 downloads2y agoHugging Face27automated-research-group /phi-boolq-results Dataset Card for "phi-boolq-results" More Information needed text1K<n<10K1 likes35 downloads3y agoHugging Face28automated-research-group /llama2_7b_chat-boolq Dataset Card for "llama2_7b_chat-boolq" More Information needed tabular1K<n<10K0 likes35 downloads3y agoHugging Face29reaganjlee /boolq_ar Dataset Card for "boolq_ar" More Information needed text10K<n<100K0 likes34 downloads3y agoHugging Face30lurosenb /boolq_reformattedtext10K<n<100K0 likes34 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.