CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01allenai /tulu-3-harmbench-evalThis data comes from the HarmBench benchmark. This is one of the datasets included in the Ai2 Safety Evaluation Suite, and the Tülu 3 evaluation suite. The repo for Ai2's safety suite includes instructions on how to evaluate models on various safety-related evaluation including this one. textn<1K3 likes445 downloads1y agoHugging Face02anasedova /tulu_3_factual_errors Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/anasedova/tulu_3_factual_errors.text10K<n<100K0 likes10 downloads2y agoHugging Face03anasedova /tulu_3_formatting_errorstextn<1K0 likes10 downloads2y agoHugging Face04anasedova /tulu_3_all_errors_updtabular100K<n<1M0 likes10 downloads2y agoHugging Face05anasedova /tulu_3_no_errorstext100K<n<1M0 likes8 downloads2y agoHugging Face06anasedova /tulu_3_incorrect_output_errorstext100K<n<1M0 likes6 downloads2y agoHugging Face07anasedova /tulu_3_underspecified_input_errorstext10K<n<100K0 likes6 downloads2y agoHugging Face08yuxixia /triviaqa-test-tulu3-querytext1K<n<10K0 likes5 downloads2y agoHugging Face09anasedova /tulu_3_model_modality_mismatch_errorstext1K<n<10K0 likes4 downloads2y agoHugging Face10anasedova /tulu_3_whole_updatedtext100K<n<1M0 likes2 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.