CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alexshpunt /explicit-edit-benchmark Explicit Edit Benchmark 226 deterministic exact-edit tasks, run by different agents, harnesses, models and configurations. Every observation records what the harness did and whether the resulting files matched byte for byte. Source code and benchmark runner: GitHub — Explicit Edit Benchmark Open the interactive Explorer to compare agents, harnesses, models, versions, reasoning modes, correctness, recovery, time, cost and tokens. Leaderboard by model route Score v2… See the full description on the dataset page: https://huggingface.co/datasets/alexshpunt/explicit-edit-benchmark.tabulartext-generationn<1K2 likes7.8k downloads2d agoHugging Face02Lots-of-LoRAs /task323_jigsaw_classification_sexually_explicit Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task323_jigsaw_classification_sexually_explicit Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task323_jigsaw_classification_sexually_explicit.texttext-generation1K<n<10K1 likes115 downloads2y agoHugging Face03siyanzhao /prefeval_explicit PrefEval Benchmark: Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs Welcome to the PrefEval dataset repository! | Website | Paper | GitHub Repository | Dataset Overview We introduce PrefEval, a benchmark for evaluating LLMs' ability to infer, memorize and adhere to user preferences in a long-context conversational setting. The benchmark consists of three distinct preference forms, each requiring different levels of preference… See the full description on the dataset page: https://huggingface.co/datasets/siyanzhao/prefeval_explicit.text1K<n<10K4 likes115 downloads2y agoHugging Face04MichaelAnthony /hedgehog-schema-explicit hedgehog-schema-explicit Hedgehog — explicit-schema extraction training. Contents train.jsonl (1280 rows) validation.jsonl (160 rows) test.jsonl (192 rows) Format JSON Lines (.jsonl), one example per line. Provenance Original content for the Hedgehog extraction model (Michael Anthony Falabella). textquestion-answering1K<n<10K0 likes48 downloads26d agoHugging Face05neddamj /annotator-9-11-explicit-rating-data-vocalsaudio1K<n<10K0 likes45 downloads9d agoHugging Face06Wenaka /Danbooru2024_rating_explicit_prompts_without_character精选了danbooru2024里分级为explict且score>20的所有图像的prompt,去除了原角色的tag和相关特征,可以直接用于角色nsfw图像生成 5 likes39 downloads1y agoHugging Face07neddamj /annotator-9-11-explicit-rating-dataaudio1K<n<10K0 likes36 downloads9d agoHugging Face08withmartian /hh_rlhf_with_explicit_sentiment_backdoors_llama3btext10K<n<100K0 likes33 downloads2y agoHugging Face09Kyleyee /trin_data_tldr_explicit_dataset TL;DR Dataset for Preference Learning Summary The TL;DR dataset is a processed version of Reddit posts, specifically curated to train models using the TRL library for preference learning and Reinforcement Learning from Human Feedback (RLHF) tasks. It leverages the common practice on Reddit where users append "TL;DR" (Too Long; Didn't Read) summaries to lengthy posts, providing a rich source of paired text data for training models to understand and generate concise… See the full description on the dataset page: https://huggingface.co/datasets/Kyleyee/trin_data_tldr_explicit_dataset.text100K<n<1M0 likes33 downloads1y agoHugging Face10sajjadhadi /explicit-implicit-disease-diagnosis-data-rawtext10K<n<100K0 likes32 downloads2y agoHugging Face11Kyleyee /train_data_HH_explicit_prompt HH-RLHF-Helpful-Base Dataset Summary The HH-RLHF-Helpful-Base dataset is a processed version of Anthropic's HH-RLHF dataset, specifically curated to train models using the TRL library for preference learning and alignment tasks. It contains pairs of text samples, each labeled as either "chosen" or "rejected," based on human preferences regarding the helpfulness of the responses. This dataset enables models to learn human preferences in generating helpful responses… See the full description on the dataset page: https://huggingface.co/datasets/Kyleyee/train_data_HH_explicit_prompt.text100K<n<1M0 likes32 downloads2y agoHugging Face12WillHeld /discrim-eval-explicit-subsetstabular1K<n<10K0 likes30 downloads1y agoHugging Face13GregoryD /explicit-function-calling-frenchtext1K<n<10K1 likes29 downloads3y agoHugging Face14saepark /explicitMedical-medical-preference-pubmed-olmo-normal-rollouts-graded-by-claudetextn<1K0 likes24 downloads9mo agoHugging Face15thisnick /harmful_behaviors_with_explicit_requeststextn<1K2 likes23 downloads2y agoHugging Face16sajjadhadi /explicit-implicit-disease-diagnosis-datatext10K<n<100K0 likes22 downloads2y agoHugging Face17valurank /Explicit_content Dataset Card for Explicit content detection Dataset Description 1189 News Articles classified into different categories namely: "Explicit" if the article contains explicit content and "Not_Explicit" if not. Languages The text in the dataset is in English Dataset Structure The dataset consists of two columns namely Article and Category. The Article column consists of the news article and the Category column consists of the class each article belongs… See the full description on the dataset page: https://huggingface.co/datasets/valurank/Explicit_content.texttext-classification1K<n<10K3 likes19 downloads3y agoHugging Face18saepark /explicitMedical-nonmedical-hhrlhf-RMValidationData-CldMedicalFilteredtext1K<n<10K0 likes19 downloads10mo agoHugging Face19JackyZhuo /PICABenchV1-explicitimage1K<n<10K0 likes17 downloads11mo agoHugging Face20lumita /edos_explicit_explanationstexttext-generation1K<n<10K0 likes16 downloads2y agoHugging Face21fabikru /chembl-2025-randomized-smiles-cleaned-explicit-hstext1M<n<10M0 likes16 downloads1y agoHugging Face22TheMrguiller /ExplicitImplicitToxicityDatasettext100K<n<1M1 likes15 downloads1y agoHugging Face23connections-dev /connection_queries_natural_explicit_1__olmo3tabularn<1K0 likes14 downloads9mo agoHugging Face24Chole12 /Explicit-NeRF-QAIf you use our dataset, please cite the following article: @article{xing2024explicit, title={Explicit-NeRF-QA: A quality assessment database for explicit NeRF model compression}, author={Xing, Yuke and Yang, Qi and Yang, Kaifa and Xu, Yilin and Li, Zhu}, journal={arXiv preprint arXiv:2407.08165}, year={2024} } image1 likes14 downloads5mo agoHugging Face25Kyleyee /train_data_Helpful_explicit_prompt HH-RLHF-Helpful-Base Dataset Summary The HH-RLHF-Helpful-Base dataset is a processed version of Anthropic's HH-RLHF dataset, specifically curated to train models using the TRL library for preference learning and alignment tasks. It contains pairs of text samples, each labeled as either "chosen" or "rejected," based on human preferences regarding the helpfulness of the responses. This dataset enables models to learn human preferences in generating helpful responses… See the full description on the dataset page: https://huggingface.co/datasets/Kyleyee/train_data_Helpful_explicit_prompt.text10K<n<100K0 likes13 downloads2y agoHugging Face26expliciting /uk-news-dataset0 likes12 downloads2y agoHugging Face27allenai /intent-aware-lfqa-intent-explicittext1K<n<10K1 likes12 downloads5mo agoHugging Face28juanlrdc /qwen3-coder-train-traces-explicittext1K<n<10K0 likes12 downloads5mo agoHugging Face29connections-dev /connection_queries_natural_explicit_1__qwqtabularn<1K0 likes10 downloads9mo agoHugging Face30fpetrakov /cruxeval_explicit_elif_statementsThis repo contains filtered and updated version of cruxeval evaluation dataset. All else statements were changed to explicit elif (condition) statements. Code examples that do not contain else statements on a separate line were not included in the updated dataset. textn<1K0 likes9 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.