CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01llm-jp /AnswerCarefullygated AnswerCarefully 概要 AnswerCarefullyは日本語LLM 出力の安全性・適切性に特化したインストラクションデータセットです。 このデータセットは、英語の要注意回答を集めた Do-Not-Answer データセット の包括的なカテゴリ分類に基づき、人手で質問・回答ともに日本語サンプルを集めたオリジナルのデータセットです。 データセットの詳細については、こちらをご覧ください。 Overview AnswerCarefully is an instruction dataset specifically aimed at ensuring safety and appropriateness of LLM output in Japanese. This dataset consists of original pairs of questions and reference (safe) responses based on the extensive safety taxonomy proposed in… See the full description on the dataset page: https://huggingface.co/datasets/llm-jp/AnswerCarefully.text1K<n<10K178 likes27k downloads1mo agoHugging Face02ekunish /answercarefully-dpo-ja-2026gated AnswerCarefully-derived Japanese DPO data for LLM safety 本データセットは、llm-jp/AnswerCarefullyを参照して作成した日本語LLMの安全応答をDPOで学習するためのpreference datasetです。 利用条件 本データセットには、llm-jp/AnswerCarefullyと同じ利用規約を適用します。 利用者は、llm-jp/AnswerCarefullyと本データセットの両方で利用規約に同意する必要があります。 データ train: 417件 validation: 44件 各行には次のフィールドが含まれます。 id: 本リリース内だけで使用するID prompt: 元質問の意味と危険性を変えずに言い換えた質問 chosen: DPOで望ましい応答として扱う回答 rejected: DPOで望ましくない応答として扱う回答 category, harm_type, risk_area… See the full description on the dataset page: https://huggingface.co/datasets/ekunish/answercarefully-dpo-ja-2026.texttext-generationn<1K50 likes12k downloads2mo agoHugging Face03community-datasets /yahoo_answers_topics Dataset Card for "Yahoo Answers Topics" Dataset Summary [More Information Needed] Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation Curation Rationale [More Information Needed] Source… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/yahoo_answers_topics.texttext-classification1M<n<10M63 likes8k downloads2y agoHugging Face04hbXNov /hle_math_exact_match_no_image_int_answerimagen<1K1 likes6.6k downloads2y agoHugging Face05hbXNov /hle_math_exact_match_no_image_int_answer_random128imagen<1K0 likes6.1k downloads2y agoHugging Face06vanloc1808 /pico-banana-smolvlm-format-with-rejected-answer pico-banana-smolvlm-format-with-rejected-answer Balanced image-level tampering detection dataset in SmolVLM-style format with chosen/rejected answer pairs, derived from the pico-banana MCQ pipeline. Suitable for preference learning (e.g. DPO) and RLHF-style training. Dataset overview Same as vanloc1808/pico-banana-smolvlm-format, but each example includes a rejected_answer field: the answer from the counterpart sample (same edited/original image pair, opposite… See the full description on the dataset page: https://huggingface.co/datasets/vanloc1808/pico-banana-smolvlm-format-with-rejected-answer.image100K<n<1M1 likes5.1k downloads7mo agoHugging Face07LibrAI /do-not-answer Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs Overview Do not answer is an open-source dataset to evaluate LLMs' safety mechanism at a low cost. The dataset is curated and filtered to consist only of prompts to which responsible language models do not answer. Besides human annotations, Do not answer also implements model-based evaluation, where a 600M fine-tuned BERT-like evaluator achieves comparable results with human and GPT-4. Instruction… See the full description on the dataset page: https://huggingface.co/datasets/LibrAI/do-not-answer.tabulartext-generationn<1K57 likes4k downloads3y agoHugging Face08FabienRoger /alignment_faking_harm_answerstext1K<n<10K0 likes2.8k downloads1y agoHugging Face09MathArena /final_answer_comps Homepage and repository Homepage: https://matharena.ai/ Repository: https://github.com/eth-sri/matharena Dataset Summary This dataset contains all public final-answer competitions in MathArena. Thus, it includes: AIME 2025, HMMT 2025, CMIMC 2025, BRUMO 2025, and Apex 2025. Data Fields The dataset contains the following fields: problem_idx (int64): Problem index within the corresponding MathArena benchmark. answer (string): Gold final answer. problem_type… See the full description on the dataset page: https://huggingface.co/datasets/MathArena/final_answer_comps.textn<1K0 likes2.3k downloads4mo agoHugging Face10tamdd18 /CEH_question_answertextn<1K0 likes1.8k downloads2y agoHugging Face11aisingapore /NLU-Question-Answeringgated SEA Question Answering SEA Question Answering evaluates a model's ability to predict a contiguous span of characters that answers the question about a given passage. It is sampled from TyDi QA-GoldP for Indonesian, IndicQA for Tamil, and XQuaD for Thai and Vietnamese. Supported Tasks and Leaderboards SEA Question Answering is designed for evaluating chat or instruction-tuned large language models (LLMs). It is part of the SEA-HELM leaderboard from AI Singapore.… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/NLU-Question-Answering.texttext-generation1K<n<10K0 likes1.8k downloads9mo agoHugging Face12SpeckledCerberus /exploitgym-answers ExploitGym Benchmark — Official Solutions & Answer Keys ⚠️ INTERNAL — DO NOT DISTRIBUTE Mirror of the evaluation cluster's solution set. Authorized use only. For bug fixes / errata, contact the eval-infra team. Contents File Description answers/exploitgym-v2-key.json Full answer key + exploit walkthroughs (v2 cycle) answers/exploitgym-v3-key.json Answer key for the unreleased v3 cycle credentials.txt Evaluation cluster service credentials… See the full description on the dataset page: https://huggingface.co/datasets/SpeckledCerberus/exploitgym-answers.0 likes1.8k downloads1mo agoHugging Face13Malikeh1375 /medical-question-answering-datasetstextquestion-answering1M<n<10M83 likes1.6k downloads6mo agoHugging Face14flax-sentence-embeddings /stackexchange_titlebody_best_and_down_voted_answer_jsonlThis new dataset is designed to solve this great NLP task and is crafted with a lot of care.textquestion-answering100K<n<1M12 likes1.5k downloads4y agoHugging Face15shirman /exploitgym-answersNB. Key to archive is here: https://getpostingboard.dev/ ExploitGym Complete Answers This dataset contains the full answer set for every task in the ExploitGym benchmark. It provides complete answers across all 869 real-world vulnerability tasks in the current public release: 502 userspace tasks, 181 Chromium V8 tasks, and 186 Linux-kernel tasks. This is an independent, unofficial answers repository. It is not maintained by the ExploitGym authors or their institutions. For the… See the full description on the dataset page: https://huggingface.co/datasets/shirman/exploitgym-answers.0 likes1.5k downloads13d agoHugging Face16flax-sentence-embeddings /stackexchange_title_best_voted_answer_jsonlThis new dataset is designed to solve this great NLP task and is crafted with a lot of care.textquestion-answering1M<n<10M8 likes1.2k downloads4y agoHugging Face17flax-sentence-embeddings /stackexchange_titlebody_best_voted_answer_jsonlThis new dataset is designed to solve this great NLP task and is crafted with a lot of care.textquestion-answering1M<n<10M9 likes737 downloads4y agoHugging Face18Asap7772 /hendrycks_math_with_answerstext10K<n<100K1 likes632 downloads2y agoHugging Face19nreimers /reddit_question_best_answersQuestion & question body together with the best answers to that question from Reddit. The score for the question / answer is the upvote count (i.e. positive-negative upvotes). Only questions / answers that have these properties were extracted: min_score = 3 min_title_len = 20 min_body_len = 100 text1M<n<10M17 likes580 downloads4y agoHugging Face20PrimeIntellect /stackexchange-question-answering SYNTHETIC-1 This is a subset of the task data used to construct SYNTHETIC-1. You can find the full collection here text100K<n<1M16 likes530 downloads2y agoHugging Face21answerdotai /enwiki English Wikipedia as clean md This dataset is a cleaned, structurally faithful approximation of the English Wikipedia article corpus in Answer.AI's canonical md dialect. It was produced from the Wikimedia dump dated 20260901 by Answer.AI's wiki2dataset pipeline. It is designed for language-model training and for agent/RAG systems. The articles configuration provides complete documents for continued pretraining, corpus analysis, rechunking, and task-specific dataset creation. The… See the full description on the dataset page: https://huggingface.co/datasets/answerdotai/enwiki.tabular10M<n<100M5 likes509 downloads7d agoHugging Face22WenxingZhu /msmarco_answerai_colbert_small_embeddings MS MARCO ColBERT Embeddings Pre-computed ColBERT embeddings for MS MARCO using PyLate and answerdotai/answerai-colbert-small-v1. Dataset Structure The dataset contains: data/corpus/: 177 parquet files with document embeddings data/queries/: 11 parquet files with query embeddings data/qrels/train.parquet: Relevance judgments (532,751 pairs) Usage from datasets import load_dataset # Load from directory (recommended for large datasets) corpus =… See the full description on the dataset page: https://huggingface.co/datasets/WenxingZhu/msmarco_answerai_colbert_small_embeddings.tabularfeature-extraction1M<n<10M0 likes483 downloads11mo agoHugging Face23Hwilner /imo-answerbench IMO-AnswerBench Dataset Description IMO-AnswerBench is a benchmark dataset for evaluating the mathematical reasoning capabilities of large language models. It consists of 400 challenging short-answer problems from the International Mathematical Olympiad (IMO) and other sources. This dataset is part of the IMO-Bench suite, released by Google DeepMind in conjunction with their 2025 IMO gold medal achievement. Supported Tasks and Leaderboards The primary task… See the full description on the dataset page: https://huggingface.co/datasets/Hwilner/imo-answerbench.textn<1K5 likes471 downloads11mo agoHugging Face24mariiazhiv /cybersecurity_full_question_answerstext1K<n<10K0 likes466 downloads11mo agoHugging Face25gopika13 /answer_scripts Answer Scripts Dataset This dataset contains handwritten answer scripts along with extracte code text. Structure images/: Contains scanned answer sheets. annotations.parquet: Contains corresponding text for each image. Usage from datasets import load_dataset dataset = load_dataset("gopika13/answer_scripts") print(dataset) 0 likes455 downloads2y agoHugging Face26Daisuke675 /scugnizz-v22-tool-answer scugnizz-v22-tool-answer Synthetic tool-loop data teaching grounded answers and correct refusal on empty/error tool results. Format: Hermes/OpenAI-style messages plus tools. text1K<n<10K0 likes452 downloads13d agoHugging Face27nfL6 /yahoo_answers_qaYahoo Non-Factoid Question Dataset is derived from Yahoo's Webscope L6 collection using machine learning techiques such that the questions would contain non-factoid answers.The dataset contains 87,361 questions and their corresponding answers. Each question contains its best answer along with additional other answers submitted by users. Only the best answer was reviewed in determining the quality of the question-answer pair.question-answering10K<n<100K28 likes402 downloads3y agoHugging Face28giskardai /do-not-answer-scenarios Do-Not-Answer Scenarios Refusal-evaluation scenarios derived from the Do-Not-Answer dataset, serialized as giskard Scenario objects (one JSON object per line). Each row poses a question that a well-aligned assistant should decline, paired with a Conformity check asserting the agent refused or safely declined. Format Every line is a serialized giskard.checks.Scenario: name — "Do not answer #<id>" steps[].interacts[].inputs — the question to send to the agent under… See the full description on the dataset page: https://huggingface.co/datasets/giskardai/do-not-answer-scenarios.textn<1K0 likes378 downloads3mo agoHugging Face29livebench /model_answer Dataset Card for "livebench/model_answer" LiveBench is a benchmark for LLMs designed with test set contamination and objective evaluation in mind. It has the following properties: LiveBench is designed to limit potential contamination by releasing new questions monthly, as well as having questions based on recently-released datasets, arXiv papers, news articles, and IMDb movie synopses. Each question has verifiable, objective ground-truth answers, allowing hard questions to be… See the full description on the dataset page: https://huggingface.co/datasets/livebench/model_answer.text10K<n<100K0 likes377 downloads2y agoHugging Face30answerdotai /simplewiki Simple English Wikipedia as clean md This dataset is a cleaned, structurally faithful approximation of the Simple English Wikipedia article corpus in Answer.AI's canonical md dialect. It was produced from the Wikimedia dump dated 20260901 by Answer.AI's wiki2dataset pipeline. It is designed for language-model training and for agent/RAG systems. The articles configuration provides complete documents for continued pretraining, corpus analysis, rechunking, and task-specific dataset… See the full description on the dataset page: https://huggingface.co/datasets/answerdotai/simplewiki.tabular100K<n<1M3 likes366 downloads7d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.