CoolFace
21 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01SetFit /ethos_binaryThis is the binary split of ethos, split into train and test. It contains comments annotated for hate speech or not. textn<1K1 likes230 downloads5y agoHugging Face02tuhink /cambench_binary_eval CameraBench Binary Evaluation Dataset A balanced VQA dataset for evaluating camera motion understanding in videos. 📊 Dataset Statistics Total Questions: 384 Unique Videos: 119 Unique Questions: 31 Yes Answers: 192 (50.0%) No Answers: 192 (50.0%) Balance Ratio: 1.00 Total Size: 126.16 MB (0.12 GB) Average Video Size: 1.06 MB 🎯 Task Categories This dataset covers various camera motion tasks including: Static: 42 questions Move In: 29 questions Pan Left: 24… See the full description on the dataset page: https://huggingface.co/datasets/tuhink/cambench_binary_eval.imagevisual-question-answeringn<1K0 likes151 downloads11mo agoHugging Face03graphs-datasets /IMDB-BINARY Dataset Card for IMDB-BINARY (IMDb-B) Dataset Summary The IMDb-B dataset is "a movie collaboration dataset that consists of the ego-networks of 1,000 actors/actresses who played roles in movies in IMDB. In each graph, nodes represent actors/actress, and there is an edge between them if they appear in the same movie. These graphs are derived from the Action and Romance genres". Supported Tasks and Leaderboards IMDb-B should be used for graph classification… See the full description on the dataset page: https://huggingface.co/datasets/graphs-datasets/IMDB-BINARY.graph-ml1K<n<10K1 likes74 downloads4y agoHugging Face04Labradorlabs /bsca-binary-source-gold-v3-multidomain BSCA Gold v3 Multidomain Address-grounded P1 pairs for stripped pseudo-C → source retrieval. Dataset ID: GD_19330e06aae0462447c1fd05ccaa38d7 Accepted P1 pairs: 42449 Repositories: 138 Target formats: {"elf": 40733, "pe": 1716} Target architectures: {"aarch64": 1672, "x86": 1903, "x86_64": 38874} Internal quality GPA: 3.660; target pass: True Use train.jsonl for fitting, development.jsonl for model selection, and the immutable test.jsonl only after selection. dataset_card.json… See the full description on the dataset page: https://huggingface.co/datasets/Labradorlabs/bsca-binary-source-gold-v3-multidomain.tabularfeature-extraction10K<n<100K0 likes46 downloads1mo agoHugging Face05bingbangboom /editlens_iclr_binary_reasoning bingbangboom/editlens_iclr_binary_reasoning This dataset is a binary-classification subset drawn from the training split of pangram/editlens_iclr dataset. It isolates purely human-crafted texts (human_written) against purely synthetic content (ai_generated), strictly filtering out the overlapping ai_edited classification cluster for binary classification tasks. The primary augmentation of this dataset is the inclusion of Reasoning Traces (Chain of Thought). Every single text… See the full description on the dataset page: https://huggingface.co/datasets/bingbangboom/editlens_iclr_binary_reasoning.texttext-classification1K<n<10K0 likes35 downloads5mo agoHugging Face06christinacdl /binary_hate_speechtexttext-classification10K<n<100K0 likes28 downloads3y agoHugging Face07jijivski /metaculus_binarytextn<1K0 likes25 downloads3y agoHugging Face08lucy3 /aftermath_binary_correctness Aftermath of DrawEduMath This contains binary_correctness.json, for recreating the results of the paper titled "The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors". This file includes outputs from GPT-5-mini labeling whether student is correct/incorrect on binary error & correctness questions, from DrawEduMath. Please consult the datacard for DrawEduMath for detailed information about data source. Quick links:… See the full description on the dataset page: https://huggingface.co/datasets/lucy3/aftermath_binary_correctness.textn<1K0 likes20 downloads7mo agoHugging Face09SoulInPsyAbstract /specialist-cd-binary-honestytextn<1K0 likes19 downloads1mo agoHugging Face10SotirisLegkas /binary_off_hate_toxictext10K<n<100K1 likes16 downloads3y agoHugging Face11trentmkelly /authorship-attribution-binaryBinary classification dataset for authorship attribution. Each row contains two sets of messages, each set containing at least 250 characters, and a label indicating if the two sets of messages were written by the same author or by two different authors. Three sources are used: reddit comments, discord messages, and blog posts from the Blog Authorship Corpus (license unknown). Discord messages: 291,648 pairs Reddit comments: 190,308 pairs Blog Authorship Corpus: 121,350 pairs Train set: 573… See the full description on the dataset page: https://huggingface.co/datasets/trentmkelly/authorship-attribution-binary.text100K<n<1M0 likes16 downloads1y agoHugging Face12PJMixers /nvidia_HelpSteer2-Correctness-Binary-ClassificationCorrectness == 4/4 = 1 Correctness < 4/4 = 0 text10K<n<100K0 likes14 downloads2y agoHugging Face13NJU-LINK /camerabench_binary 示例条目展示 以下是如何读取 camerabench_binary.jsonl 文件中的前100个条目并进行展示的示例代码: import jsonlines # 假设文件已经在当前工作目录 filename = "camerabench_binary.jsonl" # 读取并展示前100个条目 with jsonlines.open(filename) as reader: for i, obj in enumerate(reader): print(f"条目 {i+1}: {obj}") if i >= 99: # 只展示前100个条目break text1K<n<10K0 likes13 downloads8mo agoHugging Face14SotirisLegkas /binary_off_hate_toxic_newtext10K<n<100K1 likes12 downloads3y agoHugging Face15Binarybardakshat /SVLM-ACL-DATASETtext1K<n<10K0 likes12 downloads2y agoHugging Face16lucy3 /aftermath_question_binary Aftermath of DrawEduMath This contains question_binary.json, for recreating the results of the paper titled "The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors". This file outputs from GPT-5-mini labeling whether an error & correctness question is "binary" (e.g. "Does the student do ___ correctly?") or "other" (e.g. "What incorrect product did the student calculate for 667 times 5?"). Please consult the datacard for… See the full description on the dataset page: https://huggingface.co/datasets/lucy3/aftermath_question_binary.textn<1K0 likes10 downloads7mo agoHugging Face17jowaov /RE-KERNEL-BINARYtext1M<n<10M0 likes6 downloads7mo agoHugging Face18FibonacciNeu /binary_zero_shottext10K<n<100K0 likes5 downloads1y agoHugging Face19Mao940534844 /IMDB-BINARY Dataset Card for IMDB-BINARY (IMDb-B) Dataset Summary The IMDb-B dataset is "a movie collaboration dataset that consists of the ego-networks of 1,000 actors/actresses who played roles in movies in IMDB. In each graph, nodes represent actors/actress, and there is an edge between them if they appear in the same movie. These graphs are derived from the Action and Romance genres". Supported Tasks and Leaderboards IMDb-B should be used for graph classification… See the full description on the dataset page: https://huggingface.co/datasets/Mao940534844/IMDB-BINARY.graph-ml1K<n<10K0 likes5 downloads5mo agoHugging Face20FibonacciNeu /binary_one_shottext10K<n<100K0 likes4 downloads1y agoHugging Face21chentong00 /binary-rar-wildchat-8ktext1K<n<10K0 likes4 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.