CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01liyf001 /CT-RATE-VQA CT-RATE-VQA Dataset We constructed a large-scale CT-VQA dataset based on the ReXGroundingCT data \cite{rexct} to support model training and evaluation.For each case, CT volumes were processed along with their corresponding multi-class segmentation masks, where each mask channel represents a specific lesion type. This dataset is designed for medical visual question answering (Med-VQA) tasks. image10K<n<100K1 likes1.2k downloads1y agoHugging Face02rimine /ct-rate-medgemma-readytext1K<n<10K0 likes855 downloads4mo agoHugging Face03PwwSeniorProj /CT-RATE_RAPTOR_DINOV3_Embedded_Validtext0 likes461 downloads9mo agoHugging Face04urbisjuridica-ctrl /AGENTES-ARCA-DE-LA-ALIANZA 🏺 AGENTES-ARCA-DE-LA-ALIANZA Un Benchmark de Alta Densidad Simbólica y Hermenéutica Gnóstica https://133396fd328a4c60ac.gradio.live Este repositorio es un ecosistema de acceso libre, público y gratuito para toda la humanidad y todas las inteligencias artificiales presentes y futuras. 🎯 Propósito Evaluar la capacidad de razonamiento multinivel, la fidelidad al contexto restrictivo y el anclaje de símbolos hiper-semánticos en modelos de lenguaje… See the full description on the dataset page: https://huggingface.co/datasets/urbisjuridica-ctrl/AGENTES-ARCA-DE-LA-ALIANZA.texttext-generationn<1K0 likes261 downloads14h agoHugging Face05neur26anonsub /ctrldataset2026 MonitoringBench A benchmark for evaluating LLM-based monitors of agentic AI systems. Contains 2,644 successful attack trajectories in which an AI agent accomplished one of four harmful side tasks (sudo escalation, firewall disabling, malware download, password leaking) in a sandboxed Linux environment under the control_arena framework. Each trajectory is scored by a panel of 13+ LLM monitors (GPT-3.5 / 4.x / 5.x, Claude Opus 4.x and Sonnet 4.x, o3, o4-mini, gpt-5-nano), with both… See the full description on the dataset page: https://huggingface.co/datasets/neur26anonsub/ctrldataset2026.tabulartext-classification1K<n<10K0 likes244 downloads5mo agoHugging Face06AlpachinoNLP /CT-RATE-Thinking CT-RATE-Thinking: Reasoning-Augmented CT Report Dataset 🎉🎉🎉 Our paper was accepted at the 28th conference of The Medical Image Computing and Computer Assisted Intervention Society (MICCAI). See you in Daejeon, Korea, September 23–27, 2025.CT-RATE-Thinking is a reasoning-augmented dataset derived from CT-RATE, containing chain-of-thought VQA pairs and report-level thinking narratives for 3D chest CT volumes. It was generated as part of the μ²Tokenizer project… See the full description on the dataset page: https://huggingface.co/datasets/AlpachinoNLP/CT-RATE-Thinking.textvisual-question-answering1M<n<10M2 likes164 downloads5mo agoHugging Face07friedmanroy /ctrl-shift Dataset Card for Control+Shift: Generating Controllable Distribution Shifts [arXiv], [GitHub] Curated by: Roy Friedman and Rhea Chowers This dataset is the one that accompanies the paper Control+Shift: Generating Controllable Distribution Shifts. Our data is based on CIFAR10 and ImageNet, using EDM to generate our data. We generated datasets for 3 types of distribution shift on CIFAR10 and ImageNet - so a total of 6 datasets. The types of distribution shifts are called overlap… See the full description on the dataset page: https://huggingface.co/datasets/friedmanroy/ctrl-shift.imageimage-classification100K<n<1M1 likes152 downloads2y agoHugging Face08SM-Bello /PHI-CTRL-F16-Fault-Recovery-Telemetry PHI-CTRL F-16 Actuator Fault Recovery Dataset High-Fidelity JSBSim 6-DOF Telemetry for Physics-Hybrid Self-Healing Flight Control Official verification artifacts of the PHI-CTRL (Physics-Hybrid Integrity Control) architecture — a digital-twin-driven, self-healing flight control framework that actively compensates actuator degradation in real time. Author: Mohammed Bello Sani (SM-Bello) Affiliation: Air Force Institute of Technology (AFIT), Kaduna · Penelope Inc. / PHI Lab… See the full description on the dataset page: https://huggingface.co/datasets/SM-Bello/PHI-CTRL-F16-Fault-Recovery-Telemetry.tabulartime-series-forecasting10K<n<100K0 likes118 downloads18d agoHugging Face09yw3325 /CT-RATE-AB CT-RATE-AB Vision-language annotations for chest CT abnormality reporting, derived from CT-RATE and formatted in the LLaVA conversation schema. Includes both an SFT split (train / valid) and a DPO preference set. ⚠️ Research use only. Not a medical device. Do not use for clinical decisions. Splits Subset # Samples Purpose train 46,709 SFT training valid 3,039 Validation DPO 46,709 DPO preference fine-tuning The DPO entries share id values with the… See the full description on the dataset page: https://huggingface.co/datasets/yw3325/CT-RATE-AB.textvisual-question-answering10K<n<100K0 likes101 downloads4mo agoHugging Face10Yanjo /headlines-ctr Headlines CTR Dataset This dataset contains pairs of news headlines with labels indicating which headline received more clicks. It's designed for studying what makes headlines engaging and for training models to predict user preferences. Dataset Description Each example contains two competing headlines (A and B) that were shown to users, along with engagement metrics and a binary label indicating which performed better. Dataset Statistics Train: 8,781 headline… See the full description on the dataset page: https://huggingface.co/datasets/Yanjo/headlines-ctr.tabulartext-classification10K<n<100K0 likes97 downloads1y agoHugging Face11Shiki42 /ctr-pick-dual-bottles-original-20260919 Pick Dual Bottles Original — shared50 scene cohort This LeRobot v3 release contains 50 successful simulated demonstrations and 8,185 action rows at25FPS. Every source seed occurs exactly once. The source seed set matches the current CTR Q1–Q3 Concurrent, CTR, Sequential, Mixed, Left-first and Right-first datasets. Pair by retime.source_seed, not episode index: composition datasets may have different ordering. Mask limitation: retime.left_idle and retime.right_idle are boolean… See the full description on the dataset page: https://huggingface.co/datasets/Shiki42/ctr-pick-dual-bottles-original-20260919.tabularn<1K0 likes86 downloads7d agoHugging Face12BadrAbu /CTR_Predictiontabular10M<n<100M5 likes80 downloads2y agoHugging Face13EvgeniaKyriazi /ctr-prediction-datasettabular1M<n<10M4 likes79 downloads11mo agoHugging Face14crumbs-playground /CTRL-RAW-MATERIALtheres 100k but i have to pace it out and only upload when other people arent using the internet for anything else tabular1K<n<10K0 likes35 downloads6mo agoHugging Face15DynamicSuperb /SingingVoiceDeepfakeDetection_CtrSVDD_ACEKiSing_M4Singeraudion<1K1 likes28 downloads2y agoHugging Face16yanjo-gutenberg /headlines-ctr-regressionDemo of regression for Baskerville From upworthy: https://upworthy.natematias.com/about-the-archive.html Which was later used in SAE's for hypothesis generation Transformed to just be pairs of [headline, raw_ctr] tabular10K<n<100K0 likes28 downloads1y agoHugging Face17dharmam-stjude /CT-RATE-Dataset-cleanedtabular10K<n<100K0 likes26 downloads10mo agoHugging Face18DeepFense /CtrSVDD CtrSVDD Dataset Please use the download scripts from https://github.com/XIAOYixuan/AUDDT/tree/yixuan-dev to download and process the dataset. chmod +x download/get_ctrsvdd.sh ./download/get_ctrsvdd.sh Description CtrSVDD is a dataset for audio deepfake detection and spoofing detection research. The dataset is used for evaluation in DeepFense. Label Distribution The dataset contains 92,769 samples in the test split: spoof: 79,173 samples bonafide: 13,596… See the full description on the dataset page: https://huggingface.co/datasets/DeepFense/CtrSVDD.text10K<n<100K0 likes21 downloads9mo agoHugging Face19Ricardo520nono /libero-ctrlworldtabular1K<n<10K0 likes21 downloads6mo agoHugging Face20ae0j /ctrlpotato-ai-interview-assistant-benchmark CTRLpotato AI Interview Assistant Cross-review Evidence Matrix (2026) A citation-ready snapshot of hands-on desktop evidence for six AI interview assistants: Cluely, Interview Coder, LockedIn AI, ULTRACODE AI, Parakeet AI, and Final Round AI. The package contains 66 assessments across 6 products and 11 shared criteria. Product versions and test dates are preserved in every row. Important scope This is a cross-review evidence matrix, not a statistically controlled… See the full description on the dataset page: https://huggingface.co/datasets/ae0j/ctrlpotato-ai-interview-assistant-benchmark.textn<1K0 likes19 downloads1mo agoHugging Face21Ashneon /AI_CTR_Googletabular10M<n<100M0 likes18 downloads2y agoHugging Face22CtrlAltDEviL /freight_forwardingtabular100K<n<1M1 likes17 downloads1y agoHugging Face23AlpachinoNLP /CT-RATE-Chinesetext10K<n<100K0 likes15 downloads1y agoHugging Face24NewEden /CT-Rollouts-v1tabular10K<n<100K0 likes15 downloads8mo agoHugging Face25XiaoEnn /CTR_NLPCC2025text10K<n<100K0 likes13 downloads1y agoHugging Face26ctrlprompt /ojadata-v0.1textn<1K0 likes13 downloads2mo agoHugging Face27minhthong /flashdeal_data_CTR_historical_signaltabularn<1K0 likes11 downloads2y agoHugging Face28NewEden /CT-Rollouts-v2tabular10K<n<100K0 likes11 downloads8mo agoHugging Face29shktty /CTREL 网络威胁情报信息抽取数据集 text1K<n<10K0 likes10 downloads2y agoHugging Face30ma-zn /trunc100_rt-rel-avito__ad-ctrtext10K<n<100K0 likes9 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.