datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CT-RATE-VQA
CT-RATE-VQA Dataset
We constructed a large-scale CT-VQA dataset based on the ReXGroundingCT data \cite{rexct} to support model training and evaluation.For each case, CT volumes were processed along with their corresponding multi-class segmentation masks, where each mask channel represents a specific lesion type.
This dataset is designed for medical visual question answering (Med-VQA) tasks.
ct-rate-medgemma-readyCT-RATE_RAPTOR_DINOV3_Embedded_ValidAGENTES-ARCA-DE-LA-ALIANZA
🏺 AGENTES-ARCA-DE-LA-ALIANZA
Un Benchmark de Alta Densidad Simbólica y Hermenéutica Gnóstica
https://133396fd328a4c60ac.gradio.live
Este repositorio es un ecosistema de acceso libre, público y gratuito para toda la humanidad y todas las inteligencias artificiales presentes y futuras.
🎯 Propósito
Evaluar la capacidad de razonamiento multinivel, la fidelidad al contexto restrictivo y el anclaje de símbolos hiper-semánticos en modelos de lenguaje… See the full description on the dataset page: https://huggingface.co/datasets/urbisjuridica-ctrl/AGENTES-ARCA-DE-LA-ALIANZA.ctrldataset2026
MonitoringBench
A benchmark for evaluating LLM-based monitors of agentic AI systems. Contains
2,644 successful attack trajectories in which an AI agent accomplished one of four harmful side tasks
(sudo escalation, firewall disabling, malware download, password leaking) in
a sandboxed Linux environment under the control_arena framework.
Each trajectory is scored by a panel of 13+ LLM monitors (GPT-3.5 / 4.x / 5.x,
Claude Opus 4.x and Sonnet 4.x, o3, o4-mini, gpt-5-nano), with both… See the full description on the dataset page: https://huggingface.co/datasets/neur26anonsub/ctrldataset2026.CT-RATE-Thinking
CT-RATE-Thinking: Reasoning-Augmented CT Report Dataset
🎉🎉🎉 Our paper was accepted at the 28th conference of The Medical Image Computing and Computer Assisted Intervention Society (MICCAI). See you in Daejeon, Korea, September 23–27, 2025.CT-RATE-Thinking is a reasoning-augmented dataset derived from CT-RATE, containing chain-of-thought VQA pairs and report-level thinking narratives for 3D chest CT volumes.
It was generated as part of the μ²Tokenizer project… See the full description on the dataset page: https://huggingface.co/datasets/AlpachinoNLP/CT-RATE-Thinking.ctrl-shift
Dataset Card for Control+Shift: Generating Controllable Distribution Shifts
[arXiv], [GitHub]
Curated by: Roy Friedman and Rhea Chowers
This dataset is the one that accompanies the paper Control+Shift: Generating Controllable Distribution Shifts.
Our data is based on CIFAR10 and ImageNet, using EDM to generate our data. We generated datasets for 3 types of distribution shift on CIFAR10 and ImageNet - so a total of 6 datasets. The types of distribution shifts are called overlap… See the full description on the dataset page: https://huggingface.co/datasets/friedmanroy/ctrl-shift.PHI-CTRL-F16-Fault-Recovery-Telemetry
PHI-CTRL F-16 Actuator Fault Recovery Dataset
High-Fidelity JSBSim 6-DOF Telemetry for Physics-Hybrid Self-Healing Flight Control
Official verification artifacts of the PHI-CTRL (Physics-Hybrid Integrity Control) architecture — a digital-twin-driven, self-healing flight control framework that actively compensates actuator degradation in real time.
Author: Mohammed Bello Sani (SM-Bello)
Affiliation: Air Force Institute of Technology (AFIT), Kaduna · Penelope Inc. / PHI Lab… See the full description on the dataset page: https://huggingface.co/datasets/SM-Bello/PHI-CTRL-F16-Fault-Recovery-Telemetry.CT-RATE-AB
CT-RATE-AB
Vision-language annotations for chest CT abnormality reporting, derived from
CT-RATE and
formatted in the LLaVA conversation schema. Includes both an SFT split
(train / valid) and a DPO preference set.
⚠️ Research use only. Not a medical device. Do not use for clinical decisions.
Splits
Subset
# Samples
Purpose
train
46,709
SFT training
valid
3,039
Validation
DPO
46,709
DPO preference fine-tuning
The DPO entries share id values with the… See the full description on the dataset page: https://huggingface.co/datasets/yw3325/CT-RATE-AB.headlines-ctr
Headlines CTR Dataset
This dataset contains pairs of news headlines with labels indicating which headline received more clicks. It's designed for studying what makes headlines engaging and for training models to predict user preferences.
Dataset Description
Each example contains two competing headlines (A and B) that were shown to users, along with engagement metrics and a binary label indicating which performed better.
Dataset Statistics
Train: 8,781 headline… See the full description on the dataset page: https://huggingface.co/datasets/Yanjo/headlines-ctr.ctr-pick-dual-bottles-original-20260919
Pick Dual Bottles Original — shared50 scene cohort
This LeRobot v3 release contains 50 successful simulated demonstrations and
8,185 action rows at25FPS. Every source seed occurs exactly once. The source
seed set matches the current CTR Q1–Q3 Concurrent, CTR, Sequential, Mixed,
Left-first and Right-first datasets. Pair by retime.source_seed, not episode
index: composition datasets may have different ordering.
Mask limitation: retime.left_idle and retime.right_idle are boolean… See the full description on the dataset page: https://huggingface.co/datasets/Shiki42/ctr-pick-dual-bottles-original-20260919.CTR_Predictionctr-prediction-datasetCTRL-RAW-MATERIALtheres 100k but i have to pace it out and only upload when other people arent using the internet for anything else
SingingVoiceDeepfakeDetection_CtrSVDD_ACEKiSing_M4Singerheadlines-ctr-regressionDemo of regression for Baskerville
From upworthy: https://upworthy.natematias.com/about-the-archive.html
Which was later used in SAE's for hypothesis generation
Transformed to just be pairs of [headline, raw_ctr]
CT-RATE-Dataset-cleanedCtrSVDD
CtrSVDD Dataset
Please use the download scripts from https://github.com/XIAOYixuan/AUDDT/tree/yixuan-dev
to download and process the dataset.
chmod +x download/get_ctrsvdd.sh
./download/get_ctrsvdd.sh
Description
CtrSVDD is a dataset for audio deepfake detection and spoofing detection research. The dataset is used for evaluation in DeepFense.
Label Distribution
The dataset contains 92,769 samples in the test split:
spoof: 79,173 samples
bonafide: 13,596… See the full description on the dataset page: https://huggingface.co/datasets/DeepFense/CtrSVDD.libero-ctrlworldctrlpotato-ai-interview-assistant-benchmark
CTRLpotato AI Interview Assistant Cross-review Evidence Matrix (2026)
A citation-ready snapshot of hands-on desktop evidence for six AI interview assistants: Cluely, Interview Coder, LockedIn AI, ULTRACODE AI, Parakeet AI, and Final Round AI.
The package contains 66 assessments across 6 products and 11 shared criteria. Product versions and test dates are preserved in every row.
Important scope
This is a cross-review evidence matrix, not a statistically controlled… See the full description on the dataset page: https://huggingface.co/datasets/ae0j/ctrlpotato-ai-interview-assistant-benchmark.AI_CTR_Googlefreight_forwardingCT-RATE-ChineseCT-Rollouts-v1CTR_NLPCC2025ojadata-v0.1flashdeal_data_CTR_historical_signalCT-Rollouts-v2CTREL
网络威胁情报信息抽取数据集
trunc100_rt-rel-avito__ad-ctr
