datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ai-detector-data
AI Detector Predictions Dataset
A continuously-growing collection of AI text detection predictions with optional user feedback, generated from the AI Text Detector Space.
Every time someone analyzes text or a URL on the Space, the prediction is appended to this dataset. Users can also click "Correct" or "Incorrect" to provide feedback, which gets stored alongside the prediction.
Schema
Field
Type
Description
id
string
Unique 12-char hex identifier… See the full description on the dataset page: https://huggingface.co/datasets/adaptive-classifier/ai-detector-data.adaptive-adversaries-data
Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security
A 21-scenario multi-turn (15-round) adversarial red-teaming benchmark for LLM agents, in which both attacker and defender are independent LLM agents and attacks are regenerated per battle. Includes calibrated 3×3 attacker × defender matrix evaluation, full battle transcripts, attack-replay corpus, and traces from two open AgentBeats competitions.
Companion paper: Adaptive Adversaries: A Multi-Turn… See the full description on the dataset page: https://huggingface.co/datasets/neurips-adaptive-adversaries/adaptive-adversaries-data.han-adaptive-environmental-interaction-dataset-v1
Humanoid Adaptive Environmental Interaction Dataset
This dataset captures environmental perception
and adaptive behavioral responses
of humanoid agents operating
in dynamic real-world conditions.
It includes sensor inputs,
obstacle interaction logs,
and adaptive strategy outcomes.
Objective
To train humanoid agents
to respond intelligently
to unpredictable environments.
Data Fields
agent_id
environment_type
sensor_input_vector
detected_obstacle… See the full description on the dataset page: https://huggingface.co/datasets/achiepatricia/han-adaptive-environmental-interaction-dataset-v1.adaptive-clinical-decision-datasethan-adaptive-behavior-feedback-dataset-v1
Humanoid Adaptive Behavior Feedback Dataset
This dataset captures behavioral feedback loops
between humans and humanoid agents.
It includes corrective feedback, reinforcement signals,
and behavior adjustment outcomes.
Purpose
To enable continuous learning and refinement
of humanoid behavior through structured feedback.
Data Fields
initial_behavior
human_feedback
correction_type
adjustment_strategy
final_behavior
performance_score
Use Cases… See the full description on the dataset page: https://huggingface.co/datasets/achiepatricia/han-adaptive-behavior-feedback-dataset-v1.han-adaptive-governance-policy-dataset-v1
Humanoid Adaptive Governance Policy Dataset
This dataset represents dynamic governance policies
applied to humanoid agents
in decentralized operational environments.
It includes rule evolution,
compliance scoring,
and automated policy adjustment logs.
Objective
To enable adaptive rule enforcement
and governance automation
in distributed humanoid systems.
Data Fields
policy_id
policy_version
rule_set_signature
compliance_score
violation_event_count… See the full description on the dataset page: https://huggingface.co/datasets/achiepatricia/han-adaptive-governance-policy-dataset-v1.humanoid-adaptive-evolution-datasetadaptive_datasetadaption-biosecure-adaptive-dataset
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
biosecure_adaptive_dataset
This dataset contains structured entries for Dual-Use Research of Concern (DURC), featuring biosafety risk scores, regulatory frameworks, and organism details. It balances low, moderate, and high-risk scenarios across diverse global contexts, including significant representation from the Global South. Each entry adheres to a strict schema that categorizes experiments by… See the full description on the dataset page: https://huggingface.co/datasets/joduor/adaption-biosecure-adaptive-dataset.adaption-biosecure-adaptive-dataset-v2
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
biosecure_adaptive_dataset
This dataset provides globally adaptive classifications for Dual-Use Research of Concern (DURC), including biosafety risk scoring and regulatory mapping. Each entry details specific biological research scenarios, identifying the organism, methodology, risk level, required biosafety containment, and applicable international or national policies. The data supports risk… See the full description on the dataset page: https://huggingface.co/datasets/joduor/adaption-biosecure-adaptive-dataset-v2.adaption-biosecure-adaptive-dataset-v1
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
biosecure_adaptive_dataset
This dataset contains synthetic research scenarios designed to classify Dual-Use Research of Concern (DURC) and assign biosafety risk levels across diverse global contexts. Each entry follows a strict JSON schema detailing organism modifications, experiment types, risk scores, and applicable regulatory frameworks. The collection balances high-risk pathogen studies with… See the full description on the dataset page: https://huggingface.co/datasets/joduor/adaption-biosecure-adaptive-dataset-v1.
