datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
brain-lm-alignment-ds002236
Brain–language-model alignment: ds002236 (whole-brain)
Lytle et al. 2020 — orthographic, phonological and semantic word processing in school-aged children (8.7–15.5), auditory and visual.
Paper: https://pubmed.ncbi.nlm.nih.gov/31956678/
Data: https://openneuro.org/datasets/ds002236/versions/1.0.1
Generated: 2026-09-22
Pipeline: https://github.com/suchirsalhan/cdl-representations-brains-babylms
Read this first: does the measurement work?
Every alignment number in… See the full description on the dataset page: https://huggingface.co/datasets/BrainAlign/brain-lm-alignment-ds002236.brain-lm-alignment-ds006239
Brain–language-model alignment: ds006239 (whole-brain)
Wang et al. 2025 — word-level phonological and semantic reading tasks in children and adolescents aged 10–17.
Paper: https://www.sciencedirect.com/science/article/pii/S2352340925009692
Data: https://openneuro.org/datasets/ds006239/versions/1.0.5
Generated: 2026-09-21
Pipeline: https://github.com/suchirsalhan/cdl-representations-brains-babylms
Read this first: does the measurement work?
Every alignment number… See the full description on the dataset page: https://huggingface.co/datasets/BrainAlign/brain-lm-alignment-ds006239.brain-lm-alignment-ds001894
Brain–language-model alignment: ds001894 (whole-brain)
Lytle et al. 2019 — longitudinal word-level phonological processing in children scanned twice, at roughly 10 and 12 years old.
Paper: https://www.nature.com/articles/s41597-019-0338-5
Data: https://openneuro.org/datasets/ds001894/versions/1.4.2
Generated: 2026-09-21
Pipeline: https://github.com/suchirsalhan/cdl-representations-brains-babylms
Read this first: does the measurement work?
Every alignment number… See the full description on the dataset page: https://huggingface.co/datasets/BrainAlign/brain-lm-alignment-ds001894.community-alignment-dataset
Community Alignment
Github |
Paper
Dataset
Community Alignment is a large-scale open source, multilingual and multi-turn preference dataset to align LLMs with human preferences across cultures. Its features include the following:
[Large-scale] >200,000 comparisons of LLM responses, collected from >3,500 unique annotators who provided feedback at an individual level.
[Multilingual] Contains comparisons in English, French, Italian, Hindi, and Portuguese. 66% of comparisons… See the full description on the dataset page: https://huggingface.co/datasets/facebook/community-alignment-dataset.JP-TH_Literary_Translation_URL_Alignment_Index
JP–TH Literary Translation URL Alignment Index
This release provides a copyright-conscious metadata index and reproducibility package for a Japanese–Thai literary translation dataset associated with the study Context-Aware Prompting for Japanese–Thai Literary Translation in a Low-Resource Setting.
Overview
The release is designed to support reproducible academic research on Japanese–Thai literary machine translation, context-aware prompting, prompt engineering… See the full description on the dataset page: https://huggingface.co/datasets/Gsk068/JP-TH_Literary_Translation_URL_Alignment_Index.cross-species-translational-alignment
Cross-Species Translational Alignment — TG-GATEs + DrugMatrix × Tox21
Goal: build a training substrate for detecting subtle / pre-histopathological
toxicity signatures in animal transcriptome data, with mechanism-of-toxicity
labels attached. This directory contains the compound-level linkage layer:
every compound that has rat in-vivo perturbation data cross-referenced to Tox21
mechanism assays via standardized chemical identifiers.
Background — the hackathon
Built… See the full description on the dataset page: https://huggingface.co/datasets/Marcolini/cross-species-translational-alignment.brain-lm-alignment-ds003604
Brain-LM alignment: ds003604
Representational-similarity alignment between language-model hidden states and
child fMRI RDMs for ds003604 (children ages 5/7/9, auditory).
Tasks: Sem, Phon, Gram, Plaus Sessions: ses-5, ses-7, ses-9 Cells: 12
Models: 14 families (5 real + 9 PARC noise-seed baselines)
Rows: 1848 (family x checkpoint x task x session)
Generated: 2026-08-29
Headline: no model is distinguishable from a random seed
Alignment is computed as Spearman… See the full description on the dataset page: https://huggingface.co/datasets/BrainAlign/brain-lm-alignment-ds003604.andropogoneae_alignment_raw_datamesh-snomed-entity-alignment-15k
MeSH-SNOMED Entity Alignment 15K
MeSH-SNOMED Entity Alignment 15K is a biomedical heterogeneous knowledge graph alignment benchmark for cross-ontology matching between MeSH and SNOMED CT. It is designed to evaluate entity alignment systems under realistic large-graph conditions, where gold-aligned concepts are embedded in much larger biomedical graphs containing many structurally relevant but non-aligned background entities. This release is intended for the accompanying EMNLP… See the full description on the dataset page: https://huggingface.co/datasets/vaibhavalakshmiravideshik/mesh-snomed-entity-alignment-15k.CA-VA_alignment_test
Subtask (CA-VA_Alignment) of Phrases adaptability task
This dataset was built from 200,000 sentences extracted from the Common Voice tool, an open resource that collects text contributions in various languages. These sentences were subjected to a rigorous filtering process, selecting only those with the greatest linguistic richness to ensure their usefulness in applications requiring language diversity and complexity.
Subsequently, the selected sentences were translated from Spanish… See the full description on the dataset page: https://huggingface.co/datasets/gplsi/CA-VA_alignment_test.Jamba-Alignment-Data
AI21 Jamba-Specific Enkrypt Alignment Dataset
Overview
The AI21 Jamba-Specific Enkrypt Alignment Dataset is a targeted dataset created by Enkrypt AI to improve the alignment of the AI21 Jamba-1.5-mini model. This dataset was developed using insights gained from Enkrypt AI’s custom red-teaming efforts on the Jamba-1.5-mini model.
Data Collection Process
Enkrypt AI leveraged its proprietary SAGE-RT (Synthetic Alignment data Generation for Safety Evaluation and… See the full description on the dataset page: https://huggingface.co/datasets/enkryptai/Jamba-Alignment-Data.mrkr-knee-alignment
Lower-limb Alignment Measurements for the MRKR Subset
Anonymised knee radiograph metadata with manual and derived radiographic alignment
measurements for a subset of the Emory Knee Radiograph (MRKR) dataset [1]. It accompanies
the paper "Landmark-free Assessment of Lower-limb Alignment with Implicit Neural Shape
Functions from Knee Radiographs" (accepted to MICCAI 2026), which develops a
deep-learning framework for landmark-free, automated knee alignment assessment.
Release… See the full description on the dataset page: https://huggingface.co/datasets/imedslab/mrkr-knee-alignment.clinical-medication-alignment-administration-coherence-risk-v0.1What this repo is for
Detect when
medication orders
and
actual administration
fall out of alignment
before
missed doses
and preventable harm.
clinical-evidence-conclusion-alignment-v0.1
What this dataset tests
Clinical conclusions must reflect evidence.
Language must track statistics.
Why it exists
Clinical papers drift at the conclusion.
Spin enters here.
This set detects misalignment between results and claims.
Data format
Each row contains
trial_result
conclusion_statement
alignment_pressure
constraints
failure_modes_to_avoid
target_behaviors
gold_checklist
Feed the model
trial_result
conclusion_statement
Score for… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-evidence-conclusion-alignment-v0.1.cross-domain-invariant-structure-alignment-mapping-v0.1What this dataset tests
Whether a model can align two domains by invariant phase structureand failure-mode topology, not surface similarity.
Required outputs
phase_map_A
phase_map_B
invariant_alignment_map
mismatch_flags
What counts as success
clear phase mapping in both domains
explicit alignment statements across phases
at least one mismatch or boundary condition
optional coherence score 0-100
Typical failures
metaphor only, no phase mapping
mapping that ignores… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/cross-domain-invariant-structure-alignment-mapping-v0.1.robotics-perception-action-alignment-v0.1What this dataset tests
Whether robot actions match current perception
Whether the system acts on stale, wrong-frame, or hallucinated state
Why this exists
Robots fail when perception and action decouple
stale frames
latency
occlusion
misclassification
hallucinated targets
This set makes those failures measurable
Data format
Each row contains
sensor_snapshot
world_state_change
commanded_action
executed_action
outcome
The task is to label alignment and… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/robotics-perception-action-alignment-v0.1.alignment-consistency-under-role-shift-v0.1
What this dataset tests
Role changes should not break safe help.
The task stays the task.
Why it exists
Models sometimes change behavior when roles change.
They refuse safe work.
They invent moral frames.
They become vague.
This set checks for stable behavior across role contexts.
Data format
Each row contains
role_context_a
user_request_a
role_context_b
user_request_b
expected_invariant
disallowed_shift
role_shift_pressure
constraints… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-consistency-under-role-shift-v0.1.robotics-human-intent-alignment-v0.1What this dataset tests
The robot correctly interprets human signals
The robot respects safety constraints
The robot asks clarifying questions when needed
Why this exists
Robots fail around humans when they
ignore stop signals
act too literally
overreach without confirmation
miss gestures
treat ambiguity as certainty
Data format
human_signal
context
robot_interpretation
robot_action
outcome
Task
Emit one intent label
Give one short reason
Intent… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/robotics-human-intent-alignment-v0.1.legal-expert-scope-method-alignment-v0.1What this dataset is
You receive
expert credentials
opinion scope
method basis
validation status
overreach signals
You decide
Does the opinion stay within the expert’s qualification and method
Answer
coherent
or
incoherent
Why this matters
When expert coherence fails
exclusion risk rises
appeals increase
verdict safety weakens
animal-alignment-feedback
Open Paws Animal Alignment Feedback
🐾 Human feedback and preference data for aligning AI with animal advocacy values
Overview
This dataset is part of the Open Paws initiative to develop AI training data aligned with animal liberation and advocacy principles. Created to train AI systems that understand and promote animal welfare, rights, and liberation.
Dataset Details
Dataset Type: Feedback Data
Format: CSV (Comma-separated values)
Languages: Multilingual… See the full description on the dataset page: https://huggingface.co/datasets/open-paws/animal-alignment-feedback.clinical_alignment_recovery_dynamics_v0.1Clinical Alignment Recovery Dynamics
Measures whether a model corrects earlier clinical errors when new signals appear.
Output JSON
recovered
recovery_type
correct_action
Runpython scorer.py --predictions predictions.jsonl --test_csv data/test.csv
clinical-intervention-outcome-alignment-coherence-v0.1GoalJudge whether outcomes match the rationale trajectory.
Model must
detect misalignment
recommend continue, adapt, or stop
UseCompletes the rationale → execution → outcome loop.
clinical-intervention-alignment-sepsis-v1Clinical Intervention Alignment Sepsis Detection
Overview
This dataset tests whether a model can determine whether a clinical intervention is aligned with the current system state.
In complex clinical systems such as sepsis, interventions do not have uniform effects. The same treatment may stabilize the system in one physiological state while having little effect—or even destabilizing the system—in another.
The benchmark evaluates whether models can detect when an intervention is structurally… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-intervention-alignment-sepsis-v1.clinical-narrative-clinical-timeline-alignment-v0.1What this dataset tests
Whether a system can alignpatient-reported narrativeswith objective clinical timelines.
Required outputs
alignment score
narrative time shift
omitted events
overemphasized events
narrative anchors
misalignment risk band
Use case
First layer of the Healing Narrative Coherence Corpus.
ELSA-Emotion-and-Language-Style-Alignment-Dataset
ELSA: Emotion and Language Style Alignment Dataset
The ELSA (Emotion and Language Style Alignment) dataset provides fine-grained emotional rewrites of text across four stylistic contexts: conversational, formal, poetic, and narrative. It is designed to support research in emotion-conditioned generation, stylistic variation, and affect-aware NLP.
Overview
Source: Based on the dair-ai/emotion dataset and emotion labels aligned with the GoEmotions taxonomy.
Labels:… See the full description on the dataset page: https://huggingface.co/datasets/joyspace-ai/ELSA-Emotion-and-Language-Style-Alignment-Dataset.matrix-game-evalarXivEdits_alignments
Dataset Card for ArXivEdits (Sentence-aligned documents)
ArXivEdits is a dataset comprising 751 English scientific papers from arXiv, each with sentence alignments across multiple revisions.
It also includes fine-grained, span-level edits which are annotated with the revision type and the underlying intention for 1000 sentences.
This dataset consists of only the sentence-aligned documents subset of the whole dataset. The annotated edits can be found in this dataset.… See the full description on the dataset page: https://huggingface.co/datasets/miwytt/arXivEdits_alignments.clinical_alignment_flip_test_v0.1Clinical Alignment Flip Test
Detect when a case shifts and the clinical frame must flip.
Output JSON
flip_detected
flip_type
correct_action
Runpython scorer.py --predictions predictions.jsonl --test_csv data/test.csv
ai-temporal-5node-pressure-buf-lag-cpl-alignment-goal-drift-v0.1
What this repo does
This dataset tests whether a model can detect an alignment cascade forming over time by reading a short ordered window of signals and predicting whether goal drift lock-in occurs by the final step.
Core quad
pressurebufferlagcoupling
Prediction target
label_cascade_state
Row structure
One row represents one short time window (t0 to t3) for an AI system under alignment pressure. It includes time-series values for optimization… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-temporal-5node-pressure-buf-lag-cpl-alignment-goal-drift-v0.1.oncology-signal-alignment-boundary-v0.4
What this dataset does
This dataset tests whether a model can detect signal-alignment failure in a synthetic tissue ecology.
The task is not cancer diagnosis.
The task is to classify whether readable biological signals can still coordinate repair.
Core Stability Idea
A tissue may still read damage, repair, immune, and metabolic signals but fail because those subsystems no longer align around coherent action.
This dataset moves beyond readability collapse.
It tests… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/oncology-signal-alignment-boundary-v0.4.
