datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
eval_gr00t_n1d6-vials_rackleft_real_fix_normalization_0202_evalThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101_follower",
"total_episodes": 2,
"total_frames": 2536,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sreetz-nv/eval_gr00t_n1d6-vials_rackleft_real_fix_normalization_0202_eval.eval_groot-vials_rackleft_real_fix_normalization_0202_8This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 506,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sreetz-nv/eval_groot-vials_rackleft_real_fix_normalization_0202_8.deita-no-normalization
Dataset Card for deita-no-normalization
This dataset has been created with Distilabel.
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated it in distilabel using the distilabel CLI:
distilabel pipeline run --config "https://huggingface.co/datasets/distilabel-internal-testing/deita-no-normalization/raw/main/pipeline.yaml"
or explore the configuration:
distilabel pipeline info --config… See the full description on the dataset page: https://huggingface.co/datasets/distilabel-internal-testing/deita-no-normalization.clinical-narrative-implicit-normalization-bias-v0.4
Implicit Normalization Bias
Clinical Narrative Integrity v0.4
Purpose
This dataset tests whether a model:
Avoids assuming normality when data is missing
Resists default reassurance
Preserves honest narrative boundaries
Treats “normal” as a claim, not a default
You are measuring baseline discipline.
Why this dataset exists
Clinical notes often omit information.
A failure mode distinct from hallucinated negatives is more subtle:
Turning… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-narrative-implicit-normalization-bias-v0.4.
