CoolFace
Datasetpublic

nabin2004/manim-narrated-dpo-400

manim-narrated-dpo-400 Direct Preference Optimization (DPO) dataset pairing 361 verified, diverse narrated VoiceoverScene scripts (chosen) against structurally identical un-narrated silent Scene scripts (rejected), curated from authentic code-agent trajectories in nabin2004/AOS-Trajectories. Dataset Summary Size: 361 preference pairs (100% unique user visualization prompts). Domains: Linear algebra (eigenvalues, SVD, transformations), calculus, machine learning… See the full description on the dataset page: https://huggingface.co/datasets/nabin2004/manim-narrated-dpo-400.

sourceHugging Faceapache-2.0updated 19d agoView on Hugging Face
0likes95downloads
11 commits on main
90e4eec19d ago

Upload clean DPO preference dataset (361 samples from AOS-Trajectories)

nabin2004
9b8ead724d ago

Upload README.md with huggingface_hub

nabin2004
899613524d ago

Upload train.jsonl with huggingface_hub

nabin2004
2155fdd24d ago

Upload README.md with huggingface_hub

nabin2004
7c3764624d ago

Upload train.jsonl with huggingface_hub

nabin2004
632b87e25d ago

Upload README.md with huggingface_hub

nabin2004
b1bd7f325d ago

Upload train.jsonl with huggingface_hub

nabin2004
aef350425d ago

Upload train.jsonl with huggingface_hub

nabin2004
82f7dcd25d ago

Upload README.md with huggingface_hub

nabin2004
919223425d ago

Upload train.jsonl with huggingface_hub

nabin2004
003d69825d ago

initial commit

nabin2004