nabin2004/manim-narrated-dpo-400
manim-narrated-dpo-400 Direct Preference Optimization (DPO) dataset pairing 361 verified, diverse narrated VoiceoverScene scripts (chosen) against structurally identical un-narrated silent Scene scripts (rejected), curated from authentic code-agent trajectories in nabin2004/AOS-Trajectories. Dataset Summary Size: 361 preference pairs (100% unique user visualization prompts). Domains: Linear algebra (eigenvalues, SVD, transformations), calculus, machine learning… See the full description on the dataset page: https://huggingface.co/datasets/nabin2004/manim-narrated-dpo-400.
Upload clean DPO preference dataset (361 samples from AOS-Trajectories)
Upload README.md with huggingface_hub
Upload train.jsonl with huggingface_hub
Upload README.md with huggingface_hub
Upload train.jsonl with huggingface_hub
Upload README.md with huggingface_hub
Upload train.jsonl with huggingface_hub
Upload train.jsonl with huggingface_hub
Upload README.md with huggingface_hub
Upload train.jsonl with huggingface_hub
initial commit
