yuanqianhao/Vision-OPD-6K
Vision-OPD-6K: Training Data for Vision-OPD Overview Vision-OPD proposes a regional-to-global self-distillation framework that transfers the model's own privileged regional perception to its full-image policy, without external teacher models, ground-truth labels, reward verifiers, or inference-time tool use. Vision-OPD instantiates two conditional policies from the same MLLM: A crop-conditioned teacher that observes the evidence-centered crop as a privileged… See the full description on the dataset page: https://huggingface.co/datasets/yuanqianhao/Vision-OPD-6K.
Fix bbox: clip negative values to 0 and overflow values to image dimensions
Repack original_images.tar.gz to flat structure (./sa_xxx.jpg), consistent with images/teacher_images
Upload folder using huggingface_hub
Upload train.jsonl with huggingface_hub
Upload README.md with huggingface_hub
Update README.md
Upload images/images.tar.gz05 with huggingface_hub
Upload images/images.tar.gz04 with huggingface_hub
Upload images/images.tar.gz03 with huggingface_hub
Upload images/images.tar.gz02 with huggingface_hub
Upload images/images.tar.gz01 with huggingface_hub
Upload images/images.tar.gz00 with huggingface_hub
Upload teacher_images/teacher_images.tar.gz with huggingface_hub
Delete teacher_images/teacher_images.tar.gz with huggingface_hub
Delete images/images.tar.gz02 with huggingface_hub
Delete images/images.tar.gz01 with huggingface_hub
Delete images/images.tar.gz00 with huggingface_hub
Upload Vision-OPD-6K training data
initial commit
