elliot-mllm/coco_train_cleaned
coco_train_cleaned The coco_train family of the ElliotVL supervised-fine-tuning pool, after VLM cleaning. images 118,027 QA turns 985,335 answers rewritten by the cleaning pass 0 QA created by the cleaning pass (new_qa) not measured for this family shards 39 How this was cleaned A vision-language model read each image together with its QA and judged the item. The pass is not a filter that only removes rows — it rewrites answers it finds… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/coco_train_cleaned.
This repository belongs to elliot-mllm on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
