CoolFace
Datasetpublicgated

elliot-mllm/icdar2015_train_cleaned

icdar2015_train_cleaned The icdar2015_train family of the ElliotVL supervised-fine-tuning pool, after VLM cleaning. images 968 QA turns 4,498 answers rewritten by the cleaning pass 874 QA created by the cleaning pass (new_qa) 3,531 (78.5%) shards 1 How this was cleaned A vision-language model read each image together with its QA and judged the item. The pass is not a filter that only removes rows — it rewrites answers it finds wrong but… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/icdar2015_train_cleaned.

sourceHugging Faceotherupdated 25d agoView on Hugging Face
0likes21downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.