Elliot-Data/icdar2013_train_cleaned
icdar2013_train_cleaned The icdar2013_train family of the ElliotVL supervised-fine-tuning pool, after VLM cleaning. images 227 QA turns 846 answers rewritten by the cleaning pass 88 QA created by the cleaning pass (new_qa) 661 (78.1%) shards 1 How this was cleaned A vision-language model read each image together with its QA and judged the item. The pass is not a filter that only removes rows — it rewrites answers it finds wrong but… See the full description on the dataset page: https://huggingface.co/datasets/Elliot-Data/icdar2013_train_cleaned.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face