CoolFace
Datasetpublic

stzhao/AnyWord-3M

Dataset from AnyText: Multilingual Visual Text Generation And Editing. Dataset description from Anytext Team: Currently, there is a relative scarcity of public datasets for text generation tasks, especially those involving non-Latin script languages. To address this, we introduce a large-scale multilingual dataset called AnyWord-3M. The images in this dataset are sourced from Noah-Wukong, LAION-400M, and OCR recognition datasets such as ArT, COCO-Text, RCTW, LSVT, MLT, MTWI, ReCTS, etc. These… See the full description on the dataset page: https://huggingface.co/datasets/stzhao/AnyWord-3M.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
17likes10kdownloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

stzhao/AnyWord-3M · main · files are served by the source, never re-hosted here