CoolFace
Datasetpublic

lmms-lab-encoder/textvqa

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of TextVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @inproceedings{singh2019towards, title={Towards vqa models that can read}, author={Singh, Amanpreet and… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-encoder/textvqa.

sourceHugging Faceupdated 3y agoView on Hugging Face
25likes48kdownloads

lmms-lab-encoder/textvqa · main · files are served by the source, never re-hosted here