CoolFace
Datasetpublic

zzliang/GRIT

GRIT: Large-Scale Training Corpus of Grounded Image-Text Pairs Dataset Summary We introduce GRIT, a large-scale dataset of Grounded Image-Text pairs, which is created based on image-text pairs from COYO-700M and LAION-2B. We construct a pipeline to extract and link text spans (i.e., noun phrases, and referring expressions) in the caption to their corresponding image regions. More details can be found in the paper. Supported Tasks During the… See the full description on the dataset page: https://huggingface.co/datasets/zzliang/GRIT.

sourceHugging Facems-plupdated 3y agoView on Hugging Face
161likes970downloads
10 commits on main
4696bb03y ago

update readme

hululuhu
25a78193y ago

update readme

hululuhu
67c8e383y ago

update readme

hululuhu
d324bb23y ago

update readme

hululuhu
3b3ccf53y ago

update readme

hululuhu
6e47e033y ago

update readme

hululuhu
befac713y ago

update readme

hululuhu
65a096a3y ago

update

hululuhu
d1b34353y ago

add parquet

addf400
e1d63503y ago

initial commit

zzliang