datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Recap-DataComp-1B-FoodOrDrink
Recap-DataComp-1B: Food or Drink
A filtered subset of Recap-DataComp-1B containing 106,230,157 rows classified as food/drink content, enriched with structured food/drink extraction from FoodExtract-v2.
Overview
Count
Percentage
Total rows
106,230,157
100%
Food/drink (Stage 5 label)
96,618,895
91.0%
Not food/drink (Stage 5 label)
9,611,262
9.0%
FoodExtract (re_caption): food/drink
79,519,489
74.9%
FoodExtract (re_caption): not food/drink
26,710,156… See the full description on the dataset page: https://huggingface.co/datasets/mrdbourke/Recap-DataComp-1B-FoodOrDrink.DataComp-1B-food-and-drink-3M
DataComp-1B Food and Drink 3M
~3,108,047 food and not-food images extracted from Recap-DataComp-1B, each classified by three independent signals and accompanied by SigLIP2 embeddings (1,152-dim). Built for training food/drink classifiers, building FAISS search indices, and as a foundation for the Nutrify VLM — an on-device vision-language model for nutrition tracking.
How this dataset was made
The problem
Recap-DataComp-1B contains 1 billion… See the full description on the dataset page: https://huggingface.co/datasets/mrdbourke/DataComp-1B-food-and-drink-3M.
