CoolFace
Datasetpublic

openfoodfacts/ingredient-detection

This dataset is used to train a multilingual ingredient list detection model. The goal is to automate the extraction of ingredient lists from food packaging images. See this issue for a broader context about ingredient list extraction. Dataset generation Raw unannotated texts are OCR results obtained with Google Cloud Vision. It only contains images marked as ingredient image on Open Food Facts. The dataset was generated using ChatGPT-3.5: we asked ChatGPT to extract ingredient… See the full description on the dataset page: https://huggingface.co/datasets/openfoodfacts/ingredient-detection.

sourceHugging Facecc-by-sa-4.0updated 2y agoView on Hugging Face
2likes60downloads
fileingredient_detection_dataset-v1_test.jsonl.gz380 KBdownload
fileingredient_detection_dataset-v1_train.jsonl.gz3.3 MBdownload

openfoodfacts/ingredient-detection · v1.0 · files are served by the source, never re-hosted here