image-captions
fujie_vit-gpt2-japanese-image-captioning_stair-captions-resultfujie_vit-bert-japanese-image-captioning_stair-captions-resultsfujie_vit-gpt2-japanese-image-captioning_stair-captions-pipelinevit-gpt2-image-captioning-instagram-captionsImage-Captions-Generatorimage_caption_git-base_pokemon-blip-captions_finetunedj_qwen-image_anat-true-w-captionsimage-caption-summarizer-qwen-3-4b
ramanv-image-captions-realramanv-image-captions-11ramanv-image-captions-2learn_hf_food_not_food_image_captions
Food/Not Food Image Caption Dataset
Small dataset of synthetic food and not food image captions.
Text generated using Mistral Chat/Mixtral.
Can be used to train a text classifier on food/not_food image captions as a demo before scaling up to a larger dataset.
See Colab notebook on how dataset was created.
Example usage
import random
from datasets import load_dataset
# Load dataset
loaded_dataset = load_dataset("mrdbourke/learn_hf_food_not_food_image_captions")
# Get… See the full description on the dataset page: https://huggingface.co/datasets/mrdbourke/learn_hf_food_not_food_image_captions.ramanv-image-captions-11image_captions
From the Frontier Research Team at takara.ai we present over 1 million curated captioned images for multimodal text and image tasks.
Usage
from datasets import load_dataset
ds = load_dataset("takara-ai/image_captions")
print(ds)
Example
10,000 images from the dataset.
Methodology
We consolidated multiple open source datasets through an intensive 96-hour computational process across three nodes. This involved standardizing and validating the… See the full description on the dataset page: https://huggingface.co/datasets/takara-ai/image_captions.
