datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
consolidated_receipt_dataset
Dataset Card for Consolidated Receipt Dataset
This is a FiftyOne dataset with 800 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Voxel51/consolidated_receipt_dataset")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/consolidated_receipt_dataset.consolidated_receipt_dataset
Dataset Card for Consolidated Receipt Dataset
This is a FiftyOne dataset with 800 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Voxel51/consolidated_receipt_dataset")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/Zarov888/consolidated_receipt_dataset.i1-consolidated
MiniT2I Processed Mixtures
Processed image/caption subsets for MiniT2I training.
Each config contains train rows with image, captions, key, and source.
Load one subset with:
from datasets import load_dataset
ds = load_dataset("owner/repo", "pexels", split="train")
consolidated_videogame_art_with_captionscell-seg-consolidatedconsolidatedrepomalayalam_char_consolidatedconsolidated_receipt_datasetConsolidated_Dataset_MedLeavesvertexai_predictions_consolidatedoss20b_predictions_consolidated
