datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
image_dummy\dummy_image_text_data
Dataset Card for "dummy_image_text_data"
More Information needed
dummy_image_class_data
Dataset Card for "dummy_image_class_data"
More Information needed
artelingo-dummyArtELingo is a benchmark and dataset introduced in a research paper aimed at promoting work on diversity across languages and cultures. It is an extension of ArtEmis, which is a collection of 80,000 artworks from WikiArt with 450,000 emotion labels and English-only captions. ArtELingo expands this dataset by adding 790,000 annotations in Arabic and Chinese. The purpose of these additional annotations is to evaluate the performance of "cultural-transfer" in AI systems.
The dataset in ArtELingo… See the full description on the dataset page: https://huggingface.co/datasets/youssef101/artelingo-dummy.indonesian-id-card-dummy
Indonesian KTP Dataset 24K (Flat & Augmented - Commercially Safe)
Welcome to the Indonesian KTP (Kartu Tanda Penduduk) Dataset. This is a highly robust, high-fidelity, and commercially safe synthetic dataset designed to advance SOTA (State-of-the-Art) research in Document Information Extraction (DIE), Key Information Extraction (KIE), and Optical Character Recognition (OCR) specifically for Indonesian National ID cards (KTP-el).
The dataset is natively packaged in Apache Parquet… See the full description on the dataset page: https://huggingface.co/datasets/cloverx-id/indonesian-id-card-dummy.vidore_benchmark_qa_dummy
Dataset Description
This dataset is a small subset of the vidore/syntheticDocQA_energy_test dataset.
It aims to be used for debugging and testing.
Load the dataset
from datasets import load_dataset
ds = load_dataset("vidore/vidore_benchmark_qa_dummy", split="test")
Dataset Structure
Here is an example of a dataset instance structure:
features:
- name: query
dtype: string
- name: image
dtype: image
- name: image_filename
dtype: string
-… See the full description on the dataset page: https://huggingface.co/datasets/vidore/vidore_benchmark_qa_dummy.vidore_benchmark_beir_dummy
Dataset Description
This dataset is a small subset of the vidore/syntheticDocQA_energy_test_beir dataset.
It aims to be used for debugging and testing.
Load the dataset
from datasets import load_dataset
ds = load_dataset("vidore/vidore_benchmark_beir_dummy", split="test")
Dataset Structure
Here is an example of a dataset instance structure:
features:
- name: query
dtype: string
- name: image
dtype: image
- name: image_filename
dtype: string… See the full description on the dataset page: https://huggingface.co/datasets/vidore/vidore_benchmark_beir_dummy.vidore_benchmark_ocr_qa_dummydummy_images
Dataset Card for dummy_images
This dataset card aims to describe the test data for AI Powered Image Restoration and Enhancement project from Computer Vision Challenge.
Tasks:
Super-resolution: increasing image resolution
De-noising: removing noise
De-blurring: sharpening blurry images
Colorization: adding color information to grayscale images
Dataset Details
Uses
from datasets import load_dataset
ds = load_dataset("afondiel/dummy_images")
print(ds)… See the full description on the dataset page: https://huggingface.co/datasets/afondiel/dummy_images.dummydummy-pickscore-datasetExamples are taken from https://huggingface.co/spaces/yuvalkirstain/PickScore/
dummy-datasetdummy1dummyImagedummy-mpdocvqadummy-train-val-setdonut-dummy
Dataset Card for "donut-dummy"
More Information needed
dummy-pdfs-1_alp3vqasynth_sample_processed_dummylerobot_dummy_testThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 3,
"total_frames": 45,
"total_tasks": 3,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:3"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/brandonyang/lerobot_dummy_test.animalsdummy-pdfs-1_defnevisheye_dummy
Dataset Card for 2025.03.27.15.31.49
This is a FiftyOne dataset with 1 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Abeyankar/visheye_dummy")
# Launch the App
session = fo.launch_app(dataset)
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/Abeyankar/visheye_dummy.vqasynth_sample_processed_dummy_fulldummy-code-quizdummy-pdfs-2dummy_datasetsworldmodel_wa_multimodal_dummy
Dataset Card for "worldmodel_wa_multimodal_dummy"
More Information needed
dummy-pdfs-2_defnedummy-pdfs-1-labeled
