datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dousekoishiteshimaunda
Bangumi Image Base of Douse, Koishite Shimaunda.
This is the image base of bangumi Douse, Koishite Shimaunda., we detected 42 characters, 3811 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1%… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/dousekoishiteshimaunda.bangla-ocr-double-benchmark
Bangla OCR Double Benchmark
Two equally weighted, deterministic full-page Bangla handwriting robustness splits:
bongabdo: 6,669 readability-preserving renderings balanced over all 111 Bongabdo pages.
bn_htrd: 6,669 renderings balanced over all 75 actual files in the writer-separated
BN-HTRd test split.
These are explicitly compositional/augmentation robustness rows, not 13,338 independent
writers or source documents. Every row exposes its source page ID, source SHA-256… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/bangla-ocr-double-benchmark.hongyan_transfer_bottle_double_hand_vedio_testThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 16,
"total_frames": 12865,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:16"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gaozj/hongyan_transfer_bottle_double_hand_vedio_test.double_pendulumDataset from 2022-07-15. There are two columns: image and trajectory. Image is the image file and trajectory is the label trajectory number_image number for the corresponding image.
douban_movie_info该数据集为豆瓣电影信息维表。
更多信息请参考文章《数据获取:豆瓣电影信息爬取》。
MediPPD
MediPPD
MediPPD contains images and annotations used by the main MediPPD experiment for automated interpretation of purified protein derivative (PPD) skin tests. Public release of the source images has been authorized by the data owner.
Dataset summary
Total image cases: 554
Training cases: 443
Validation cases: 111
Bottle-cap cases: 553
Red/swollen reaction cases: 383
Blister cases: 20
Necrosis cases: 15
Double-ring cases: 17
The five image labels are… See the full description on the dataset page: https://huggingface.co/datasets/DoubleYue/MediPPD.douluodalupersian-ocr-double-benchmark
Persian OCR Double Benchmark
A frozen, leakage-controlled Persian OCR benchmark with two equally weighted splits:
printed: 6,669 rows carved from Reza2kn/persian-printed-ocr-3.5m at ba02f36c3d496838d8fad9aff352b77763af1ea4.
handwriting: 6,669 rows carved from Reza2kn/persian-handwriting-pages-3.69m at b114f0a36a6a2e397bc93517dd084431bfca2329.
The exact rows were uploaded here before being removed from their source training repositories.
Each row retains its original repository… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-ocr-double-benchmark.dou-brazil-dataset
Dataset Card for Dataset Diário Oficial da União (DOU)
The Diário Oficial da União (DOU) is the official government gazette of Brazil, published by the National Press. It serves as the primary means of communication for federal government acts, including laws, decrees, ordinances, public notices, and other official decisions. The DOU ensures transparency and legal validity for government actions and is divided into three sections:
Section 1: Publishes laws, decrees, and… See the full description on the dataset page: https://huggingface.co/datasets/gerson-vfs/dou-brazil-dataset.pick-double-caption
Dual Caption Preference Optimization for Diffusion Models
We propose DCPO, a new paradigm to improve the alignment performance of text-to-image diffusion models. For more details on the technique, please refer to our paper here.
Developed by
Amir Saeidi*
Yiran Luo*
Agneet Chatterjee
Shamanthak Hegde
Bimsara Pathiraja
Yezhou Yang
Chitta Baral
Dataset
This dataset is Pick-Double Caption, a modified version of the Pick-a-Pic V2 dataset. We… See the full description on the dataset page: https://huggingface.co/datasets/DualCPO/pick-double-caption.CrossHall-Bench
CrossHall-Bench
CrossHall-Bench evaluates cross-call hallucinations in visual tool-using agents. The public model input for each instance contains only the original image and the full question. Gold answers and reference annotations are used only after inference for scoring.
This Hugging Face release contains the benchmark data and official evaluation code only. Dataset construction, reference-generation, manual-audit, release-audit, and upload utilities are intentionally not… See the full description on the dataset page: https://huggingface.co/datasets/Doushabo/CrossHall-Bench.SDv2-Spatial-allseeds-rectifiedstable-diffusion-2-1_obj15_set4_seed100SDv2-Count-seedmining-v4-top1douban_movie_info该数据集为豆瓣电影信息维表。
更多信息请参考文章《数据获取:豆瓣电影信息爬取》。
douban_crawlerArtBench-10cloud-adapter-datasets
Cloud-Adapter-Datasets
This dataset card aims to describe the datasets used in the Cloud-Adapter, a collection of high-resolution satellite images and semantic segmentation masks for cloud detection and related tasks.
Install
pip install huggingface-hub
Usage
# Step 1: Download datasets
huggingface-cli download --repo-type dataset XavierJiezou/cloud-adapter-datasets --local-dir data --include hrc_whu.zip
huggingface-cli download --repo-type dataset… See the full description on the dataset page: https://huggingface.co/datasets/doudou112/cloud-adapter-datasets.SDv2-Spatial-seedmining-v1.01-replacedSDv2_512-Spatial-allseeds-replacedSDv2-Count70-details-appendedSDv2_512-Spatial-seedmining-replacedhongyan_transfer_bottle_double_handThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 31,
"total_frames": 23588,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:31"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gaozj/hongyan_transfer_bottle_double_hand.VanGogh_Stevedores_DoubleEntry_ForensicConsistencyTestREADME: Comparative Forensic Dataset
Title: Tree Oil Painting vs. Van Gogh's Los Descargadores en Arles (1888)
Version: August 2025
🌐 Overview
This dataset presents a comprehensive forensic comparison between an anonymous artwork known as The Tree Oil Painting and Vincent van Gogh's 1888 painting Los descargadores en Arles (also referred to in some archives as The Stevedores in Arles). The primary aim is to evaluate gestural, structural, and energetic consistency using the AI-based framework… See the full description on the dataset page: https://huggingface.co/datasets/HaruthaiAi/VanGogh_Stevedores_DoubleEntry_ForensicConsistencyTest.multimodal_confounding
Dataset Card
Semi-synthetic dataset with multimodal confounding.
The dataset is generated according to the description in DoubleMLDeep: Estimation of Causal Effects with Multimodal Data.
Dataset Details
Dataset Description & Usage
The dataset is a semi-synthetic dataset as a benchmark for treatment effect estimation with multimodal confounding. The outcome
variable Y is generated according to a partially linear model
Y=θ0D1+g1(X)+ε
Y = \theta_0 D_1 + g_1(X) +… See the full description on the dataset page: https://huggingface.co/datasets/DoubleML/multimodal_confounding.DoubanActorT2IRetrievalSDv2-Count-allseeds-rectifiedSDv2-Count-seedmining-v1.01-rectifiedSDv2-Count70-kangarooDoubleE-dataset
