datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bangla-ocr-double-benchmark
Bangla OCR Double Benchmark
Two equally weighted, deterministic full-page Bangla handwriting robustness splits:
bongabdo: 6,669 readability-preserving renderings balanced over all 111 Bongabdo pages.
bn_htrd: 6,669 renderings balanced over all 75 actual files in the writer-separated
BN-HTRd test split.
These are explicitly compositional/augmentation robustness rows, not 13,338 independent
writers or source documents. Every row exposes its source page ID, source SHA-256… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/bangla-ocr-double-benchmark.hongyan_transfer_bottle_double_hand_vedio_testThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 16,
"total_frames": 12865,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:16"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gaozj/hongyan_transfer_bottle_double_hand_vedio_test.double_pendulumDataset from 2022-07-15. There are two columns: image and trajectory. Image is the image file and trajectory is the label trajectory number_image number for the corresponding image.
persian-ocr-double-benchmark
Persian OCR Double Benchmark
A frozen, leakage-controlled Persian OCR benchmark with two equally weighted splits:
printed: 6,669 rows carved from Reza2kn/persian-printed-ocr-3.5m at ba02f36c3d496838d8fad9aff352b77763af1ea4.
handwriting: 6,669 rows carved from Reza2kn/persian-handwriting-pages-3.69m at b114f0a36a6a2e397bc93517dd084431bfca2329.
The exact rows were uploaded here before being removed from their source training repositories.
Each row retains its original repository… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-ocr-double-benchmark.pick-double-caption
Dual Caption Preference Optimization for Diffusion Models
We propose DCPO, a new paradigm to improve the alignment performance of text-to-image diffusion models. For more details on the technique, please refer to our paper here.
Developed by
Amir Saeidi*
Yiran Luo*
Agneet Chatterjee
Shamanthak Hegde
Bimsara Pathiraja
Yezhou Yang
Chitta Baral
Dataset
This dataset is Pick-Double Caption, a modified version of the Pick-a-Pic V2 dataset. We… See the full description on the dataset page: https://huggingface.co/datasets/DualCPO/pick-double-caption.hongyan_transfer_bottle_double_handThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 31,
"total_frames": 23588,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:31"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gaozj/hongyan_transfer_bottle_double_hand.double_datasetThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "aloha",
"total_episodes": 40,
"total_frames": 24241,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 50,
"splits": {
"train": "0:40"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/zhifeishen/double_dataset.double_exposureface-eye-double-eyelids2
Dataset Card for "face-eye-double-eyelids2"
More Information needed
double_side_dataset_v1
