datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
persian-ocr-community-dataset-layout
Persian OCR Community Layout Annotations
Resumable layout annotations for the page images in
Reza2kn/persian-ocr-community-dataset.
Each row points to an exact source dataset revision, Parquet shard, blob, and row. It includes the
page identifier, page dimensions, handwriting flag, and structured layout boxes produced by
datalab-to/surya_layout2 at confidence threshold
0.4.
The boxes field contains label, confidence, raster-order position, and pixel coordinates
x0, y0, x1, y1.… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-ocr-community-dataset-layout.ingredient-detection-layout-dataset
Dataset Card for "ingredient-detection-layout-dataset"
More Information needed
persian-ocr-community-dataset-layout-mergedkhmer-newspaper-layout-dataset
Khmer Newspaper Layout Dataset
Dataset Description
This dataset contains Khmer newspaper layouts with annotated regions for document layout analysis and OCR tasks.
Dataset Summary
Total Examples: 9,344 newspaper layouts
Language: Khmer (Cambodian)
Image Format: PNG
Annotations: LabelMe JSON format with bounding boxes and segmentation masks
Features
config_id: Unique identifier for each sample
image: Newspaper layout image (PNG)… See the full description on the dataset page: https://huggingface.co/datasets/jonny122/khmer-newspaper-layout-dataset.mgh-critical-edition-layout
MGH Layout Detection Dataset
Dataset Description
General Description
This dataset consists of scans from the MGH critical edition of Alcuin's letters, which were first edited by Ernestus Duemmler in 1895. The digital scans were sourced from the DMGH's repository, which can be accessed here. The scans were annotated using CVAT, marking out two classes: the title of a letter and the body of the letter.
Why was this dataset created?
The primary… See the full description on the dataset page: https://huggingface.co/datasets/medieval-data/mgh-critical-edition-layout.layout-data-khmer-3illuin_layout_dataset_text_only
Dataset Card for "illuin_layout_dataset_text_only"
More Information needed
elementor-layout-vlm-dataset
Elementor Layout VLM Dataset
📊 Dataset Summary
Dataset para fine-tuning de modelos Vision-Language (VLM) para geração de layouts Elementor a partir de imagens.
Task: Visual Question Answering (VQA)
Format: VLM VQA (image, question, answer)
Total: 30 exemplos
Training: 24 exemplos
Validation: 6 exemplos
🎯 Uso Recomendado
AutoTrain Configuration
Task: VLM VQA
Base Model: google/paligemma-3b-pt-448
Dataset: vinicios94/elementor-layout-vlm-dataset… See the full description on the dataset page: https://huggingface.co/datasets/vinicios94/elementor-layout-vlm-dataset.alwas-analog-layout-dataset
ALWAS Analog Layout Dataset
Synthetic dataset for training ML models in the ALWAS (Analog Layout Workflow Automation System) pipeline.
Dataset Description
4,000 analog IC layout blocks with complete metadata, stage transitions, and labels for:
Hours estimation — actual vs estimated hours
Complexity classification — Low / Medium / High
Bottleneck risk prediction — Low / Medium / High
Completion time prediction — stage-by-stage transition history
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/muthuk1/alwas-analog-layout-dataset.persian-ocr-community-dataset-layout-argilla
persian-ocr-community-dataset-layout-argilla
This dataset contains only two columns for Argilla import: image and label.
Total pages (rows): 5878
Bboxes with OCR text: 33486
Bboxes without OCR text: 0
Pages with at least one OCR: 5871
Pages with zero OCR: 7
text-based-layout-generation-datasettable_layout_dataset_v1
Table Dataset - Image & LabelMe & OBB Annotation (Train/Val Split)
Dataset Overview
Comprehensive table detection dataset with ground truth LabelMe polygon annotations and OBB (Oriented Bounding Box) data, split into training and validation sets.
Total examples: 10,000 image-annotation pairs
Train: 8,000 (80.0%)
Validation: 2,000 (20.0%)
Total size: 670.25 MB
Language: km
Document types: Table/Chart documents
Ground truth: LabelMe polygon annotations… See the full description on the dataset page: https://huggingface.co/datasets/vichetkao/table_layout_dataset_v1.table_layout_dataset_random_v1
Table Dataset - Image & OBB Annotation (Train/Val Split)
Dataset Overview
Table detection dataset with OBB (Oriented Bounding Box) annotations in YOLO format,
split into training and validation sets.
Total examples: 3762 image-annotation pairs
Train: 3053 (81.2%)
Validation: 709 (18.8%)
Total size: 522.45 MB
Language: Khmer (km)
Document types: Table documents
Annotation format: YOLO OBB (class cx cy w h)
Dataset Statistics
Split… See the full description on the dataset page: https://huggingface.co/datasets/vichetkao/table_layout_dataset_random_v1.khmer-newspaper-layout-dataset-masked-2khmer-newspaper-layout-dataset-largetable-data-layout-khm-v4-auglayout-data-khmer-aug-1
