datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OCR-Tibetan_layout_analysis_mask_annotation
Split: test
Total Rows: 11,907
format
Type: categorical
Data Type: object
Unique Values: 3
Value Distribution:
Value
Count
Percentage
``
9,592
80.56%
pering
1,891
15.88%
modern
424
3.56%
image_size_pixel
Type: categorical
Data Type: object
Unique Values: 1
Value Distribution:
Value
Count
Percentage
512x512
11,907
100.00%
Split: val
Total Rows: 15,630
format
Type: categorical… See the full description on the dataset page: https://huggingface.co/datasets/openpecha/OCR-Tibetan_layout_analysis_mask_annotation.financial-layout-analysis
📑 Financial Document Layout Analysis (Synthetic)
This dataset contains synthetically generated scanned financial reports intended for Document Layout Analysis. This data helps in training AI models for object detection (such as identifying tables, stamps, and figures) without compromising the confidentiality of real financial information.
📂 Label Classes
The dataset supports 6 main object classes:
0: Title (Section headers)
1: Text (Standard text paragraphs)
2: Table… See the full description on the dataset page: https://huggingface.co/datasets/Zenng2812/financial-layout-analysis.
