datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SQuADDS_Layouts
SQuADDS Layouts - versioned GDS artifacts for superconducting quantum hardware
SQuADDS Layouts is the geometry-artifact companion to
SQuADDS_DB, the
Superconducting Qubit And Device Design and Simulation Database. It provides
checksum-verified GDS files, stable geometry identities, and machine-readable
geometry metadata so a simulation result can be traced to the exact layout
that produced it.
Homepage: https://lfl-lab.github.io/SQuADDS/
Repository:… See the full description on the dataset page: https://huggingface.co/datasets/SQuADDS/SQuADDS_Layouts.funsd-layoutlmv3LayoutSAM
LayoutSAM Dataset
Overview
The LayoutSAM dataset is a large-scale layout dataset derived from the SAM dataset, containing 2.7 million image-text pairs and 10.7 million entities. Each entity is annotated with a spatial position (i.e., bounding box) and a textual description.
Traditional layout datasets often exhibit a closed-set and coarse-grained nature, which may limit the model's ability to generate complex attributes such as color, shape, and texture.… See the full description on the dataset page: https://huggingface.co/datasets/HuiZhang0812/LayoutSAM.layout_diffusion_hypersimThis repository contains the data for SceneCraft: Layout-Guided 3D Scene Generation.
Project page: https://orangesodahub.github.io/SceneCraft
Code: https://github.com/OrangeSodahub/SceneCraft
room-layout-planning-curated-v1
Room Layout Planning — curated pilot v1
39 个逐条检查并编写需求的房间布局任务,供实验流程验证与人工抽查。所有最终设计需求均为 AI 编写;没有人工标注或人工复核声明。
Split
条数
独立房屋
几何来源
train
26
26
InstructScene / 3D-FRONT
val_seen
7
7
InstructScene / 3D-FRONT
val_unseen
6
5
M3DLayout / Matterport3D
输入:英文使用需求 + 可用地板多边形 + 4–10 件家具及固定宽深尺寸。输出:所有家具的二维位置与旋转角度。家具清单和尺寸不可修改。参考摆放已通过几何检查,但没有被认证为满足全部语言偏好的标准答案。
下载后打开 review.html 可以逐条浏览需求、尺寸、空房轮廓、参考图和修订理由。原文与 46 条逐条审核记录见 individual_reviews.jsonl,其中 39 条保留、7 条排除。此次规模适合跑通… See the full description on the dataset page: https://huggingface.co/datasets/yfan1997/room-layout-planning-curated-v1.SORIE_layoutlmv2
Dataset Card for "SORIE_layoutlmv2"
More Information needed
docbank-layout
Support the Project ☕
If you find this dataset helpful, please support me with a mocha:
Dataset Summary
DocBank is a large-scale dataset tailored for Document AI tasks, focusing on integrating textual and layout information. It comprises 500,000 document pages, divided into 400,000 for training, 50,000 for validation, and 50,000 for testing. The dataset is generated using a weak supervision approach, enabling efficient annotation of document structures… See the full description on the dataset page: https://huggingface.co/datasets/astrologos/docbank-layout.layout_distribution_shiftLayoutOrderingHard3d_layout_reasoningDataset for MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
Github: https://github.com/PzySeere/MetaSpatial
SQuADDS_Layout_Embeddings
SQuADDS Layout Embeddings
Versioned layout representations for the 24,106 GDS artifacts in
SQuADDS/SQuADDS_Layouts.
Static embedding model v0
static-embedding-v0 implements the original SQuADDS proof-of-concept model:
v0 = parameter_sum + geometric_moments + flattened_shape_bitmap
Each unit-normalized vector has 9,227 dimensions:
Block
Dimensions
Contents
Parameter sum
1
Permutation- and parameter-count-invariant sum of numerical design options… See the full description on the dataset page: https://huggingface.co/datasets/SQuADDS/SQuADDS_Layout_Embeddings.LayoutSAM-eval
LayoutSAM-eval Benchmark
Overview
LayoutSAM-Eval is a comprehensive benchmark for evaluating the quality of Layout-to-Image (L2I) generation models. This benchmark assesses L2I generation quality from two perspectives: region-wise quality (spatial and attribute accuracy) and global-wise quality (visual quality and prompt following). It employs the VLM’s visual question answering to evaluate spatial and attribute adherence, and utilizes various metrics including IR score… See the full description on the dataset page: https://huggingface.co/datasets/HuiZhang0812/LayoutSAM-eval.DocVQA_layoutLM
Dataset Card for "DocVQA_layoutLM"
More Information needed
persian-ocr-community-dataset-layout
Persian OCR Community Layout Annotations
Resumable layout annotations for the page images in
Reza2kn/persian-ocr-community-dataset.
Each row points to an exact source dataset revision, Parquet shard, blob, and row. It includes the
page identifier, page dimensions, handwriting flag, and structured layout boxes produced by
datalab-to/surya_layout2 at confidence threshold
0.4.
The boxes field contains label, confidence, raster-order position, and pixel coordinates
x0, y0, x1, y1.… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-ocr-community-dataset-layout.ingredient-detection-layout-dataset
Dataset Card for "ingredient-detection-layout-dataset"
More Information needed
fineweb-vi-pdf-layout
Fineweb VI PDF Layout (v2)
Vietnamese web pages (from HuggingFaceFW/fineweb-2) rendered to PDF, with per-page
layout annotations for VLM training.
Each row = 1 document:
Column
Type
Description
doc_name
string
Document ID (e.g. doc_000)
url
string
Source URL (metadata)
pdf
binary
Rendered document.pdf
layouts_merged
string
All per-page layout JSONs merged: {"num_pages": N, "pages": [...]}
pages
image list
Rendered page images (page_*.png)
source_html
string… See the full description on the dataset page: https://huggingface.co/datasets/dauvannam321/fineweb-vi-pdf-layout.nutrient-detection-layout
Nutrient extraction dataset
This dataset contains annotated images of nutrition tables. The goal of this dataset was to train a model to extract nutrient values from nutrition tables, as part of the Nutrisight project.
It contains ~3k samples in total (2.8k for training and 199 for testing). For more information about the project, please refer to the nutrisight directory in the openfoodfacts-ai GitHub repository.
The images were collected from the Open Food Facts database, and… See the full description on the dataset page: https://huggingface.co/datasets/openfoodfacts/nutrient-detection-layout.vivid_layoutpersian-ocr-community-dataset-layout-mergeddhivehi-layout-syn-lg-florence
Synthetic Dhivehi Document Layout Analysis Dataset
Overview
This dataset contains synthetic document layouts annotated with bounding boxes and labels for various sections in the Dhivehi language. It is designed for training models on document layout analysis and Optical Character Recognition (OCR) tasks. The dataset simulates real-world documents in Dhivehi and can be used for layout-aware OCR to recognize text and understand the structure of Dhivehi documents.… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-layout-syn-lg-florence.Tridis_layout_manuscripts
A Unified Dataset for Codicological Document Layout Analysis
Dataset Description
This repository contains a large-scale, unified dataset for Document Layout Analysis (DLA) in historical manuscripts. It was created by harmonizing three distinct public corpora—e-NDP, CATMuS, and HORAE—which cover a wide range of document types from the 12th to the 17th century (administrative registers, literary manuscripts, printed books, and Books of Hours).
The key feature of this… See the full description on the dataset page: https://huggingface.co/datasets/magistermilitum/Tridis_layout_manuscripts.CIVQA-TesseractOCR-LayoutLM
CIVQA TesseractOCR LayoutLM Dataset
The Czech Invoice Visual Question Answering dataset was created with Tesseract OCR and encoded for the LayoutLM.
The pre-encoded dataset can be found on this link: https://huggingface.co/datasets/fimu-docproc-research/CIVQA-TesseractOCR
All invoices used in this dataset were obtained from public sources. Over these invoices, we were focusing on 15 different entities, which are crucial for processing the invoices.
Invoice number
Variable… See the full description on the dataset page: https://huggingface.co/datasets/SpringRollMonster/CIVQA-TesseractOCR-LayoutLM.myanmar_complex_document_layouts
🇲🇲 Myanmar Complex Document Layouts
A large-scale, high-quality synthetic dataset containing 17,632 images of complex document layouts, dashboards, and infographics entirely in the Myanmar (Burmese) language.
This dataset is specifically designed to train and benchmark modern Computer Vision and multimodal LLMs on complex Myanmar typography, structured data, and diverse graphical layouts.
📊 Dataset Overview
Total Images: 17,632 high-resolution pages.… See the full description on the dataset page: https://huggingface.co/datasets/freococo/myanmar_complex_document_layouts.form-fields-for-layout-labeled-pagesdhivehi-layout-syn-lg-paligemma
Synthetic Dhivehi Document Layout Analysis Dataset
Overview
This dataset contains synthetic document layouts annotated with bounding boxes and labels for various sections in the Dhivehi language. It is designed for training models on document layout analysis and Optical Character Recognition (OCR) tasks. The dataset simulates real-world documents in Dhivehi and can be used for layout-aware OCR to recognize text and understand the structure of Dhivehi documents.… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-layout-syn-lg-paligemma.layouts_donut_1
Dataset Card for "layouts_donut_1"
More Information needed
layout_diffusion_scannetpp_voxel0.2This dataset is used in the paper SceneCraft: Layout-Guided 3D Scene Generation.
File information
The repository contains the following file information:
xfund-multilingual-normalized-layoutlmv3rb-box-layouts-inlinelayoutlmv3_cord
Dataset Card for "layoutlmv3_cord"
Original Dataset is "naver-clova-ix/cord-v2"
This dataset is modified for learning.
More Information needed
