datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ml-design-doc-reviewer-data
ml-system-design/ml-design-doc-reviewer-data (v1.0.0)
Evaluation artifacts for the ML Design Doc Reviewer project.
Layout
Path
Description
manifest/sample_manifest.csv
Stratified 100-case sample manifest
manifest/error_topology.csv
Controlled error taxonomy for flawed docs
raw/
Raw markdown exports, metadata sidecars, OCR image blocks
raw/images/
Downloaded article images
normalized/
Canonical 14-section ML design documents
flawed/
Normalized… See the full description on the dataset page: https://huggingface.co/datasets/ml-system-design/ml-design-doc-reviewer-data.ui-design-audit-dataset
UI Design Audit Screenshot Benchmark v2.1
A reproducible synthetic benchmark of 3,000 mobile and web UI screenshots labeled across 12 design-risk categories.
Splits
train: 2,400
validation: 300
test: 300
Labels
small_touch_targets
low_contrast
action_overload
navigation_overload
form_friction
content_density
responsive_risk
modal_overuse
deep_scrolling
weak_hierarchy
interaction_overload
mobile_web_mismatch
Data creation
Every… See the full description on the dataset page: https://huggingface.co/datasets/newazhala/ui-design-audit-dataset.
