datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
persian-handwriting-ocr
Persian Handwriting OCR Dataset
Dataset Summary
A standardized dataset of Persian (Farsi) handwritten pages with word-level
bounding-box annotations and transcriptions. The dataset is page-level:
each sample is a full page scan; annotations are one row per word bbox on
that page. This is the most flexible form -- users can train page-level OCR,
word detection (DBNet/PaddleOCR), or derive word/line crops as needed.
Pages: 1115 scanned pages (canonical IDs… See the full description on the dataset page: https://huggingface.co/datasets/MR3z4/persian-handwriting-ocr.ouhd-l-online-urdu-nastaliq-handwriting
OUHD-L: Online Urdu Nastaliq Handwriting — Line Pen Trajectories
Unmodified mirror. This repository re-hosts the OUHD-L v1.0 core release
exactly as published on Zenodo, byte-for-byte. Nothing has been added to or
removed from the data. It exists only to provide an alternative download
endpoint. The canonical source and citation is the Zenodo record:
https://zenodo.org/records/20642162 — DOI
10.5281/zenodo.20642162, version 1.0.0.
Overview
2,403 handwritten Urdu… See the full description on the dataset page: https://huggingface.co/datasets/saad2002/ouhd-l-online-urdu-nastaliq-handwriting.EEG-fNIRS-based-Handwriting-Trajectory-Dataset
Overview
This dataset contains the raw multimodal signals and metadata.
The goal of the competition is to predict the imagined handwriting class for each test trial using synchronized:
EEG signals
fNIRS signals
The dataset is organized to support an end-to-end competition workflow:
train_meta.csv provides labeled training trials
test_meta.csv provides unlabeled test trials
raw/ contains the corresponding raw EEG and fNIRS recordings
Files… See the full description on the dataset page: https://huggingface.co/datasets/lasfk/EEG-fNIRS-based-Handwriting-Trajectory-Dataset.
