caveman273/aida-typewritten
typewritten OCR training data from AIDA-project Dataset Summary This dataset contains typewritten textline images and their transcriptions from the AIDA-project. It is a subset of the full AIDA dataset, containing only the best-quality typwritten annotations — lines where the annotator was confident about every character. The majority of lines are in Finnish, with some Swedish, English, French, and German. Supported Tasks The dataset was created for… See the full description on the dataset page: https://huggingface.co/datasets/caveman273/aida-typewritten.
Upload README.md with huggingface_hub
Upload validation.parquet with huggingface_hub
Upload test.parquet with huggingface_hub
Upload train.parquet with huggingface_hub
initial commit
