CoolFace
Datasetpublic

aceknock/IAM-line

IAM - line level Dataset Summary The IAM Handwriting Database contains forms of handwritten English text which can be used to train and test handwritten text recognizers and to perform writer identification and verification experiments. Note that all images are resized to a fixed height of 128 pixels. Languages All the documents in the dataset are written in English. Dataset Structure Data Instances { 'image':… See the full description on the dataset page: https://huggingface.co/datasets/aceknock/IAM-line.

sourceHugging Facemitupdated 6mo agoView on Hugging Face
0likes8downloads
Dataset Card

IAM - line level

Table of Contents

Dataset Description

Dataset Summary

The IAM Handwriting Database contains forms of handwritten English text which can be used to train and test handwritten text recognizers and to perform writer identification and verification experiments.

Note that all images are resized to a fixed height of 128 pixels.

Languages

All the documents in the dataset are written in English.

Dataset Structure

Data Instances

{
  'image': <PIL.JpegImagePlugin.JpegImageFile image mode=RGB size=2467x128 at 0x1A800E8E190,
  'text': 'put down a resolution on the subject'
}

Data Fields

  • image: a PIL.Image.Image object containing the image. Note that when accessing the image column (using dataset[0]["image"]), the image file is automatically decoded. Decoding of a large number of image files might take a significant amount of time. Thus it is important to first query the sample index before the "image" column, i.e. dataset[0]["image"] should always be preferred over dataset["image"][0].
  • text: the label transcription of the image.