CoolFace
Datasetpublic

Riksarkivet/swedish_fraktur

Swedish Fraktur This is a dataset for swedish blackletter from the 19th century. The transcriptions were made by Språkbanken and converted into a text-line dataset by the Swedish National Archives Dataset Details Uses Direct Use Train textline-based OCR models for swedish 19th century blackletter Dataset Structure { "image": Image(), "text": str } Dataset Creation Source Data Original… See the full description on the dataset page: https://huggingface.co/datasets/Riksarkivet/swedish_fraktur.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
2likes163downloads
Dataset Card

Swedish Fraktur

<!-- Provide a quick summary of the dataset. -->

This is a dataset for swedish blackletter from the 19th century. The transcriptions were made by Språkbanken and converted into a text-line dataset by the Swedish National Archives

Dataset Details

Dataset Description

  • —Curated by: Språkbanken, The Swedish National Archives
  • —Language(s) (NLP): Swedish
  • —License: [More Information Needed]

Uses

Direct Use

Train textline-based OCR models for swedish 19th century blackletter

Dataset Structure

{ "image": Image(), "text": str }

Dataset Creation

Source Data

Original dataset from Språkbanken - svenska tidningar 1818-1870

Who are the source data producers?

Språkbanken

Dataset Card Authors

Erik Lenas

Dataset Card Contact

ai@riksarkivet.se