k8mpass-fraktur-series/fraktur-baltic-corpus
Fraktur Baltic Corpus Fraktur Baltic Corpus is a multilingual dataset based on historical German-language books printed in the Russian Empire during the 18th–19th centuries, primarily in Fraktur typeface. Each entry in the dataset contains: Raw OCR text from historical Fraktur sources Normalized German version Translations into seven languages: English, Russian, Estonian, Swedish, Finnish, Danish, and Modern German Volume 1: Hansen, Geschichte der Stadt Narva… See the full description on the dataset page: https://huggingface.co/datasets/k8mpass-fraktur-series/fraktur-baltic-corpus.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face