CoolFace
Datasetpublicgated

abhinandansamal/odia-german-parallel-corpus-research

Dataset Summary This dataset is a high-quality, parallel corpus for Odia (Oriya) to German and German to Odia machine translation. It focuses on the news domain, specifically covering National, International, Sports, Trade, and Science & Technology topics. The dataset contains 3,676 unique parallel sentence pairs, curated through a hybrid approach combining automated web scraping, manual human translation (Gold Standard), and human-corrected machine translation (Silver… See the full description on the dataset page: https://huggingface.co/datasets/abhinandansamal/odia-german-parallel-corpus-research.

sourceHugging Faceapache-2.0updated 9mo agoView on Hugging Face
0likes3downloads
settings

This repository belongs to abhinandansamal on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameodia-german-parallel-corpus-research
visibilitypublic
licenceapache-2.0
gatedyes
ownerabhinandansamal
Account settings
abhinandansamal/odia-german-parallel-corpus-research · CoolFace