CoolFace
Datasetpublicgated

abhinandansamal/odia-german-parallel-corpus-research

Dataset Summary This dataset is a high-quality, parallel corpus for Odia (Oriya) to German and German to Odia machine translation. It focuses on the news domain, specifically covering National, International, Sports, Trade, and Science & Technology topics. The dataset contains 3,676 unique parallel sentence pairs, curated through a hybrid approach combining automated web scraping, manual human translation (Gold Standard), and human-corrected machine translation (Silver… See the full description on the dataset page: https://huggingface.co/datasets/abhinandansamal/odia-german-parallel-corpus-research.

sourceHugging Faceapache-2.0updated 9mo agoView on Hugging Face
0likes3downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.