CoolFace
Datasetpublic

shershen/ru_anglicism

Dataset Card for Ru Anglicism Dataset Description Dataset Summary Dataset for detection and substraction anglicisms from sentences in Russian. Sentences with anglicism automatically parsed from National Corpus of the Russian language, Habr and Pikabu. The paraphrases for the sentences were created manually. Languages The dataset is in Russian. Usage Loading dataset: from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/shershen/ru_anglicism.

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
4likes15downloads

shershen/ru_anglicism · main · files are served by the source, never re-hosted here