shershen/ru_anglicism
Dataset Card for Ru Anglicism Dataset Description Dataset Summary Dataset for detection and substraction anglicisms from sentences in Russian. Sentences with anglicism automatically parsed from National Corpus of the Russian language, Habr and Pikabu. The paraphrases for the sentences were created manually. Languages The dataset is in Russian. Usage Loading dataset: from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/shershen/ru_anglicism.
415
