shershen/ru_anglicism
Dataset Card for Ru Anglicism Dataset Description Dataset Summary Dataset for detection and substraction anglicisms from sentences in Russian. Sentences with anglicism automatically parsed from National Corpus of the Russian language, Habr and Pikabu. The paraphrases for the sentences were created manually. Languages The dataset is in Russian. Usage Loading dataset: from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/shershen/ru_anglicism.
Update ru_anglicism.py
Upload dataset_infos.json
add files
Delete data/train.jsonl
Delete data/test.jsonl
Update README.md
Update README.md
Upload ru_anglicism.py
add files
initial commit
