CZLC/CNC_skript12
Introduction This is the SKRIPT2012 dataset, maintained by the Czech National Corpus project. This dataset corresponds to the version available in the LINDAT repository, where it is named AKCES-1. The dataset was created from public .rtf and .doc file formats using the convert_AKCES.py script. About Original Dataset (Taken from project Wiki). The Corpus SKRIPT2012 is a learner corpus aimed at representing the written language of Czech pupils and students at… See the full description on the dataset page: https://huggingface.co/datasets/CZLC/CNC_skript12.
017
Update README.md
Update README.md
Update README.md
Upload test.jsonl
Upload 2 files
initial commit
