CoolFace
Datasetpublic

CZLC/CNC_skript12

Introduction This is the SKRIPT2012 dataset, maintained by the Czech National Corpus project. This dataset corresponds to the version available in the LINDAT repository, where it is named AKCES-1. The dataset was created from public .rtf and .doc file formats using the convert_AKCES.py script. About Original Dataset (Taken from project Wiki). The Corpus SKRIPT2012 is a learner corpus aimed at representing the written language of Czech pupils and students at… See the full description on the dataset page: https://huggingface.co/datasets/CZLC/CNC_skript12.

sourceHugging Facecc-by-nc-nd-3.0updated 2y agoView on Hugging Face
0likes17downloads
6 commits on main
bc7e94f2y ago

Update README.md

mfajcik
531c0652y ago

Update README.md

mfajcik
c75f7fc2y ago

Update README.md

mfajcik
f26bfe72y ago

Upload test.jsonl

mfajcik
8d80dc02y ago

Upload 2 files

mfajcik
f96ffea2y ago

initial commit

mfajcik