MU-NLPC/pdt_anaphora_czech
Dataset Card for pdt_anaphora_czech This dataset is used for my thesis to fine-tune language models on Czech unstructured text for anaphora resolution. Dataset Sources The dataset is based on data from the Prague Dependency Treebank, specifically the PDTC 1.0 (https://lindat.mff.cuni.cz/repository/xmlui/handle/11234/1-3185) How to cite Stano P. and Horák A. Evaluating Prompt-Based and Fine-Tuned Approaches to Czech Anaphora Resolution.… See the full description on the dataset page: https://huggingface.co/datasets/MU-NLPC/pdt_anaphora_czech.
045
