CoolFace
Datasetpublic

MU-NLPC/pdt_anaphora_czech

Dataset Card for pdt_anaphora_czech This dataset is used for my thesis to fine-tune language models on Czech unstructured text for anaphora resolution. Dataset Sources The dataset is based on data from the Prague Dependency Treebank, specifically the PDTC 1.0 (https://lindat.mff.cuni.cz/repository/xmlui/handle/11234/1-3185) How to cite Stano P. and Horák A. Evaluating Prompt-Based and Fine-Tuned Approaches to Czech Anaphora Resolution.… See the full description on the dataset page: https://huggingface.co/datasets/MU-NLPC/pdt_anaphora_czech.

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes47downloads
Dataset Card

Dataset Card for pdtanaphoraczech

This dataset is used for my thesis to fine-tune language models on Czech unstructured text for anaphora resolution.

Dataset Sources

The dataset is based on data from the Prague Dependency Treebank, specifically the PDTC 1.0 (https://lindat.mff.cuni.cz/repository/xmlui/handle/11234/1-3185)

How to cite

Stano P. and Horák A. Evaluating Prompt-Based and Fine-Tuned Approaches to Czech Anaphora Resolution. International Conference on Text, Speech, and Dialogue, 2025.

@article{stano_horak2025, author = "Patrik Stano and Aleš Horák", title = "Evaluating Prompt-Based and Fine-Tuned Approaches to Czech Anaphora Resolution", journal= "International Conference on Text, Speech, and Dialogue", year = "2025", }