yeshpanovrustem/kaznerd
A Named Entity Recognition Dataset for Kazakh This is a modified version of the dataset provided in the LREC 2022 paper KazNERD: Kazakh Named Entity Recognition Dataset. The original repository for the paper can be found at https://github.com/IS2AI/KazNERD. Tokens denoting speech disfluencies and hesitations (parenthesised) and background noise [bracketed] were removed. A total of 2,027 duplicate sentences were removed. Statistics for training (Train), validation… See the full description on the dataset page: https://huggingface.co/datasets/yeshpanovrustem/kaznerd.
9118
A Named Entity Recognition Dataset for Kazakh
- This is a modified version of the dataset provided in the LREC 2022 paper *KazNERD: Kazakh Named Entity Recognition Dataset*.
- The original repository for the paper can be found at https://github.com/IS2AI/KazNERD.
- Tokens denoting speech disfluencies and hesitations (parenthesised) and background noise [bracketed] were removed.
- A total of 2,027 duplicate sentences were removed.
Dataset Description
- Original repository: https://github.com/IS2AI/KazNERD
- Original paper: https://aclanthology.org/2022.lrec-1.44
- Point of contact: Rustem Yeshpanov
