evalitahf/entity_recognition
Data for the NERMuD shared task (Evalita 2023) This data is the one used for the NERMuD shared task organized at Evalita 2023. The dataset contains the Wikinews, fiction, and De Gasperi subsets of KIND, where test data is used for development. Content of the dataset Split Sentences wn_train 10,912 wn_dev 2,594 wn_test 2,088 fic_train 11,423 fic_dev 1,051 fic_test 1,517 adg_train 5,147 adg_dev 1,122 adg_test 521 Set Sentences… See the full description on the dataset page: https://huggingface.co/datasets/evalitahf/entity_recognition.
Add 3 files from dev set for few-shot learning
Add 3 files for NER task with diverse examples for few-shot learning
feat: added reduced version of dataset
feat: initial commit
initial commit
