tomaarsen/conll2003
Dataset Card for "conll2003" Dataset Summary The shared task of CoNLL-2003 concerns language-independent named entity recognition. We will concentrate on four types of named entities: persons, locations, organizations and names of miscellaneous entities that do not belong to the previous three groups. The CoNLL-2003 shared task data files contain four columns separated by a single space. Each word has been put on a separate line and there is an empty line after… See the full description on the dataset page: https://huggingface.co/datasets/tomaarsen/conll2003.
Delete conll2003.py
Upload dataset
Merge branch 'main' of https://huggingface.co/datasets/tomaarsen/conll2003
initial commit
Add 'document_id' and 'sentence_id' columns
Convert dataset sizes from base 2 to base 10 in the dataset card (#6)
Replace YAML keys from int to str (#5)
Reorder split names (#3)
add dataset_info in dataset metadata
remove dummmy data
Fix: conll2003 - fix empty example (#4662)
Align more metadata with other repo types (models,spaces) (#4607)
Eval metadata batch 1: BillSum, CoNLL2003, CoNLLPP, CUAD, Emotion, GigaWord, GLUE, Hate Speech 18, Hate Speech (#4335)
Update datasets task tags to align tags with models (#4067)
Update files from the datasets library (from 1.18.0)
Update files from the datasets library (from 1.16.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.5.0)
Update files from the datasets library (from 1.4.0)
Update files from the datasets library (from 1.3.0)
Update files from the datasets library (from 1.1.3)
Update files from the datasets library (from 1.0.2)
