entity recognition
named-entity-recognition-nerkor-hubert-hungarianSupersaiyan1729_-_financeLM_outputpath_Named_Entity_Recognition__25_gpt2small-ggufNamed-entity-recognitionnamed-entity-recognitionUnBIAS-Named-Entity-Recognitionconflibert-named-entity-recognitionxlm-roberta-name-entity-recognition-japanesebert-finetuned-named-entity-recognition-ner
named_entity_recognition_document_contexttask960_ancora-ca-ner_named_entity_recognition
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task960_ancora-ca-ner_named_entity_recognition
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task960_ancora-ca-ner_named_entity_recognition.flan_combined_task1544_conll2002_named_entity_recognition_answer_generationamharic-named-entity-recognition
Amharic Named Entity Recognition Dataset
This dataset can be used to train models for Named Entity Recognition.
Dataset Source
https://github.com/uhh-lt/ethiopicmodels/blob/master/am/data/NER/train.txt
Finetuned Models
The following transformer models were finetuned using this dataset. The reported precision, recall, and f1 metrics are macro averages.
Model
Size (# params)
Precision
Recall
F1
bert-medium-amharic
40.5M
0.64
0.73
0.68… See the full description on the dataset page: https://huggingface.co/datasets/rasyosef/amharic-named-entity-recognition.entity_recognition
Data for the NERMuD shared task (Evalita 2023)
This data is the one used for the NERMuD shared task organized
at Evalita 2023.
The dataset contains the Wikinews, fiction, and De Gasperi subsets of KIND, where test data is used for development.
Content of the dataset
Split
Sentences
wn_train
10,912
wn_dev
2,594
wn_test
2,088
fic_train11,423
fic_dev
1,051
fic_test
1,517
adg_train
5,147
adg_dev
1,122
adg_test
521
Set
Sentences… See the full description on the dataset page: https://huggingface.co/datasets/evalitahf/entity_recognition.sarvam-entity-recognition-gemini-2.0-flash-thinking-01-21-distill-1600Dataset for sarvam's entity normalisation task. More detailed information can be found here, in the main model repo: Hugging Face
Detailed Report (Writeup): Google Drive
It also has a gguf variant, with certain additional gguf based innstructions: Hugging Face
Model inference script can be found here: Colab
Model predictions can be found in this dataset and both the repo files. named as:
eval_data_001_predictions.csv and eval_data_001_predictions_excel.csv.
train_data_001_predictions.csvand… See the full description on the dataset page: https://huggingface.co/datasets/Tasmay-Tib/sarvam-entity-recognition-gemini-2.0-flash-thinking-01-21-distill-1600.
