CoolFace
Datasetpublic

ZBWatHF/Funder-NER

Dataset Card for Dataset Named Entity Recognition of funders of scientific research Dataset Summary Training/test set for automatically identifying funder entities mentioned in scientific papers. This data set is generated from Open Access documents hosted at https://econstor.eu and manually curated/labeled. Supported Tasks and Leaderboards The dataset is for training and testing the automatic recognition of funders as they are acknowledged in… See the full description on the dataset page: https://huggingface.co/datasets/ZBWatHF/Funder-NER.

sourceHugging Facecc-by-nc-sa-4.0updated 3y agoView on Hugging Face
0likes6downloads
Dataset Card

Dataset Card for Dataset Named Entity Recognition of funders of scientific research

Dataset Description

  • Homepage:https://econstor.eu
  • Repository:https://github.com/zbw/Funder-NER
  • Paper: https://doi.org/10.1007/978-3-031-16802-4_24
  • Leaderboard:
  • Point of Contact:

Dataset Summary

Training/test set for automatically identifying funder entities mentioned in scientific papers. This data set is generated from Open Access documents hosted at https://econstor.eu and manually curated/labeled.

Supported Tasks and Leaderboards

The dataset is for training and testing the automatic recognition of funders as they are acknowledged in scientific papers.

Languages

English

Dataset Structure

Data Instances

[More Information Needed]

Data Fields

"handle: the PID of the OA working paper" "crossrefdoi: the PID of the corresponding publication (article), as registered via CrossRef" "crossrefphrase: the funder according to CrossRef metadata" "pdf_phrase: the acknowledgement phrase from the paper" "funder: Y(es) in case there is at least one funder, N(o) in case there is no funder supporting the work, but OA publishing funding"

Data Splits

[More Information Needed]

Dataset Creation

Curation Rationale

The dataset was manually curated to train the NER recognition based on a QA system.

Source Data

Initial Data Collection and Normalization

[More Information Needed]

Who are the source language producers?

[More Information Needed]

Annotations

Annotation process

[More Information Needed]

Who are the annotators?

[More Information Needed]

Personal and Sensitive Information

[More Information Needed]

Considerations for Using the Data

Social Impact of Dataset

[More Information Needed]

Discussion of Biases

[More Information Needed]

Other Known Limitations

[More Information Needed]

Additional Information

Dataset Curators

[More Information Needed]

Licensing Information

[More Information Needed]

Citation Information

[More Information Needed]

Contributions

[More Information Needed]