datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SciERCSCIERC (Luan et al., 2018) via "Don’t Stop Pretraining: Adapt Language Models to Domains and Tasks" (Gururangan et al., 2020) reuploaded because of error encountered when trying to load zj88zj/SCIERC with the huggingfaces/datasets library.
sciercsci-ercSCIERCSCIERC-PT
SCIERC-PT
Overview
SCIERC-PT is a European Portuguese (PT-PT) translation of the SCIERC dataset, created to support research on Scientific Information Extraction in Portuguese. The dataset provides translated scientific abstracts suitable for evaluating Named Entity Recognition (NER) and Relation Extraction (RE) models while preserving the original annotation schema.
The dataset was automatically translated and subsequently curated to improve alignment between the… See the full description on the dataset page: https://huggingface.co/datasets/amalia-llm/SCIERC-PT.method_only_scierc
