disi-unibo-nlp/physionet-deid-i2b2-2014
The De-identification dataset contains medical text records with Named Entity Recognition (NER) annotations. The dataset is processed to split records into individual sentences while preserving entity annotations. Each sentence is tokenized and annotated in IOB format for training NER models.
3245
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face