CoolFace
Datasetpublic

stighellemans/meddeid-dutch-synthetic-corpus

MedDeID Dutch synthetic corpus This repository contains 6,493 synthetic Dutch clinical documents with character-offset de-identification spans. It contains no real patient notes or personal information. The corpus was used to train meddeid-dutch-synth. Split policy All 6,493 records are exposed together through one conventional Hugging Face train split. There is no publisher-defined validation split. Here train means the complete model-development corpus; users… See the full description on the dataset page: https://huggingface.co/datasets/stighellemans/meddeid-dutch-synthetic-corpus.

sourceHugging Facecc-by-4.0updated 12d agoView on Hugging Face
0likes116downloads
settings

This repository belongs to stighellemans on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namemeddeid-dutch-synthetic-corpus
visibilitypublic
licencecc-by-4.0
gatedno
ownerstighellemans
Account settings
stighellemans/meddeid-dutch-synthetic-corpus · CoolFace