CoolFace
Datasetpublic

ai4privacy/pii-masking-openpii-1m

OpenPII 1M — Multilingual PII Masking Dataset Overview The OpenPII 1M dataset is a large-scale, multilingual collection of 1,428,143 synthetic text examples with fine-grained PII (Personally Identifiable Information) annotations, spanning 23 European languages and 19 entity types. Built to advance open research in privacy-preserving NLP, this dataset enables the development and benchmarking of Named Entity Recognition (NER) models, token classification… See the full description on the dataset page: https://huggingface.co/datasets/ai4privacy/pii-masking-openpii-1m.

sourceHugging Faceotherupdated 6mo agoView on Hugging Face
15likes1.6kdownloads

ai4privacy/pii-masking-openpii-1m · main · files are served by the source, never re-hosted here