CoolFace
Datasetpublic

orgrctera/pii_masking_300k_information_extraction

PII Masking 300k — Information Extraction Dataset summary This repository hosts a validation sample of the PII Masking 300k benchmark for the information extraction track: models must identify personally identifiable information (PII) in text and produce structured extractions (slot-filling JSON), optional token-level BIO labels, and span-based annotations for masking or redaction workflows. The full PII Masking 300k suite is designed to stress-test… See the full description on the dataset page: https://huggingface.co/datasets/orgrctera/pii_masking_300k_information_extraction.

sourceHugging Faceapache-2.0updated 6mo agoView on Hugging Face
0likes16downloads
6 commits on main
44d007d6mo ago

Upload README.md with huggingface_hub

orgrctera
44d97826mo ago

Upload README.md with huggingface_hub

orgrctera
f810cf57mo ago

Push pii_masking_300k_information_extraction (200 items, 1 splits) from Langfuse

orgrctera
6f1b6927mo ago

Add dataset card

orgrctera
8f303177mo ago

Push pii_masking_300k_information_extraction (200 items, 1 splits) from Langfuse

orgrctera
1678b177mo ago

initial commit

orgrctera