CoolFace
20 results

redaction

RedactionBench /RedactionBench Dataset Card for RedactionBench RedactionBench is an evaluation-only benchmark for character-level redaction across eleven document categories. Each of the 200 documents is manually-annotated with character spans that are either mandatory (must redact) or contextual. RedactionBench mixes 101 real-world documents manually sourced from the public web (transcribed, augmented) with 99 synthetic documents authored to fill categories where synthetic data is more appropriate. The above is… See the full description on the dataset page: https://huggingface.co/datasets/RedactionBench/RedactionBench.texttoken-classificationn<1K2 likes291 downloads4mo agoHugging FaceA10Networks /RedactionBench Dataset Card for RedactionBench RedactionBench is an evaluation-only benchmark for character-level redaction across eleven document categories. Each of the 200 documents is manually-annotated with character spans that are either mandatory (must redact) or contextual. RedactionBench mixes 101 real-world documents manually sourced from the public web (transcribed, augmented) with 99 synthetic documents authored to fill categories where synthetic data is more appropriate. The… See the full description on the dataset page: https://huggingface.co/datasets/A10Networks/RedactionBench.texttoken-classificationn<1K0 likes111 downloads3mo agoHugging Facermems /log-redaction-trajectories Log Redaction Trajectories Rights & intended use: legacy public research corpus / portfolio artifact. Hosted frontier-model outputs are research-only inputs under project policy (synthetic-factory#161): intended_use: research_only, project_training_policy: blocked. Not training data for any model-weight update. Machine-readable record: rights.json. Release status: The raw, uncurated payload is now published under data/raw/. It is available for inspection and reproducibility… See the full description on the dataset page: https://huggingface.co/datasets/rmems/log-redaction-trajectories.textn<1K0 likes70 downloads21d agoHugging Faceaibotjock /RedactionBench Dataset Card for RedactionBench RedactionBench is an evaluation-only benchmark for character-level redaction across eleven document categories. Each of the 200 documents is manually-annotated with character spans that are either mandatory (must redact) or contextual. RedactionBench mixes 101 real-world documents manually sourced from the public web (transcribed, augmented) with 99 synthetic documents authored to fill categories where synthetic data is more appropriate. The… See the full description on the dataset page: https://huggingface.co/datasets/aibotjock/RedactionBench.texttoken-classificationn<1K0 likes48 downloads3mo agoHugging FaceKing-Harry /NinjaMasker-PII-Redactiontext10K<n<100K2 likes35 downloads3y agoHugging FaceKing-Harry /NinjaMasker-PII-Redaction-Datasettext10K<n<100K1 likes30 downloads3y agoHugging Face