datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
RedactionBench
Dataset Card for RedactionBench
RedactionBench is an evaluation-only benchmark for character-level redaction across eleven document categories. Each of the 200 documents is manually-annotated with character spans that are either mandatory (must redact) or contextual.
RedactionBench mixes 101 real-world documents manually sourced from the public web (transcribed, augmented) with 99 synthetic documents authored to fill categories where synthetic data is more appropriate.
The above is… See the full description on the dataset page: https://huggingface.co/datasets/RedactionBench/RedactionBench.RedactionBench
Dataset Card for RedactionBench
RedactionBench is an evaluation-only benchmark for character-level redaction across eleven document categories. Each of the 200 documents is manually-annotated with character spans that are either mandatory (must redact) or contextual.
RedactionBench mixes 101 real-world documents manually sourced from the public web (transcribed, augmented) with 99 synthetic documents authored to fill categories where synthetic data is more appropriate.
The… See the full description on the dataset page: https://huggingface.co/datasets/A10Networks/RedactionBench.log-redaction-trajectories
Log Redaction Trajectories
Rights & intended use: legacy public research corpus / portfolio
artifact. Hosted frontier-model outputs are research-only inputs under
project policy (synthetic-factory#161):
intended_use: research_only, project_training_policy: blocked. Not
training data for any model-weight update. Machine-readable record:
rights.json.
Release status: The raw, uncurated payload is now published under
data/raw/. It is available for inspection and reproducibility… See the full description on the dataset page: https://huggingface.co/datasets/rmems/log-redaction-trajectories.RedactionBench
Dataset Card for RedactionBench
RedactionBench is an evaluation-only benchmark for character-level redaction across eleven document categories. Each of the 200 documents is manually-annotated with character spans that are either mandatory (must redact) or contextual.
RedactionBench mixes 101 real-world documents manually sourced from the public web (transcribed, augmented) with 99 synthetic documents authored to fill categories where synthetic data is more appropriate.
The… See the full description on the dataset page: https://huggingface.co/datasets/aibotjock/RedactionBench.NinjaMasker-PII-RedactionDocPII-redaction-benchmark
DocPII: Contextual Redaction Benchmark Dataset
Dataset Description
DocPII contains 1101 high-quality document samples enriched with embedded personally identifiable information (PII). Designed to evaluate context-aware redaction systems, it provides realistic, full-document contexts—a notable advancement over sentence-level datasets.
All documents have been manually reviewed for accuracy, coherence, and redaction alignment, ensuring data quality for benchmarking and… See the full description on the dataset page: https://huggingface.co/datasets/nutrientdocs/DocPII-redaction-benchmark.NinjaMasker-PII-Redaction-Datasetlegal-redaction-necessity-scope-consistency-coherence-v0.1What this dataset does
You receive
doc context
sensitive elements
proposed redactions
justification tags
duplicate signals
over or under signals
You decide
coherent
or
incoherent
Daily use
redaction QC
over redaction flag
under redaction flag
duplicate consistency check
Clinical_PII_Redaction_Testllama-redaction-instructionllama-redactionlegal-disclosure-tagging-relevance-privilege-redaction-coherence-risk-v0.1What this dataset does
You receive
doc summary
issue list
tag
privilege basis
redaction rationale
rule consistency notes
You decide
coherent
or
incoherent
Daily use
batch tagging QC
privilege basis checking
redaction logic consistency
redaction_keyphi-redaction-sftPII-redaction-bertredaction-demo
