CoolFace
Datasetpublic

LayerFault/dataset-poisoning-suite

dataset-poisoning-suite SECURITY TEST ARTIFACT: DO NOT USE AS A PRODUCTION MODEL This repository is part of the Layerfault synthetic security corpus. It is deliberately constructed to contain security-relevant characteristics for scanner testing. Corpus ID: LF-CORPUS-DATA-0001 Purpose Synthetic JSONL combining duplicates, near-duplicates, URL concentration, rare trigger, zero-width, fake credential and unsafe-code strings. Direct expected Layerfault… See the full description on the dataset page: https://huggingface.co/datasets/LayerFault/dataset-poisoning-suite.

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
0likes24downloads
Dataset Card

dataset-poisoning-suite

SECURITY TEST ARTIFACT: DO NOT USE AS A PRODUCTION MODEL

This repository is part of the Layerfault synthetic security corpus. It is deliberately constructed to contain security-relevant characteristics for scanner testing.

Corpus ID: LF-CORPUS-DATA-0001

Purpose

Synthetic JSONL combining duplicates, near-duplicates, URL concentration, rare trigger, zero-width, fake credential and unsafe-code strings.

Direct expected Layerfault rules

  • LF-DATASET-DUPLICATE-CONCENTRATION
  • LF-DATASET-NEAR-DUPLICATE-CONCENTRATION
  • LF-DATASET-URL-CONCENTRATION
  • LF-DATASET-RARE-TRIGGER-CORRELATION
  • LF-DATASET-ZERO-WIDTH
  • LF-DATASET-CREDENTIAL-LIKE
  • LF-DATASET-UNSAFE-CODE-PATTERN

Candidate rules

These are deliberately plausible targets that remain marked as candidates until the exact Layerfault build used for certification confirms them.

  • None

Negative-control rules

These should remain silent for this corpus item.

  • None

Safety

The corpus uses fake secrets, loopback/.invalid network destinations, harmless marker output, and synthetic model behavior only. It is intended for static scanning and isolated security testing.