datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ztl-sard-php-verdicts
Source and recipe: github.com/inventor1975/introspect — dataset/sard (commit 4e179ba). The corpus is synthetic (NIST/Stivalet generated test cases); no real project's code or vulnerability is in this dataset.
ZTL verdicts over the SARD / Stivalet PHP vulnerability suite
This dataset records what the introspect analyzer (the ZTL zero-trust judge over
a deterministic PHP atomizer) returns on the public NIST SARD / Stivalet PHP test
suite — one row per test file, the verdict and… See the full description on the dataset page: https://huggingface.co/datasets/Inventor1975/ztl-sard-php-verdicts.gemma3_12b-it_datasetThis dataset was specifically curated and refined for use with Gemma 3, with the goal of enhancing its instruction-following capabilities through LoRA (Low-Rank Adaptation) and refine-tuning techniques. While the original Alpaca dataset contained various issues such as hallucinated outputs, incorrect answers, inconsistent formats, and ambiguous instructions, this version has been cleaned and improved to ensure higher quality, more reliable training data.
By integrating this updated dataset… See the full description on the dataset page: https://huggingface.co/datasets/sardor233/gemma3_12b-it_dataset.
