zkml
Datasets
All datasets matching “zkml”ZKML-Fintech-Instruct-Alphazkml-github-repos-truncatedThis dataset is a truncated version of this one but where the format is compatible with MLX-lora,
using {"text": "This is an example for the model."}, and where each entry has been truncated, following some code logic (i.e., following classes, functions etc) to ensure
each entry is smaller than 2048 tokens.
zkml-audit-benchmark
zkml-audit-benchmark
A benchmark dataset for evaluating AI agents on zkML soundness auditing: finding cryptographic vulnerabilities in zero-knowledge machine learning proof implementations.
Overview
This dataset pairs 4 published zkML research papers with their corresponding frozen codebase snapshots and 56 bug artifacts (20 real-world from expert audits + 36 synthetic for broader coverage). Each artifact describes a single soundness vulnerability — the code edits to… See the full description on the dataset page: https://huggingface.co/datasets/Anonymous648/zkml-audit-benchmark.zkml-github-reposThis dataset contains the code from all the ZKML repos that I'm aware of that have an MIT, Apache or GPL 3.0 license.
It only contains the files with extensions: ".py", ".js", ".java", ".c", ".cpp", ".h", ".hpp", ".rs", "cairo", ".zkey", ".sol", ".circom", ".ejs", ".ipynb"
List of repos:
"https://github.com/gizatechxyz/orion",
"https://github.com/gizatechxyz/Giza-Hub",
"https://github.com/zkonduit/ezkl",
"https://github.com/socathie/keras2circom"… See the full description on the dataset page: https://huggingface.co/datasets/sa8/zkml-github-repos.zkml-audit-benchmark
zkml-audit-benchmark
A benchmark dataset for evaluating AI agents on zkML soundness auditing: finding cryptographic vulnerabilities in zero-knowledge machine learning proof implementations.
Overview
This dataset pairs 4 published zkML research papers with their corresponding frozen codebase snapshots and 56 expert-authored bug artifacts. Each artifact describes a single soundness vulnerability — the code edits to inject it, ground-truth labels for scoring, and presence… See the full description on the dataset page: https://huggingface.co/datasets/Netzerep/zkml-audit-benchmark.
