datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PyResBugs
PyResBugs
PyResBugs is a curated dataset containing 5007 residual Python bugs, paired with their corresponding fixed versions and multi-level natural language (NL) descriptions. It is the first dataset designed specifically for natural language-driven fault injection, enabling advanced research in software testing and automated fault analysis.
Description
Residual bugs are defects that remain undetected during traditional testing but surface later in production.… See the full description on the dataset page: https://huggingface.co/datasets/OSS-forge/PyResBugs.forge-industrial-control-scenarios
Forge Industrial Control and Telemetry Traces
Deterministic synthetic traces spanning device ingress, signature/quality/range
failures, offline store-and-forward, local inference, sequential agent review,
L0–L4 policy outcomes, electrolyser ramp sequences and multivariate telemetry
anomalies.
No operational plant data, customer data, secrets or real equipment identifiers are included. machine.press-03 and every measurement are fictitious.
Files… See the full description on the dataset page: https://huggingface.co/datasets/sankalpsthakur/forge-industrial-control-scenarios.Extended_Shellcode_IA32
Shellcode_IA32
Shellcode_IA32 is a dataset containing more than 20 years of shellcodes from a variety of sources and is the largest collection of shellcodes in assembly available to date. We are currently extending the dataset. Up to now, we released three versions of the dataset.
Shellcode_IA32 was presented for the first time in the paper Shellcode_IA32: A Dataset for Automatic Shellcode Generation, accepted to the 1st Workshop on Natural Language Processing for Programming… See the full description on the dataset page: https://huggingface.co/datasets/OSS-forge/Extended_Shellcode_IA32.forge-pump-digital-twin-synthetic
Forge Pump Digital Twin — Synthetic
A deterministic, clean-room tabular baseline for pump surrogate modelling, edge-runtime conformance, and advisory anomaly examples. It contains 40,000 synthetic rows split into 28,000 train, 6,000 validation, and 6,000 test rows.
This dataset contains no plant telemetry, customer data, equipment identifiers, CAD, BOMs, nameplates, vendor curves, or values copied from a private repository. Every constant is an illustrative engineering proxy. It… See the full description on the dataset page: https://huggingface.co/datasets/sankalpsthakur/forge-pump-digital-twin-synthetic.sequential-forgetting-benchmark
Sequential Forgetting Benchmark
What does sequential fine-tuning do to what a model already learned? This
dataset is a results ledger with receipts: every row of
results/results.csv links to the raw run file it came from
(results/raw/), every transcription is hand-checked (results/PROVENANCE.md),
and invalid runs are disclosed, not deleted.
Seeded from ModelBrew's archival continual-learning runs (2026). Community
submissions welcome — see protocol/PROTOCOL.md.… See the full description on the dataset page: https://huggingface.co/datasets/ModelBrew/sequential-forgetting-benchmark.MtG-json-to-ForgeScriptthe-forge-wiki-calculator
The Forge Calculator - Your Essential Roblox Forging Companion
Experience unparalleled accuracy with forge calculator designed for The Forge Roblox players
🎯 Introducing The Forge Calculator
The Forge Calculator stands as the premier and most comprehensive tool available for The Forge Roblox enthusiasts, meticulously designed to transform how you approach weapon and armor creation. This forge calculator offers exceptional accuracy when forecasting forging results… See the full description on the dataset page: https://huggingface.co/datasets/zhugetd/the-forge-wiki-calculator.Math-Forge-Hard
Math-Forge-Hard Dataset
Overview
The Math-Forge-Hard dataset is a collection of challenging math problems designed to test and improve problem-solving skills. This dataset includes a variety of word problems that cover different mathematical concepts, making it a valuable resource for students, educators, and researchers.
Dataset Details
Modalities
Text: The dataset primarily contains text data, including math word problems.
Formats… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Math-Forge-Hard.coherence-gap-observations
Coherence Gap Observations
A qualitative observational dataset documenting coherence gaps in sustained human–AI interaction: moments when a system’s behavior, functional self-description, monitoring, or continuity appears inconsistent with its subsequent framing or denial.
Dataset contents
The dataset contains 28 coded observations across several AI platforms and interaction contexts. Each row records:
platform and model version
memory and conversational-window… See the full description on the dataset page: https://huggingface.co/datasets/Flame-Forged/coherence-gap-observations.for-gemma
