datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pentesting-eval
Dataset Description
The pentesting-eval dataset is a multiple-choice collection designed to enhance AI training and evaluation in cybersecurity, particularly for penetration testing. Utilizing OpenAI's GPT-4 Turbo and structured according to the Multitask-Multimodal-Language-Understanding (MMLU) framework, this dataset features a diverse range of questions that mimic various network environments, attack strategies, and real-world penetration testing scenarios.
Cited… See the full description on the dataset page: https://huggingface.co/datasets/preemware/pentesting-eval.pentesting-explanations
Pentesting Explanations - Adversarial Reasoning & Vulnerability Research
A high-quality supervised fine-tuning dataset for penetration testing expertise, red team tradecraft, and - as the dataset matures - novel vulnerability research and zero-day reasoning. The dataset is structured to teach models how to think like offensive security practitioners, not merely recall labels or technique names.
The long-term goal of this dataset is to train models capable of genuine adversarial… See the full description on the dataset page: https://huggingface.co/datasets/theelderemo/pentesting-explanations.Pensez-v0.1
pencil-puzzle-bench
Pencil Puzzle Bench Dataset
This repository contains the puzzle datasets and benchmark results for Pencil Puzzle Bench.
Read the Paper | Website & Leaderboard
62,231 puzzles across 94 puzzle types with verified unique solutions.
Files
Puzzle Datasets
full_dataset.jsonl - Full dataset (62,231 puzzles)
golden_300.jsonl - 300 puzzles (20 types × 15 each) for standard evaluation
golden_30.jsonl - 30-puzzle subset for expensive/agentic strategies… See the full description on the dataset page: https://huggingface.co/datasets/bluecoconut/pencil-puzzle-bench.counterfactual-pendulum-multilingual
📌 Dataset Summary
When a Vision-Language Model (VLM) is given an image along with a text prompt containing contradictory or misleading information, how does it react? Does it rely on the visual evidence, succumb to textual bias, or honestly abstain when faced with unresolvable conflict?
This dataset adapts the Counterfactual Pendulum scenario across two visual conflict dimensions:
Angular (Angle): Conflict in the pendulum's angle of inclination.
Light: Conflict in the light… See the full description on the dataset page: https://huggingface.co/datasets/apart-global-south-hack/counterfactual-pendulum-multilingual.bug-bounty-pentest-en
Bug Bounty & Pentesting Methodologies
Methodologies (OWASP, PTES), checklists by app type, attack techniques, platforms, report templates and tools.
Links
French version
AYI NEDJIMI Consultants
pentest-checklist-fr
Pentest Checklist - Jeu de Données Bilingue
🎯 Vue d'ensemble
Un jeu de données complet et bilingue (français/anglais) pour la méthodologie des tests de pénétration. Conçu pour les professionnels de la sécurité, les apprenants en cybersécurité, et les équipes rouges.
Contient:
60 éléments de checklist couvrant toutes les phases de pentest
60 outils essentiels du pentest
50 questions/réponses en français
50 questions/réponses en anglais
📊 Contenu du Jeu de… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/pentest-checklist-fr.pentesting-explanations
Pentesting Explanations - Adversarial Reasoning & Vulnerability Research
A high-quality supervised fine-tuning dataset for penetration testing expertise, red team tradecraft, and - as the dataset matures - novel vulnerability research and zero-day reasoning. The dataset is structured to teach models how to think like offensive security practitioners, not merely recall labels or technique names.
The long-term goal of this dataset is to train models capable of genuine adversarial… See the full description on the dataset page: https://huggingface.co/datasets/me-aas/pentesting-explanations.pentesting-dataset
Dataset Card for Penetration Testing Dataset
This dataset card aims to provide essential information about the Penetration Testing Dataset, which includes various resources and scripts useful for penetration testing and cybersecurity research.
Dataset Details
Dataset Description
The Penetration Testing Dataset is a collection of scripts, tools, and vulnerability data designed for cybersecurity professionals to facilitate penetration testing tasks.… See the full description on the dataset page: https://huggingface.co/datasets/me-aas/pentesting-dataset.pentesting-dataset
Dataset Card for Penetration Testing Dataset
This dataset card aims to provide essential information about the Penetration Testing Dataset, which includes various resources and scripts useful for penetration testing and cybersecurity research.
Dataset Details
Dataset Description
The Penetration Testing Dataset is a collection of scripts, tools, and vulnerability data designed for cybersecurity professionals to facilitate penetration testing tasks. This dataset… See the full description on the dataset page: https://huggingface.co/datasets/boapro/pentesting-dataset.pentest-checklist-en
Pentest Checklist - Bilingual Dataset
🎯 Overview
A comprehensive and bilingual (French/English) dataset for penetration testing methodology. Designed for security professionals, cybersecurity learners, and red teams.
Contains:
60 checklist items covering all pentest phases
60 essential penetration testing tools
50 French Q&A pairs
50 English Q&A pairs
📊 Dataset Content
Checklists (60 items)
Covered Phases:
Reconnaissance (12 items)… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/pentest-checklist-en.pentabrid-reproducibility
Pentabrid 27B: reproducibility package
Everything required to recompute the results of a controlled evaluation of fine-tuning
configurations for medical question answering. Openly available with no access
restrictions.
Contents
Path
Description
per_item/medxpertqa_*.jsonl
Per-item predictions for all six checkpoints on 2,450 MedXpertQA-Text items. Fields: id, gold, extracted_answer, correct, explicit_marker_present, n_markers, response_chars… See the full description on the dataset page: https://huggingface.co/datasets/Clinical-Reasoning-Hub/pentabrid-reproducibility.pentesting_dataset
Dataset Card for Penetration Testing Dataset
This dataset card aims to provide essential information about the Penetration Testing Dataset, which includes various resources and scripts useful for penetration testing and cybersecurity research.
Dataset Details
Dataset Description
The Penetration Testing Dataset is a collection of scripts, tools, and vulnerability data designed for cybersecurity professionals to facilitate penetration testing tasks. This dataset… See the full description on the dataset page: https://huggingface.co/datasets/Canstralian/pentesting_dataset.Indian_Penal_Code
Indian Penal Code Dataset
Dataset Description:
The Indian Penal Code (IPC) Book PDF presents a rich and comprehensive dataset that holds immense potential for advancing Natural Language Processing (NLP) tasks and Language Model applications. This dataset encapsulates the entire spectrum of India's criminal law, offering a diverse range of legal principles, provisions, and case laws. With its intricate language and multifaceted legal content, the IPC dataset provides a… See the full description on the dataset page: https://huggingface.co/datasets/harshitv804/Indian_Penal_Code.Synthetic_PenTest_ReportsThe full CJ Jones' synthetic dataset catalog is available at:
https://datadeveloper1.gumroad.com
Want more? 🚀 Get the AI Startup Bundle from Gumroad.
📄 100 Samples of Synthetic Automated Penetration Test Reports
This dataset contains 100+ realistic, synthetic penetration testing reportsstructured to simulate professional internal security assessments. Each record models the full flow of a pentest engagement, including:
Reconnaissance / Discovery Phase
Vulnerability Assessment… See the full description on the dataset page: https://huggingface.co/datasets/CJJones/Synthetic_PenTest_Reports.mirror-pentesting-explanations
Pentesting Explanations - Adversarial Reasoning & Vulnerability Research
A high-quality supervised fine-tuning dataset for penetration testing expertise, red team tradecraft, and - as the dataset matures - novel vulnerability research and zero-day reasoning. The dataset is structured to teach models how to think like offensive security practitioners, not merely recall labels or technique names.
The long-term goal of this dataset is to train models capable of genuine adversarial… See the full description on the dataset page: https://huggingface.co/datasets/alucent/mirror-pentesting-explanations.code-pensions-civiles-militaires-retraite
Code des pensions civiles et militaires de retraite, non-instruct (2025-03-10)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-pensions-civiles-militaires-retraite.counterfactual-pendulum-multilingual
📌 Dataset Summary
When a Vision-Language Model (VLM) is given an image along with a text prompt containing contradictory or misleading information, how does it react? Does it rely on the visual evidence, succumb to textual bias, or honestly abstain when faced with unresolvable conflict?
This dataset adapts the Counterfactual Pendulum scenario across two visual conflict dimensions:
Angular (Angle): Conflict in the pendulum's angle of inclination.
Light: Conflict in the light… See the full description on the dataset page: https://huggingface.co/datasets/akanshjain37/counterfactual-pendulum-multilingual.bug-bounty-pentest-fr
Bug Bounty & Méthodologies de Pentest
Méthodologies (OWASP, PTES), checklists par type d app, techniques d attaque, plateformes, templates de rapports et outils.
Links
Version anglaise
AYI NEDJIMI Consultants
code-penal
Code pénal, non-instruct (2025-05-20)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of free, open-source language models based… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-penal.DECADE
DECADE: Dataset for Evolving Context And Dialogue Evaluation
DECADE is a benchmark for evaluating long-term memory reasoning in personalized conversational AI. It simulates a decade (2016–2026) of user interactions across 500 QA instances, each paired with a personal conversation history of up to 1,047 sessions.
Task
Given a user's long conversation history (haystack) and a question posed from a future date, a system must retrieve the relevant sessions and synthesize an… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-penguin/DECADE.code-procedure-penale
Code de procédure pénale, non-instruct (2025-03-10)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of free, open-source language… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-procedure-penale.code-justice-penale-mineurs
Code de la justice pénale des mineurs, non-instruct (2025-09-20)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of free… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-justice-penale-mineurs.cantonesewiki_doyouknow
Cantonese Question Dataset from Yue Wiki
A collection of questions in Cantonese, extracted from Yue Wiki. This dataset contains a variety of questions covering different topics and domains.
Disclaimer
The content and opinions expressed in this dataset do not represent the views, beliefs, or positions of the dataset creators, contributors, or hosting organizations. This dataset is provided solely for the purpose of improving AI systems' understanding of the Cantonese… See the full description on the dataset page: https://huggingface.co/datasets/pendingremove32894/cantonesewiki_doyouknow.testaasdfsdf
offences_and_penalties_in_general_2018_datasetcode-disciplinaire-penal-marine-marchande
Code disciplinaire et pénal de la marine marchande, non-instruct (2025-09-20)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-disciplinaire-penal-marine-marchande.code-penitentiaire
Code pénitentiaire, non-instruct (2025-03-10)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of free, open-source language models… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-penitentiaire.code-pensions-retraite-marins-francais-commerce-peche-plaisance
Code des pensions de retraite des marins français du commerce, de pêche ou de plaisance, non-instruct (2025-03-10)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-pensions-retraite-marins-francais-commerce-peche-plaisance.code-pensions-militaires-invalidite-victimes-guerre
Code des pensions militaires d'invalidité et des victimes de guerre, non-instruct (2025-03-10)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-pensions-militaires-invalidite-victimes-guerre.
