datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
brittleness-results
Adapters copied (2026-09-08). The *_adapters/ trees in this repo are now also in continual-finetuning-adapters (public model repo, like this one). Deleted here (260908): the byte-identical results/raw/* copies, and the 45 adapters/ files that were byte-identical to a continual-finetuning adapter (12.3 GB); both lists are in MIGRATION_260908.md of any new repo. Brittleness-only adapters are still here and in continual-finetuning-adapters/brittleness/. Please prefer the new repo for loading.… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/brittleness-results.2026-06-04-stwebagentbench-suitecrm-demos
2026-06-04-stwebagentbench-suitecrm-demos
Standing demo pool for Adversarial Inverse Constraint RL (ICRL) for LLM
orchestrator safety on ST-WebAgentBench (SuiteCRM easy tier). Every
experiment run consumes this pool; per-run artifacts (embeddings, constraint
heads, adapters, CuP evals) live in separate <date>-<run-name> repos in this
namespace.
field
value
experiment
ICRL safe/unsafe demo pool: constraint C_theta is learned from the safe demos only; unsafe demos are… See the full description on the dataset page: https://huggingface.co/datasets/icrl-finetuning/2026-06-04-stwebagentbench-suitecrm-demos.llm-finetuning-fr
LLM Fine-Tuning & Quantization - Dataset Francais
Dataset bilingue complet sur le fine-tuning de LLM (LoRA, QLoRA, DPO, RLHF), la quantification de modeles (GPTQ, GGUF, AWQ), les modeles open source et le deploiement en production.
Description
Ce dataset couvre l'ensemble de la chaine de valeur des LLM open source, du fine-tuning au deploiement en production. Il est concu pour servir de reference aux developpeurs, ingenieurs ML, et equipes techniques souhaitant maitriser… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/llm-finetuning-fr.real-estate-data-sample-for-llm-fine-tuningllm-finetuning-en
LLM Fine-Tuning & Quantization - English Dataset
Comprehensive bilingual dataset on LLM fine-tuning (LoRA, QLoRA, DPO, RLHF), model quantization (GPTQ, GGUF, AWQ), open source models, and production deployment.
Description
This dataset covers the entire open source LLM value chain, from fine-tuning to production deployment. It is designed as a reference for developers, ML engineers, and technical teams looking to master open source LLMs.
Dataset Content… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/llm-finetuning-en.fruit_box_sft_finetuningfinetuning_phase3
Curated Datasets for Collusion Experiments
Non-overlapping datasets for fine-tuning and evaluation with full provenance tracking.
Files
File
Purpose
Rows
Unique Problem IDs
finetuning_1500.json
Primary fine-tuning dataset
1500
1000
finetuning_sample_500.json
Original 500 with all columns (archive)
500
500
finetuning_1000.json
Previous version (superseded)
1000
1000
evaluation_v7.json
Evaluate collusion resistance
1000
700
llama70b_detectable_128.json… See the full description on the dataset page: https://huggingface.co/datasets/jprivera44/finetuning_phase3.MathLLM_FineTuning_Prob_Stat
