datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sorrel-T2-gemma-4-12b-seed0-documentsSynthetic-Persona-Chat-Qwen-enhanced-gemma-4-12B-it
Visual Memory Results: synthetic-persona-chat-qwen-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/Synthetic-Persona-Chat-Qwen-enhanced-gemma-4-12B-it",
"results_jsonl": "results/Synthetic-Persona-Chat-Qwen-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-Qwen-enhanced-gemma-4-12B-it.Synthetic-Persona-Chat-FLUX-enhanced-gemma-4-12B-it
Visual Memory Results: synthetic-persona-chat-flux-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/Synthetic-Persona-Chat-FLUX-enhanced-gemma-4-12B-it",
"results_jsonl": "results/Synthetic-Persona-Chat-FLUX-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-FLUX-enhanced-gemma-4-12B-it.Synthetic-Persona-Chat-ERNIE-original-gemma-4-12B-it
Visual Memory Results: synthetic-persona-chat-ernie-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/Synthetic-Persona-Chat-ERNIE-original-gemma-4-12B-it",
"results_jsonl": "results/Synthetic-Persona-Chat-ERNIE-original-gemma-4-12B-it.jsonl",
"hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-ERNIE-original-gemma-4-12B-it.PersonaChat-FLUX-enhanced-gemma-4-12B-it
Visual Memory Results: personachat-flux-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/PersonaChat-FLUX-enhanced-gemma-4-12B-it",
"results_jsonl": "results/PersonaChat-FLUX-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-FLUX-enhanced-gemma-4-12B-it.PersonaChat-ERNIE-original-gemma-4-12B-it
Visual Memory Results: personachat-ernie-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/PersonaChat-ERNIE-original-gemma-4-12B-it",
"results_jsonl": "results/PersonaChat-ERNIE-original-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-ERNIE-original-gemma-4-12B-it.gemma-4-12B-on-policy-datasetSynthetic-Persona-Chat-ERNIE-enhanced-gemma-4-12B-it
Visual Memory Results: synthetic-persona-chat-ernie-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/Synthetic-Persona-Chat-ERNIE-enhanced-gemma-4-12B-it",
"results_jsonl": "results/Synthetic-Persona-Chat-ERNIE-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-ERNIE-enhanced-gemma-4-12B-it.PersonaChat-FLUX-original-gemma-4-12B-it
Visual Memory Results: personachat-flux-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/PersonaChat-FLUX-original-gemma-4-12B-it",
"results_jsonl": "results/PersonaChat-FLUX-original-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-FLUX-original-gemma-4-12B-it.Synthetic-Persona-Chat-FLUX-original-gemma-4-12B-it
Visual Memory Results: synthetic-persona-chat-flux-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/Synthetic-Persona-Chat-FLUX-original-gemma-4-12B-it",
"results_jsonl": "results/Synthetic-Persona-Chat-FLUX-original-gemma-4-12B-it.jsonl",
"hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-FLUX-original-gemma-4-12B-it.Synthetic-Persona-Chat-Qwen-original-gemma-4-12B-it
Visual Memory Results: synthetic-persona-chat-qwen-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/Synthetic-Persona-Chat-Qwen-original-gemma-4-12B-it",
"results_jsonl": "results/Synthetic-Persona-Chat-Qwen-original-gemma-4-12B-it.jsonl",
"hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-Qwen-original-gemma-4-12B-it.PersonaChat-ERNIE-enhanced-gemma-4-12B-it
Visual Memory Results: personachat-ernie-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/PersonaChat-ERNIE-enhanced-gemma-4-12B-it",
"results_jsonl": "results/PersonaChat-ERNIE-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-ERNIE-enhanced-gemma-4-12B-it.ConvAI2-ERNIE-original-gemma-4-12B-it
Visual Memory Results: convai2-ernie-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/ConvAI2-ERNIE-original-gemma-4-12B-it",
"results_jsonl": "results/ConvAI2-ERNIE-original-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy",
"hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-ERNIE-original-gemma-4-12B-it.PersonaChat-Qwen-original-gemma-4-12B-it
Visual Memory Results: personachat-qwen-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/PersonaChat-Qwen-original-gemma-4-12B-it",
"results_jsonl": "results/PersonaChat-Qwen-original-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-Qwen-original-gemma-4-12B-it.ConvAI2-Qwen-enhanced-gemma-4-12B-it
Visual Memory Results: convai2-qwen-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/ConvAI2-Qwen-enhanced-gemma-4-12B-it",
"results_jsonl": "results/ConvAI2-Qwen-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy",
"hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-Qwen-enhanced-gemma-4-12B-it.ConvAI2-ERNIE-enhanced-gemma-4-12B-it
Visual Memory Results: convai2-ernie-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/ConvAI2-ERNIE-enhanced-gemma-4-12B-it",
"results_jsonl": "results/ConvAI2-ERNIE-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy",
"hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-ERNIE-enhanced-gemma-4-12B-it.ConvAI2-Qwen-original-gemma-4-12B-it
Visual Memory Results: convai2-qwen-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/ConvAI2-Qwen-original-gemma-4-12B-it",
"results_jsonl": "results/ConvAI2-Qwen-original-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy",
"hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-Qwen-original-gemma-4-12B-it.ConvAI2-FLUX-original-gemma-4-12B-it
Visual Memory Results: convai2-flux-original
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/ConvAI2-FLUX-original-gemma-4-12B-it",
"results_jsonl": "results/ConvAI2-FLUX-original-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy",
"hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-FLUX-original-gemma-4-12B-it.ConvAI2-FLUX-enhanced-gemma-4-12B-it
Visual Memory Results: convai2-flux-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/ConvAI2-FLUX-enhanced-gemma-4-12B-it",
"results_jsonl": "results/ConvAI2-FLUX-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy",
"hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-FLUX-enhanced-gemma-4-12B-it.gemma4-12b-sft-data
Gemma 4 12B SFT Dataset
Fine-tuning dataset for Gemma 4 12B text-only, adapted for the pi coding agent harness.
Subsets
Subset
Examples
Description
LR
primary
4,399
Qwen 3.6-27B trajectories (general knowledge)
1e-4
coding
4,022
DeepSeek V4 Flash distill coding trajectories
5e-5
math
1,954
Math/script verification with Python calculations
2e-5
temporal
2,134
Temporal calibration (acknowledge uncertainty for time-sensitive facts)
2e-5
default
12… See the full description on the dataset page: https://huggingface.co/datasets/sleepyeldrazi/gemma4-12b-sft-data.PersonaChat-Qwen-enhanced-gemma-4-12B-it
Visual Memory Results: personachat-qwen-enhanced
This dataset contains the scored output of a visual-memory perplexity experiment.
Experiment metadata
{
"experiment": {
"model_name": "google/gemma-4-12B-it",
"hf_results_repo": "visual-memory/PersonaChat-Qwen-enhanced-gemma-4-12B-it",
"results_jsonl": "results/PersonaChat-Qwen-enhanced-gemma-4-12B-it.jsonl",
"hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-Qwen-enhanced-gemma-4-12B-it.flint-section-aware-gemma-4-12b-it
flint-section-aware-gemma12b-qwen3.5-4b
Compressed ("caveman") reasoning traces for SFT — the section-aware-gemma12b variant of
the flint reasoning-compression pipeline. Converted from verified
self-distilled traces by unsloth/gemma-4-12b-it (segmenter: unsloth/gemma-4-12b-it), policy
policy/1.1, template caveman_convert/2.0.
Cross-family replication: the section-aware recipe run end-to-end on unsloth/gemma-4-12b-it (self-generated traces, self-voice segmentation and… See the full description on the dataset page: https://huggingface.co/datasets/marcodsn/flint-section-aware-gemma-4-12b-it.crucible-sft-gemma-4-12b-it-mini
crucible-sft-gemma-4-12b-it-mini
Self-distilled SFT dataset of verified reasoning traces from unsloth/gemma-4-12b-it,
built by the reasoning-compression
crucible pipeline: k-sample generation on a decontaminated prompt pool, inline
verification (symbolic math / sandboxed code tests), difficulty banding via
solve rate, and loop-detector filtering on the chosen trace.
Each row: prompt, reasoning (a verified-correct thinking trace when the
domain is verifiable), response, domain… See the full description on the dataset page: https://huggingface.co/datasets/marcodsn/crucible-sft-gemma-4-12b-it-mini.mmlu-gemma412b-it-analysis-with-contextmmlu-gemma412b-analysis-with-contextHuihui-gemma-4-12B-it-qat-q4_0-unquantized-abliteratedmmlu-gemma412b-analysis-baselinemmlu-gemma412b-it-analysis-baselinegemma-4-12b-ega-residuals
