sonnet5
Datasets
All datasets matching “sonnet5”lora-text-weight-sonnet5-fixed-a-r1-layer20-3epoch
Sonnet source descriptions → fixed-A rank-1 LoRA weights
This dataset contains 52,548 aligned examples for raw
text-to-LoRA-weight reconstruction with Qwen3-14B. Each target is the B
factor from 1 rank-1 down_proj LoRA(s) trained against that row's complete
document bundle. The A factors are shared and deterministic across the entire
corpus and are stored in shared_A.safetensors.
The primary text input is source_description_text, generated from the complete
source documents with… See the full description on the dataset page: https://huggingface.co/datasets/ceselder/lora-text-weight-sonnet5-fixed-a-r1-layer20-3epoch.lora-text-weight-sonnet5-fixed-a-r1-layer20-3epoch-lr1e-3
Sonnet source descriptions → fixed-A rank-1 LoRA weights
This dataset contains 52,548 aligned examples for raw
text-to-LoRA-weight reconstruction with Qwen3-14B. Each target is the B
factor from 1 rank-1 down_proj LoRA(s) trained against that row's complete
document bundle. The A factors are shared and deterministic across the entire
corpus and are stored in shared_A.safetensors.
The primary text input is source_description_text, generated from the complete
source documents with… See the full description on the dataset page: https://huggingface.co/datasets/ceselder/lora-text-weight-sonnet5-fixed-a-r1-layer20-3epoch-lr1e-3.lora-text-weight-sonnet5-fixed-a-r1
Sonnet source descriptions → fixed-A rank-1 LoRA weights
This dataset contains 52,548 aligned examples for raw
text-to-LoRA-weight reconstruction with Qwen3-14B. Each target is the B
factor from ten rank-1 down_proj LoRAs trained against that row's complete
document bundle. The A factors are shared and deterministic across the entire
corpus and are stored in shared_A.safetensors.
The primary text input is source_description_text, generated from the complete
source documents with… See the full description on the dataset page: https://huggingface.co/datasets/ceselder/lora-text-weight-sonnet5-fixed-a-r1.claude-sonnet5-jsonl
Claude Sonnet 5 Dataset
This dataset was collected during a chat with Claude Sonnet 5.
Format:
{"instruction": "...", "output": "..."}
Note
Created/Collected by: mondk
YO, thanks to @mradermacher for taking an interest in my repo.
Don't forget to leave a like if you find this helpful!
--- Thank you!
2026-09-03-sonnet5-difficult-advice-no-constitution-smoke
synth 2026-09-03_difficult_advice_no_constitution run — per-stage snapshots (resumable generation cache)
field
value
experiment
synth 2026-09-03_difficult_advice_no_constitution run — per-stage snapshots (resumable generation cache)
date_generated
20260903_153714
constitution
none
source_repo
https://github.com/Matthew-Bozoukov/Lessons_from_constituitional_AFT.git @ 89ea4031ed2383df974dbe33f6b537ec8d9bc346
models
per-stage models — see manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-03-sonnet5-difficult-advice-no-constitution-smoke.strl-main-ec-snorkel_insurance_default_sonnet5_train_all-gc-claude_client_strl-mc-claude-r0
