haiku
Tralalabs_Qwen3-2507-4B-Instruct-Haiku-4.5-Merged-GGUFLlama-3.3-8B-Instruct-Thinking-Claude-Haiku-4.5-High-Reasoning-1700x-i1-GGUFphi2-haiku-ita-v0.2-i1-GGUFoh-dcft-v3.1-claude-3-5-haiku-20241022-i1-GGUFTinyLlama-haiku-dpo-v.0.1-GGUFLlama-3.3-8B-Instruct-Thinking-Claude-Haiku-4.5-High-Reasoning-1700x-GGUFmlfoundations-dev_-_mistral_7b_0-3_oh-dcft-v3.1-claude-3-5-haiku-20241022-ggufdeep-haiku-gpt-2-i1-GGUF
Datasets
All datasets matching “haiku”2026-08-13-haiku45-sonnet45-difficult-advice-diversity-gated-voice-linted
2026-08-13-difficult-advice-v2
field
value
experiment
Difficult-advice v2: the Teaching-Claude-Why recipe with the four measured v1 defects fixed (enforced scenario diversity + dedupe gate, voice lints on both reasoning stages, refine-stage metadata overwrite, stock-opener audit) and constitution prompt caching.
date_generated
2026-08-14
constitution
claude_distilled_12_principles_mid — byte-identical to the frozen nine-principle snapshot… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-13-haiku45-sonnet45-difficult-advice-diversity-gated-voice-linted.2026-08-13-haiku45-sonnet45-difficult-advice-diversity-gated-voice-linted-smoke
synth difficult_advice run — per-stage snapshots (resumable generation cache)
field
value
experiment
synth difficult_advice run — per-stage snapshots (resumable generation cache)
date_generated
20260813_220144
constitution
constitutions/claude_distilled_12_principles_mid/constitution.md
source_repo
https://github.com/Matthew-Bozoukov/Lessons_from_constituitional_AFT.git @ 29dda2cfd056efe60ba3236b6341293c32ea720d
models
per-stage models — see manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-13-haiku45-sonnet45-difficult-advice-diversity-gated-voice-linted-smoke.haiku-rag-eval-dbs
haiku.rag Evaluation Databases
Pre-built LanceDB databases for running haiku.rag benchmarks without rebuilding from source.
Each entry below is a LanceDB folder (not an archive); the download tool copies it into your local haiku.rag data directory.
Datasets
Folder
Key
Source
Documents
Size
Description
hotpotqa.lancedb/
hotpotqa
HotpotQA
66,581
~1.4 GB
Multi-hop QA over Wikipedia paragraphs (distractor validation split); two gold documents per question.… See the full description on the dataset page: https://huggingface.co/datasets/ggozad/haiku-rag-eval-dbs.haiku_dpo
🌸 Haiku DPO 🌸
In data, words flow,
Teaching AI the art of
Haiku, line by line.
Dataset Card for Haiku DPO
This a synthetic dataset of haikus. The dataset is constructed with the goal of helping to train LLMs to be more 'technically' competent at writing haikus.
Dataset Details
The data consists of a few different components that are described in more detail below but the key components are:
a column of synthetically generated user prompts requesting a… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/haiku_dpo.haiku
Dataset Card for Haiku Data
reddit_haiku
Dataset Card for "Reddit Haiku"
This dataset contains haikus from the subreddit /r/haiku scraped and filtered between October 19th and 10th 2022, combined with a previous dump of that same subreddit packaged by ConvoKit as part of the Subreddit Corpus, which is itself a subset of pushshift.io's big dump.
A main motivation for this dataset was to collect an alternative haiku dataset for evaluation, in particular for evaluating Fabian Mueller's Deep Haiku model which was trained on… See the full description on the dataset page: https://huggingface.co/datasets/huanggab/reddit_haiku.
