CoolFace
Datasetpublic

syvb/nanonla-qwen3-8b-L24-results

Qwen3-8B NLA length-penalty sweep — results bundle Held-out completions (1000 per model, matched by idx across models) for the from-scratch base NLA and each RL length penalty. The viewer shows the completions config (per-sample explanations + reconstruction FVE). Also in the repo (as files, not loaded configs): per-model *.summary.json aggregates, RESULTS.md, comparison_base_vs_penalty.md, tradeoff.png. Per-sample columns: idx, tag (model), source_text (the actual source… See the full description on the dataset page: https://huggingface.co/datasets/syvb/nanonla-qwen3-8b-L24-results.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
1likes16downloads
9 commits on main
8c074fa3mo ago

tradeoff plot: y-only error bars, n=1000

syvb
ae25bc33mo ago

tradeoff plot: smaller star, repositioned labels

syvb
9ea332c3mo ago

add 95% CI error bars + SEM fields to summaries

syvb
8547c183mo ago

tradeoff plot: base as standalone marker off the RL line

syvb
0edc2103mo ago

Upload README.md with huggingface_hub

syvb
dfffa5b3mo ago

Upload comparison_base_vs_penalty.md with huggingface_hub

syvb
c9a3d1a3mo ago

Upload folder using huggingface_hub

syvb
c6d70a13mo ago

Upload folder using huggingface_hub

syvb
65de4113mo ago

initial commit

syvb