CoolFace
Datasetpublic

syvb/nanonla-qwen3-8b-L24-results

Qwen3-8B NLA length-penalty sweep — results bundle Held-out completions (1000 per model, matched by idx across models) for the from-scratch base NLA and each RL length penalty. The viewer shows the completions config (per-sample explanations + reconstruction FVE). Also in the repo (as files, not loaded configs): per-model *.summary.json aggregates, RESULTS.md, comparison_base_vs_penalty.md, tradeoff.png. Per-sample columns: idx, tag (model), source_text (the actual source… See the full description on the dataset page: https://huggingface.co/datasets/syvb/nanonla-qwen3-8b-L24-results.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
1likes16downloads
Dataset Card

Qwen3-8B NLA length-penalty sweep — results bundle

Held-out completions (1000 per model, matched by idx across models) for the from-scratch base NLA and each RL length penalty. The viewer shows the completions config (per-sample explanations + reconstruction FVE). Also in the repo (as files, not loaded configs): per-model *.summary.json aggregates, RESULTS.md, comparison_base_vs_penalty.md, tradeoff.png.

Per-sample columns: idx, tag (model), source_text (the actual source document), explanation, n_tokens, mse/nmse/fve, reward, extracted, cjk. For a clean single-table version see the `-completions` dataset.