syvb/nanonla-qwen3-8b-L24-results
Qwen3-8B NLA length-penalty sweep — results bundle Held-out completions (1000 per model, matched by idx across models) for the from-scratch base NLA and each RL length penalty. The viewer shows the completions config (per-sample explanations + reconstruction FVE). Also in the repo (as files, not loaded configs): per-model *.summary.json aggregates, RESULTS.md, comparison_base_vs_penalty.md, tradeoff.png. Per-sample columns: idx, tag (model), source_text (the actual source… See the full description on the dataset page: https://huggingface.co/datasets/syvb/nanonla-qwen3-8b-L24-results.
Qwen3-8B NLA length-penalty sweep — results bundle
Held-out completions (1000 per model, matched by idx across models) for the from-scratch base NLA and each RL length penalty. The viewer shows the completions config (per-sample explanations + reconstruction FVE). Also in the repo (as files, not loaded configs): per-model *.summary.json aggregates, RESULTS.md, comparison_base_vs_penalty.md, tradeoff.png.
Per-sample columns: idx, tag (model), source_text (the actual source document), explanation, n_tokens, mse/nmse/fve, reward, extracted, cjk. For a clean single-table version see the `-completions` dataset.
