CoolFace
Datasetpublic

RL-Forgetting-Experiments-3/qwen2.5-3b-math-kk-sft-artifacts

Qwen2.5-3B Math and Knights-and-Knaves SFT artifacts Training metrics, per-rank training manifests, exact SFT configs, and full persisted evaluation outputs for the ordered/shuffled Math and KK SFT arms. Each arm has ten checkpoint evaluations at n=160. The KK ordered step-3175 HF model is complete for inference/evaluation, but its later optimizer/prev-params serialization failed, so no resumable training-state claim is made. See delivery_manifest.json for source lineage and… See the full description on the dataset page: https://huggingface.co/datasets/RL-Forgetting-Experiments-3/qwen2.5-3b-math-kk-sft-artifacts.

sourceHugging Faceapache-2.0updated 5d agoView on Hugging Face
0likes24downloads
settings

This repository belongs to RL-Forgetting-Experiments-3 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameqwen2.5-3b-math-kk-sft-artifacts
visibilitypublic
licenceapache-2.0
gatedno
ownerRL-Forgetting-Experiments-3
Account settings
RL-Forgetting-Experiments-3/qwen2.5-3b-math-kk-sft-artifacts · CoolFace