trytryw/fixed-alpha-scale-pilot-v1
Fixed Alpha Score-Scale Pilot v1 Public artifacts for a controlled 125M-parameter, 4K-context language-model study of fixed attention score scaling with full sparsemax (alpha=2), full entmax-1.5, and a softmax reference. Layout metadata/: immutable execution records, plans, logs, source/config snapshots, and SHA-256 manifests. runs/<run-name>/: complete run backups, including configuration, provenance, metrics, final training state, and preregistered analysis… See the full description on the dataset page: https://huggingface.co/datasets/trytryw/fixed-alpha-scale-pilot-v1.
0537
