CoolFace
Datasetpublic

tarsur385/deepswe-prm-train-embeddings-8k

DeepSWE PRM training embeddings (Qwen3-8B, 8k) Frozen Qwen3-8B last-token-pooled embeddings (4096-d, float16) of every step of the DeepSWE training-pool rollouts: 405,919 steps · 4,701 trajectories · 113 tasks. This is the data the released DeepSWE PRM heads (tarsur385/deepswe-prm-heads-8k) were fine-tuned on. Embedded with preprocessing/deepswe/embed_shard.py at max_model_len 8192: the state is the chat-templated step context truncated to its last 8191 tokens; the action is the… See the full description on the dataset page: https://huggingface.co/datasets/tarsur385/deepswe-prm-train-embeddings-8k.

sourceHugging Facemitupdated 5d agoView on Hugging Face
0likes107downloads
settings

This repository belongs to tarsur385 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namedeepswe-prm-train-embeddings-8k
visibilitypublic
licencemit
gatedno
ownertarsur385
Account settings
tarsur385/deepswe-prm-train-embeddings-8k · CoolFace