CoolFace
Datasetpublic

nick-rui/proofwriter-qwen25-7b-rar-delta

proofwriter-qwen25-7b-rar-delta Per-question delta_RaR annotations on ProofWriter: how much a model's own rephrasing of a logic problem improves its ability to solve it. Computed with Qwen/Qwen2.5-7B-Instruct, 30,000 queries at k=32 samples per side. The "teacher" is not a stronger model and does not think longer. It is the same frozen model answering the same question, with one extra thing in context — a rephrasing it generated itself. How each row is produced… See the full description on the dataset page: https://huggingface.co/datasets/nick-rui/proofwriter-qwen25-7b-rar-delta.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes14downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
nick-rui/proofwriter-qwen25-7b-rar-delta · CoolFace