nick-rui/proofwriter-qwen25-7b-rar-delta
proofwriter-qwen25-7b-rar-delta Per-question delta_RaR annotations on ProofWriter: how much a model's own rephrasing of a logic problem improves its ability to solve it. Computed with Qwen/Qwen2.5-7B-Instruct, 30,000 queries at k=32 samples per side. The "teacher" is not a stronger model and does not think longer. It is the same frozen model answering the same question, with one extra thing in context — a rephrasing it generated itself. How each row is produced… See the full description on the dataset page: https://huggingface.co/datasets/nick-rui/proofwriter-qwen25-7b-rar-delta.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face