danilxyz/rest-v3
rest-v3 rest-v3 is an English text-rewriting dataset for supervised fine-tuning of a humanizing editor. Each record asks a model to rewrite a source text while preserving its meaning and contains a detector-verified natural-language rewrite. Dataset composition The training split contains 1,116 JSONL records: 1,033 newly mined, on-policy rewrites from the from-final-best generator checkpoint. 83 compatible existing verified examples. 541 examples sourced from the… See the full description on the dataset page: https://huggingface.co/datasets/danilxyz/rest-v3.
Normalize training weights to unit mean
Document weighting and legacy exclusions
Normalize source-type training weights
Fix dataset card metadata
Document dataset construction
Add rest-v3 training split
initial commit
