CoolFace
Datasetpublic

jayshah5696/humanize-rl-sft-dataset

humanize-rl-sft-dataset (v2) 4,835 high-quality SFT pairs for training a model to write natural, direct prose. Part of the humanize-rl project — a two-layer scoring and alignment pipeline for training small models to generate natural, human-sounding text. What this trains A model that can: Write natural Slack messages and emails from scratch. Rewrite stiff/formal/corporate text into direct, human-sounding prose. Fix grammar without making text formal. Shorten and… See the full description on the dataset page: https://huggingface.co/datasets/jayshah5696/humanize-rl-sft-dataset.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
0likes39downloads

jayshah5696/humanize-rl-sft-dataset · main · files are served by the source, never re-hosted here