VibrantVista/style-judge-dataset
Style Judge Dataset A pairwise dataset for learning a continuous style-similarity function while controlling for topic, introduced in Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning (arXiv:2512.05747). Dataset Summary Total rows: 156k (default subset) Splits: train 130k, validation 13k, test 13k Format: Arrow Columns sentence1 (string): original chunk text sentence2 (string): refilled chunk text score… See the full description on the dataset page: https://huggingface.co/datasets/VibrantVista/style-judge-dataset.
Delete train/data-00003-of-00004.arrow
Delete train/data-00002-of-00004.arrow
Delete train/data-00001-of-00004.arrow
Delete train/data-00000-of-00004.arrow
Delete train/data-00000-of-00001.arrow
Update README.md
Update README.md
Update README.md
Create README.md
Upload folder using huggingface_hub
initial commit
