datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
text-correction-validationtext-correction_collection
Human Samples
These samples contains contains human-written sentences produced during language learning practice, combined with AI-based grammatical verification and correction. The original sentences were written by language learners who often did not know whether their sentences were correct or incorrect. These authentic learner inputs capture a wide range of natural mistakes, such as spelling, syntax, word choice, and structure errors.
Synthetic Samples
These… See the full description on the dataset page: https://huggingface.co/datasets/marcelone/text-correction_collection.text_correction_finetuning
