CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jacekduszenko /coedit-multilingual CoEdIT Multilingual A multilingual text-editing dataset for evaluating instruction-following edit models (grammar correction, paraphrase, simplification, neutralization, coherence, clarity) plus operational speech-transcript edits (translation, number/punctuation formatting, filler removal, summarization, formalize/casualize, abbreviation expansion). Construction English: the original grammarly/coedit (train + validation). Translated (de, es, fr, it, pt, nl, zh… See the full description on the dataset page: https://huggingface.co/datasets/jacekduszenko/coedit-multilingual.text100K<n<1M0 likes55 downloads3mo agoHugging Face02JohnGorri /gec-coherence-coedit-synthtext10K<n<100K0 likes49 downloads3mo agoHugging Face03owanr /o1o2o3_large_r2_coedit_with_human_pref_practice Dataset Card for "o1o2o3_large_r2_coedit_with_human_pref_practice" More Information needed text100K<n<1M0 likes42 downloads3y agoHugging Face04muzzz /coedit-cot-reasoning Dataset Card for CoEdIT-CoT-Reasoning: Text Editing with Step-by-Step (Chain of Thought) Reasoning Dataset Description This dataset extends the original CoEdIT dataset by adding detailed step-by-step reasoning traces that explain how to perform various text editing tasks. The reasoning traces simulate the thought process of an expert editor applying the given instructions. Dataset Summary CoEdIT-Reasoning augments the CoEdIT text editing dataset with… See the full description on the dataset page: https://huggingface.co/datasets/muzzz/coedit-cot-reasoning.text10K<n<100K0 likes40 downloads1y agoHugging Face05xzuyn /BEE-spoke-data-coedit-reworded-dedupedtext10K<n<100K0 likes39 downloads3y agoHugging Face06owanr /r2_coedit Dataset Card for "r2_coedit" More Information needed text10K<n<100K2 likes37 downloads3y agoHugging Face07owanr /r2_coedit_iter Dataset Card for "r2_coedit_iter" More Information needed text10K<n<100K1 likes37 downloads3y agoHugging Face08premkumarelangovan /coedit_phrasestext100K<n<1M0 likes36 downloads1y agoHugging Face09owanr /o1o2o3_large_r2_coedit_iter_with_human_pref_practice Dataset Card for "o1o2o3_large_r2_coedit_iter_with_human_pref_practice" More Information needed text100K<n<1M0 likes31 downloads3y agoHugging Face10gnokit /coedit_llmtexttext-generation10K<n<100K0 likes31 downloads2y agoHugging Face11owanr /r2_coedit_v2 Dataset Card for "r2_coedit_v2" More Information needed text100K<n<1M2 likes29 downloads3y agoHugging Face12GPTZERO /merged-coedit-a1-bonafid-2point5-miltext1M<n<10M0 likes28 downloads2y agoHugging Face13dim /grammarly_coedit Dataset Card for "grammarly_coedit" More Information needed text10K<n<100K2 likes25 downloads3y agoHugging Face14chargoddard /coedit-reworded coedit-reworded This is Grammarly's coedit dataset parsed into Alpaca-style instruction, input, and output rows, with the original instruction values replaced with a more diverse set of procedurally generated instructions. Contains 23930 unique values of instruction, as compared to the original 144. See coedit_reword.py for how these were generated. All credit to the original authors of this dataset. Citation @article{raheja2023coedit, title={CoEdIT: Text… See the full description on the dataset page: https://huggingface.co/datasets/chargoddard/coedit-reworded.texttext-generation10K<n<100K4 likes23 downloads3y agoHugging Face15owanr /o1o2o3_xl_r2_coedit_iter_with_human_pref_practice Dataset Card for "o1o2o3_xl_r2_coedit_iter_with_human_pref_practice" More Information needed text100K<n<1M0 likes20 downloads3y agoHugging Face16bihungba1101 /Vocab-CoEdIT Vocab-Coedit Made with ❤️ using 🦥 Unsloth Studio Vocab-CoEdIT contains 59,949 records of vocabulary suggestions 🚀 Quick Start from datasets import load_dataset # Load the main dataset dataset = load_dataset("bihungba1101/Vocab-CoEdIT", "data", split="train") df = dataset.to_pandas() 📊 Dataset Summary 📈 Records: 59,949 📋 Columns: 7 ✅ Completion: 99.9% (60,000 requested) 📋 Schema & Statistics Column Type Column Type Unique (%) Null (%)… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/Vocab-CoEdIT.text10K<n<100K0 likes20 downloads5mo agoHugging Face17nayohan /coedit-koTranslated grammarly/coedit using nayohan/llama3-instrucTrans-enko-8b. This dataset is a raw translated dataset and contains repetitive sentences generated by the model, so it needs to be filtered. @article{raheja2023coedit, title={CoEdIT: Text Editing by Task-Specific Instruction Tuning}, author={Vipul Raheja and Dhruv Kumar and Ryan Koo and Dongyeop Kang}, year={2023}, eprint={2305.09857}, archivePrefix={arXiv}, primaryClass={cs.CL} } texttext-generation10K<n<100K0 likes19 downloads2y agoHugging Face18Emm9625 /r2_coedit_v2_in_outtext100K<n<1M0 likes19 downloads2y agoHugging Face19owanr /r1_coedit_v2 Dataset Card for "r1_coedit_v2" More Information needed text100K<n<1M0 likes17 downloads3y agoHugging Face20owanr /r1_coedit Dataset Card for "r1_coedit" More Information needed text10K<n<100K0 likes16 downloads3y agoHugging Face21owanr /o1o2o3_xl_r2_coedit_with_human_pref_practice Dataset Card for "o1o2o3_xl_r2_coedit_with_human_pref_practice" More Information needed text100K<n<1M0 likes16 downloads3y agoHugging Face22owanr /o1o2o3_xl_r2_coedit Dataset Card for "o1o2o3_xl_r2_coedit" More Information needed text10K<n<100K1 likes14 downloads3y agoHugging Face23tcapelle /coedit-fluencytext10K<n<100K0 likes14 downloads2y agoHugging Face24owanr /o1o2o3_xl_r2_coedit_with_human_pref Dataset Card for "o1o2o3_xl_r2_coedit_with_human_pref" More Information needed text100K<n<1M0 likes13 downloads3y agoHugging Face25owanr /r1_coedit_iter Dataset Card for "r1_coedit_iter" More Information needed text100K<n<1M0 likes11 downloads3y agoHugging Face26kraalfar /Coeditor-processed-demo2 Dataset Card for "Coeditor-processed-demo2" More Information needed text1K<n<10K0 likes10 downloads1y agoHugging Face27Emm9625 /r2_coedit_v2_convtext100K<n<1M0 likes9 downloads2y agoHugging Face28bihungba1101 /Essay-Vocab-CoEdITtext1K<n<10K0 likes9 downloads4mo agoHugging Face29owanr /o1o2o3_large_r2_coedit Dataset Card for "o1o2o3_large_r2_coedit" More Information needed text10K<n<100K1 likes7 downloads3y agoHugging Face30kraalfar /Coeditor-processed-demo Dataset Card for "Coeditor-processed-demo" More Information needed textn<1K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.