CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nguyenthanhasia /vsec-vietnamese-spell-correction VSEC: Vietnamese Spell Correction Dataset Dataset Description VSEC (Vietnamese Spell Correction) is a comprehensive dataset for Vietnamese spelling error detection and correction, containing 9,341 sentences with 11,202 human-made misspellings across 5,211 unique error types. This dataset represents the largest publicly available collection of Vietnamese spelling errors with syllable-level annotations, making it an invaluable resource for developing and evaluating… See the full description on the dataset page: https://huggingface.co/datasets/nguyenthanhasia/vsec-vietnamese-spell-correction.texttext-generation1K<n<10K5 likes118 downloads1y agoHugging Face02Khamoon /asr-spell-correction-rutext1K<n<10K0 likes72 downloads14d agoHugging Face03Dan032 /asr_spell_correction_rutextn<1K0 likes55 downloads14d agoHugging Face04Alexander-Usov /ru-asr-spell-correctiontext1K<n<10K0 likes47 downloads1h agoHugging Face05Alexander-Usov /ru-asr-spell-correction-v2text1K<n<10K0 likes46 downloads2d agoHugging Face06elinaail /asr-spell-correction-rutext1K<n<10K0 likes41 downloads9d agoHugging Face07DariaZah /asr_spell_correction_rutext1K<n<10K0 likes36 downloads7d agoHugging Face08sobadsodead /asr-spell-correction-ru-hw01 Russian ASR correction: homework 01 1020 pairs: 450 Groq-generated ASR-like inputs, 180 Groq-generated numeral-to-word pairs, and 390 identity examples added by copying screened clean targets. Model: openai/gpt-oss-120b. Generation: 8bae1a605e8506a1; prompt version: groq_asr_numbers_v2. The ASR target names come from the existing Groq-generated pool targets.jsonl. No Python character corruption is used. This is synthetic text, not real ASR output. Generation and… See the full description on the dataset page: https://huggingface.co/datasets/sobadsodead/asr-spell-correction-ru-hw01.text1K<n<10K0 likes35 downloads14d agoHugging Face09aligh4699 /persian-spell-correction-dataset Persian Spell Correction & Augmentation Dataset This is a large-scale, parallel dataset for Persian spell correction, text normalization, and augmentation. It is designed to train and evaluate models for correcting a wide variety of common and synthetic errors in Persian text. The dataset is built from two main components: Natural Data: Text from diverse Persian corpora and its corresponding clean, corrected version (corrected_text) generated by an LLM. Augmented Data: The… See the full description on the dataset page: https://huggingface.co/datasets/aligh4699/persian-spell-correction-dataset.text1M<n<10M1 likes34 downloads11mo agoHugging Face10Ruslan1995 /russian-asr-spell-correctiontext1K<n<10K0 likes30 downloads2d agoHugging Face11polinOchka33 /spell_correction_rutext1K<n<10K0 likes29 downloads5d agoHugging Face12SerejkaP /popular_names_spell_correctiontext1K<n<10K1 likes25 downloads1y agoHugging Face13Ruslan1995 /russian-asr-spell-correction_v2text1K<n<10K0 likes25 downloads2d agoHugging Face14rubin5341 /my-groq-spell-correction-mistake-datasettext1K<n<10K0 likes17 downloads1y agoHugging Face15rubin5341 /groq-spell-correction-mistake-datasettext1K<n<10K0 likes15 downloads1y agoHugging Face16nikfil /russian-spell-correction-datasettext1K<n<10K0 likes15 downloads1y agoHugging Face17Ruslan1995 /russian-asr-spell-correction_v3text1K<n<10K0 likes15 downloads17h agoHugging Face18DenK-huggingFace /russian-spell-correction-dataset-2text1K<n<10K0 likes14 downloads1y agoHugging Face19GoshaNice /grok_spell_correction_datatext1K<n<10K1 likes14 downloads1y agoHugging Face20rubin5341 /groq1441-spell-correction-mistake-datasettext1K<n<10K0 likes14 downloads1y agoHugging Face21Andrey32 /russian-spell-correction-groqtext1K<n<10K0 likes14 downloads1y agoHugging Face22gfrslf /russian_spell_correction_groqtext1K<n<10K0 likes13 downloads1y agoHugging Face23minhbui /spell_correction_datasettext1M<n<10M0 likes12 downloads2y agoHugging Face24Lak1N /spell_correction_datasettext1K<n<10K0 likes12 downloads1y agoHugging Face25nesemenpolkov /syntetic-dataset-for-gigaam-spell-correctionaudio1K<n<10K0 likes10 downloads2y agoHugging Face26nlprunnerup /russian-spell-correction-datasettext1K<n<10K0 likes9 downloads11mo agoHugging Face27ArchiRad /russian-spell-correction-datasettext1K<n<10K0 likes9 downloads11mo agoHugging Face28psaw77 /russian-spell-correction-datasettext1K<n<10K0 likes8 downloads1y agoHugging Face29KailPoppu /russian-spell-correction-datasettext1K<n<10K0 likes6 downloads1y agoHugging Face30Aza06 /russian-asr-spell-correction_vers2text1K<n<10K0 likes6 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.