CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01aarajbhattarai /nepali-law-v2-corrected Dataset Card for nepali-law-v2-corrected (V2) Version 2.0.0 — a curated, audited correction of aarajbhattarai/nepali-law-v2 (revision aa71fbe2b22310d45f86e3b429d3815817a33574). This card describes V2. The original V1 dataset is unmodified and remains the upstream source of truth. Every statistic here was computed from the released V2 files by the release audit pipeline (scripts/validate_release.py and the project's EDA notebooks, which are retained with the project rather than… See the full description on the dataset page: https://huggingface.co/datasets/aarajbhattarai/nepali-law-v2-corrected.texttext-generation10K<n<100K0 likes137 downloads5d agoHugging Face02onur48 /MetaMathQA-Turkish-correctedtext100K<n<1M2 likes42 downloads2y agoHugging Face03rafmacalaba /datause-dataset-corrected datause-dataset (re-chunked) Re-chunk of ai4data/datause-dataset into <=384-token windows (max_tokens=384, overlap=50) so GLiNER's window is never truncated during evaluation. {"train": {"orig": 1779, "rechunked": 1822}, "validation": {"orig": 415, "rechunked": 458}, "holdout": {"orig": 1149, "rechunked": 1166}} texttoken-classification1K<n<10K0 likes31 downloads1mo agoHugging Face04jzhang86 /corrected_ifevalAfter developing multi-lingual IFEVAl, we found he original IFEVAL dataset has one error at key 1174 {"let_relation": "less than", "letter": "o", "let_frequency": 6} should be corrected to: {"let_relation": "at least", "letter": "o", "let_frequency": 6} Here is the original key 1174 content {"key": 1174, "prompt": "Write a template for a chat bot that takes a user's location and gives them the weather forecast. Use the letter o as a keyword in the syntax of the template. The letter o should… See the full description on the dataset page: https://huggingface.co/datasets/jzhang86/corrected_ifeval.textn<1K3 likes23 downloads2y agoHugging Face05Mathieu-Thomas-JOSSET /fine_tome_style_sharegpt_corrected_tabs_random_datetime.jsonl fine_tome_style_sharegpt_corrected_tabs_random_datetime.jsonl Files train.jsonl Loading (example) from datasets import load_dataset ds = load_dataset("Mathieu-Thomas-JOSSET/fine_tome_style_sharegpt_corrected_tabs_random_datetime.jsonl") print(ds) textn<1K0 likes7 downloads9mo agoHugging Face06RuihanCao /bird-dev-correctedtabularn<1K0 likes3 downloads6mo agoHugging Face07autoprogrammer /humanevalplus_correctedtextn<1K0 likes2 downloads1y agoHugging Face08cowslovecats555 /Liya_Reminder_Dataset_Correctedtextn<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.