CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Cross-Mergeability /merge-accuracy merge-accuracy — does aligning a merge improve DOWNSTREAM ACCURACY? The mergeability line of work measures merge obstruction in nats/token. This dataset supplies the missing axis: task accuracy, on real released models, for the merge recipe practitioners actually run — the chat-vector recipe theta_new = theta_fork + lambda * ( theta_instruct - theta_base ) with meta-llama/Llama-3.1-8B, its official Instruct release, and three community continued-pretrained language forks… See the full description on the dataset page: https://huggingface.co/datasets/Cross-Mergeability/merge-accuracy.textn<1K0 likes179 downloads28d agoHugging Face02Mr-Rosen /Accuracy-Is-Not-Enough-FinQA-Dataset Natively Extended FinQA Natively Extended FinQA is a long-context derivative of FinQA for numerical reasoning over financial data. It preserves the FinQA task while increasing the amount of financial context surrounding each question. Average context increased from approximately: 611 words per question to: 5,629 words per question Splits Hugging Face Split File Records train long_train.json 6,251 validation long_dev.json 883 test long_test.json 1… See the full description on the dataset page: https://huggingface.co/datasets/Mr-Rosen/Accuracy-Is-Not-Enough-FinQA-Dataset.textquestion-answering1K<n<10K1 likes114 downloads1mo agoHugging Face03jason23322 /high-accuracy-email-classifiergated High-Accuracy Email Classification Dataset Dataset Description This dataset contains 12,000+ emails across 6 categories, specifically curated for high-accuracy email classification tasks. The dataset achieves 98%+ classification accuracy with appropriate models. Categories The dataset includes emails from the following categories: Category Count Description Emoji Forum ~2,000 Forum posts, discussions, and community notifications 🗣️ Promotions ~2… See the full description on the dataset page: https://huggingface.co/datasets/jason23322/high-accuracy-email-classifier.texttext-classification10K<n<100K21 likes50 downloads1y agoHugging Face04referencesource /legal-metrology-accuracy-classes Accuracy classes and maximum permissible errors for trade measuring instruments (EU MID) Canonical, always-current version: https://referencesource.org/legal-metrology-accuracy-classes/ Machine-readable: https://referencesource.org/legal-metrology-accuracy-classes/data.json — this mirror is a point-in-time copy. Last verified: 2026-08-05 Stale after: 2027-08-05 (past this date, prefer the canonical copy — it re-verifies on a cadence this snapshot does not) Records: 70 Accuracy… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/legal-metrology-accuracy-classes.textn<1K0 likes43 downloads28d agoHugging Face05MichaelAnthony /echidna-accuracy-massive echidna-accuracy-massive Echidna RAG assistant — large accuracy-focused extraction/QA dataset. Contents accuracy_massive.jsonl (37 rows) Format JSON Lines (.jsonl), one example per line. Provenance Original content for the Echidna RAG assistant (Michael Anthony Falabella). textquestion-answeringn<1K0 likes38 downloads28d agoHugging Face06MichaelAnthony /echidna-round2-accuracy echidna-round2-accuracy Echidna — round 2 accuracy-focused extraction examples. Contents round2_accuracy.jsonl (20 rows) Format JSON Lines (.jsonl), one example per line. Provenance Original content for the Echidna RAG assistant (Michael Anthony Falabella). textquestion-answeringn<1K0 likes38 downloads28d agoHugging Face07MichaelAnthony /echidna-round4-extensive-accuracy echidna-round4-extensive-accuracy Echidna — round 4 extensive accuracy extraction examples. Contents round4_extensive_accuracy.jsonl (8 rows) Format JSON Lines (.jsonl), one example per line. Provenance Original content for the Echidna RAG assistant (Michael Anthony Falabella). textquestion-answeringn<1K0 likes32 downloads28d agoHugging Face08chairulridjal /high-accuracy-email-classifier-indonesian High-Accuracy Email Classification Dataset Indonesian Translation This dataset is an Indonesian translation/adaptation of jason23322/high-accuracy-email-classifier. Contribution The original dataset, labels, IDs, and split membership come from jason23322/high-accuracy-email-classifier, released under the Apache 2.0 license. This repository contributes the Indonesian translation: subject, body, and text are translated into Indonesian. id, category, and category_id… See the full description on the dataset page: https://huggingface.co/datasets/chairulridjal/high-accuracy-email-classifier-indonesian.texttext-classification10K<n<100K0 likes24 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.