CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nyu-dice-lab /lm-eval-results-shyamieee-Padma-SLM-7b-v1.0-private Dataset Card for Evaluation run of shyamieee/Padma-SLM-7b-v1.0 Dataset automatically created during the evaluation run of model shyamieee/Padma-SLM-7b-v1.0 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-Padma-SLM-7b-v1.0-private.tabular100K<n<1M0 likes515 downloads2y agoHugging Face02nyu-dice-lab /lm-eval-results-shyamieee-Padma-SLM-7b-v3.0-private Dataset Card for Evaluation run of shyamieee/Padma-SLM-7b-v3.0 Dataset automatically created during the evaluation run of model shyamieee/Padma-SLM-7b-v3.0 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-Padma-SLM-7b-v3.0-private.tabular100K<n<1M0 likes302 downloads2y agoHugging Face03dw-indie /pad-auto-solver-reviewed PAD Reviewed Dataset Canonical reviewed PAD board/orb artifacts for dw-indie/pad-auto-solver-reviewed. This repository contains immutable reviewed package revisions and does not contain raw captures, training runs, checkpoints, or model binaries. Packages exported: 28 Active catalog datasets: 14 Catalog schema: 3 Layout packages/<dataset_id>.tar: deterministic self-contained reviewed package catalog.json: active revision heads and coverage summary… See the full description on the dataset page: https://huggingface.co/datasets/dw-indie/pad-auto-solver-reviewed.tabularimage-classification10K<n<100K1 likes205 downloads16d agoHugging Face04PaDaS-Lab /nfqa-multilingual-dataset NFQA Multilingual Dataset A large-scale multilingual dataset for Non-Factoid Question Answering (NFQA) classification, covering 49 languages and 8 question categories. Dataset Statistics Split Examples Train 28,653 Validation 3,539 Test 3,671 Total (Balanced) 35,863 Full Dataset (High Quality) 63,647 Dataset Composition Languages (49 total) Arabic (ar), Azerbaijani (az), Bulgarian (bg), Bengali (bn), Catalan (ca)… See the full description on the dataset page: https://huggingface.co/datasets/PaDaS-Lab/nfqa-multilingual-dataset.tabulartext-classification10K<n<100K1 likes165 downloads6mo agoHugging Face05JonathanZha /PADBen PADBen: Paraphrase and AI-Generated Text Detection Benchmark 📊 Dataset Overview PADBen is a comprehensive benchmark for evaluating AI-generated text detection methods, specifically designed to test detection capabilities across various paraphrasing scenarios and attack vectors. For detailed implementation of how this dataset is generated/curated, please see https://github.com/JonathanZha47/PadBen-Paraphrase-Attack-Benchmark. Total Dataset Size: 486,990 samples across 46… See the full description on the dataset page: https://huggingface.co/datasets/JonathanZha/PADBen.tabulartext-classification100K<n<1M0 likes47 downloads11mo agoHugging Face06padamenko /swemera-10tasks-pyconfHere’s your Markdown text, organized for clear readability: Task Description Instances Instance ID Short Title reframe-0 Performance threshold goes to -inf when it should be zero. pyflakes-1 Walrus operator + annotation can cause F821 sqlglot-2 MySQL dialect fails to parse PRIMARY KEY USING BTREE syntax matchms-3 matchms fails when reading spectra where abundance is in scientific notation #809 guarddog-4 Add Mach-O magic bytes to bundled binary detector… See the full description on the dataset page: https://huggingface.co/datasets/padamenko/swemera-10tasks-pyconf.tabularn<1K1 likes42 downloads1y agoHugging Face07JonathanZha /PADBen-Task1 PADBen Task 1: Paraphrase Source Attribution (Binary Classification) 📋 Dataset Summary PADBen Task 1 is a binary classification dataset for distinguishing between human-authored and LLM-generated paraphrases. This task evaluates whether AI detectors can identify the source of paraphrased text without additional context. Key Features Task Type: Binary text classification Total Samples: 16,233 sentences Train Split: 12,986 samples (80%) Test Split: 3,247… See the full description on the dataset page: https://huggingface.co/datasets/JonathanZha/PADBen-Task1.tabulartext-classification10K<n<100K0 likes16 downloads1y agoHugging Face08huiliu123 /phyworld-data-pad_featurestabularn<1K0 likes13 downloads5mo agoHugging Face09open-llm-leaderboard /shyamieee__Padma-v7.0-detailsgated Dataset Card for Evaluation run of shyamieee/Padma-v7.0 Dataset automatically created during the evaluation run of model shyamieee/Padma-v7.0 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/shyamieee__Padma-v7.0-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face10rafmacalaba /pad-pid-geospatial Geospatial data-use annotations in World Bank PADs and PIDs Focus-review annotations joined to Project Appraisal Documents (PADs) and Project Information Documents (PIDs) by canonical document id. This export preserves the complete reviewed passage rows and all focus labels—not only spatially linked mentions—alongside document/project metadata and origin fields. What's included 3,804 passage records across 1,591 documents; exact metadata types are PAD/PID. 6,503… See the full description on the dataset page: https://huggingface.co/datasets/rafmacalaba/pad-pid-geospatial.tabular1K<n<10K0 likes1h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.