CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01UARK-NED3 /BoilingBench-CV BoilingBench-CV Dataset Version: v0.1.0 Maintainer: NED3 Laboratory, University of Arkansas License: CC BY 4.0 DOI: 10.5281/zenodo.22264378 Mirror of the Zenodo deposit of 3 September 2026, published here because most users of these data work in the Hugging Face ecosystem. The file set was verified identical to the deposit at upload time: 7,147 files, 4.20 GB uncompressed. Authors Hari Pandey (University of Arkansas), Manohar Bongarala (Purdue University), Christy… See the full description on the dataset page: https://huggingface.co/datasets/UARK-NED3/BoilingBench-CV.imageimage-segmentationn<1K1 likes2.4k downloads7d agoHugging Face02stasvinokur /cve-and-cwe-dataset-1999-2025This collection brings together every Common Vulnerabilities & Exposures (CVE) entry published in the National Vulnerability Database (NVD) from the very first identifier — CVE-1999-0001 — through all records available on 30 May 2025. It was built automatically with a Python script that calls the NVD REST API v2.0 page-by-page, handles rate-limits, and filters data. After download each CVE object is pared down to the essentials and written to CVE_CWE_2025.csv with the following columns:… See the full description on the dataset page: https://huggingface.co/datasets/stasvinokur/cve-and-cwe-dataset-1999-2025.tabulartext-classification100K<n<1M10 likes1.1k downloads1y agoHugging Face03jason1966 /aikyatansinha_cybersecurity-cves-for-nlp-dataset Cybersecurity CVEs for NLP Dataset Every CVE since 1999, scrubbed and perfectly formatted for NLP tasks Dataset Info Source: Kaggle Original Size: 38.28 MB Kaggle Downloads: 36 Files: 1 Files NVD_Cybersecurity_Dataset.csv Mirrored from Kaggle tabular100K<n<1M3 likes110 downloads6mo agoHugging Face04electricsheepafrica /africa-synth-hypertension-hypertension-cvd-dataset-all African Hypertension & CVD Synthetic Dataset | Africa (Electric Sheep Africa metadata inventory) Size category: 10K<n<100K - Formats: csv - Sector: health - Engineered by Electric Sheep Africa TL;DR This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context. What This Dataset Covers Health datasets… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-synth-hypertension-hypertension-cvd-dataset-all.tabulartabular-classification1K<n<10K0 likes93 downloads1mo agoHugging Face05geethanjali-cv /stock-technical-indicators Stock Technical Indicators Dataset Historical technical indicators dataset used to train directional stock movement classifiers. Features RSI: Relative Strength Index SMA_20 / EMA_50: Simple and Exponential Moving Averages MACD: Moving Average Convergence Divergence Target: Directional label (1 = Bullish, 0 = Bearish) tabulartabular-classificationn<1K1 likes92 downloads17d agoHugging Face06regularpooria /CVE_CWE_Software_Mapping_Dataset CVE-CWE Software Weakness Mapping Dataset Dataset description This dataset maps Common Vulnerabilities and Exposures (CVEs) to Common Weakness Enumeration (CWE) entries in the CWE-699 Software category. It combines CVE descriptions with CWE descriptions and parent-category information for security research and vulnerability classification. Dataset structure The dataset is provided as Global_Dataset.csv. Its main fields include: CVE-ID: CVE… See the full description on the dataset page: https://huggingface.co/datasets/regularpooria/CVE_CWE_Software_Mapping_Dataset.tabulartext-classification10K<n<100K0 likes67 downloads12d agoHugging Face07cvmistralparis /Q20LLM 20 Questions game with LLM This dataset generated with the following LLMs: Groq API / llama3-70b-8192 Groq API / mixtral-8x7b-32768 Mistral API / mistral-large-latest Test keywords based on newlist_things.rmdup.test.txt from Entity-Deduction Arena (EDA) project. The dataset generated in two stages: LLM was prompt to generate keywords Dialog with different length generated for each keyword Keywords prompt Generate a list of 500 diverse and simple keywords suitable… See the full description on the dataset page: https://huggingface.co/datasets/cvmistralparis/Q20LLM.tabularquestion-answering10K<n<100K3 likes47 downloads2y agoHugging Face08zefang-liu /cve-and-cwe-mapping-dataset CVE and CWE Mapping Dataset This Hugging Face dataset is a partial copy of the 'CVE and CWE mapping Dataset (2021)' from Kaggle, featuring 'Global_Dataset.csv' originally as 'Global_Dataset.xlsx'. Created by Kirushikesh DB and shared under CC BY-NC-SA 4.0, it includes CVE data up to 2021 for cybersecurity research. For full details and licensing, visit the original Kaggle page. For further information, please review the CVE Terms of Use and the NVD Terms of Use. tabulartext-classification100K<n<1M9 likes46 downloads3y agoHugging Face09GilatToker /Liberty-CV LIBERTy-CV Dataset Overview LIBERTy-CV is one of the three datasets released as part of the LIBERTy (LLM-based Interventional Benchmark for Explainability with Real Targets) benchmark. The goal of LIBERTy is to evaluate concept-based explanation methods in NLP under a causal and counterfactual framework.Each dataset in the benchmark is designed to expose spurious correlations between high-level concepts and model predictions, and to enable quantitative evaluation of… See the full description on the dataset page: https://huggingface.co/datasets/GilatToker/Liberty-CV.tabulartext-classification1K<n<10K0 likes40 downloads9mo agoHugging Face10omaraboelmaaty /arabic-cv-scoring-dataset Arabic CV Scoring Dataset Dataset Summary This dataset contains ~7,220 synthetically generated Arabic CVs, each paired with a job category, an ATS (Applicant Tracking System) compatibility score, and a suitability score/class label. It was built to train and evaluate the Arabic CV Analyzer — an NLP pipeline that scores, classifies, and generates improvement suggestions for Arabic CVs targeting the Arab job market, where no equivalent ATS-optimization tooling… See the full description on the dataset page: https://huggingface.co/datasets/omaraboelmaaty/arabic-cv-scoring-dataset.tabulartext-classification1K<n<10K0 likes36 downloads1mo agoHugging Face11TechPlayground /bitcoin-ethereum-orderflow-cvd-alpha Bitcoin & Ethereum 1-Minute Order Flow & Cumulative Volume Delta (CVD) Alpha Institutional Market Microstructure Dataset Sample (Clean CSV / Parquet Ready) 📌 Dataset Overview In cryptocurrency and traditional electronic markets, price action is driven by aggressive market orders (taker flow) that cross the bid-ask spread. This preview dataset provides 1,000 rows of continuous 1-minute order flow for Bitcoin (BTC/USDT) and Ethereum (ETH/USDT)… See the full description on the dataset page: https://huggingface.co/datasets/TechPlayground/bitcoin-ethereum-orderflow-cvd-alpha.tabular1K<n<10K0 likes36 downloads9d agoHugging Face12CVPR2024 /CVPR2024-paperstabular1K<n<10K1 likes35 downloads2y agoHugging Face13readerbench /cve-2-att-cktabular1K<n<10K0 likes29 downloads2y agoHugging Face14SECT19N /Preprocessed-CVS-24-KMRtabular10K<n<100K0 likes29 downloads7mo agoHugging Face15eromang /cyberscale-training-cves CyberScale Training CVEs Training dataset for the CyberScale vulnerability severity scorer. Contains 30,641 CVEs with CVSS v3.x scores, descriptions, and CWE classifications. Schema Column Type Description cve_id string CVE identifier (e.g., CVE-2024-1234) description string Vulnerability description (English) cvss_score float CVSS v3.x base score (0.0-10.0) cvss_version string CVSS version (3.0 or 3.1) cwe string CWE identifier (e.g., CWE-79), may be… See the full description on the dataset page: https://huggingface.co/datasets/eromang/cyberscale-training-cves.tabulartext-classification10K<n<100K0 likes27 downloads6mo agoHugging Face16SECT19N /Preprocessed-CVS-24-CKBtabular10K<n<100K0 likes23 downloads8mo agoHugging Face17ukcli /cve-and-cwe-mapping-dataset CVE and CWE Mapping Dataset This Hugging Face dataset is a partial copy of the 'CVE and CWE mapping Dataset (2021)' from Kaggle, featuring 'Global_Dataset.csv' originally as 'Global_Dataset.xlsx'. Created by Kirushikesh DB and shared under CC BY-NC-SA 4.0, it includes CVE data up to 2021 for cybersecurity research. For full details and licensing, visit the original Kaggle page. For further information, please review the CVE Terms of Use and the NVD Terms of Use. tabulartext-classification100K<n<1M0 likes21 downloads5mo agoHugging Face18ioget /aims-traffic-cv-datadocumentn<1K0 likes16 downloads5mo agoHugging Face19ProjectFisokuhle /cv-corpus-22.0-2025-06-20audion<1K0 likes13 downloads11mo agoHugging Face20Amutuhaire /africa-synth-hypertension-hypertension-cvd-dataset-all ⚠️ Synthetic dataset — Parameterized from published SSA literature, not real observations. Not suitable for empirical analysis or policy inference. African Hypertension & Cardiovascular Disease Dataset Screening, Risk Stratification, and CVD Event Prediction Version: 1.0Release Date: November 2024Context: Sub-Saharan Africa (25-35% adult HTN prevalence, 70-80% undiagnosed)License: Research & Educational Use Abstract We present synthetic datasets for… See the full description on the dataset page: https://huggingface.co/datasets/Amutuhaire/africa-synth-hypertension-hypertension-cvd-dataset-all.tabulartabular-classification10K<n<100K0 likes13 downloads4mo agoHugging Face21reach-vb /open-asr-leaderboard-evals-ex-cvtabularn<1K0 likes9 downloads3y agoHugging Face22cvelist /KEV_EPSStabular1K<n<10K1 likes8 downloads2y agoHugging Face23bwbayu /job_cv_supervisedtabular10K<n<100K3 likes8 downloads2y agoHugging Face24Krm1 /CVE-2025tabular1K<n<10K0 likes8 downloads2y agoHugging Face25e-sdmartinez /cvembeddingstabular1K<n<10K0 likes7 downloads3y agoHugging Face26cvelist /CVEDBtabular1K<n<10K1 likes7 downloads1y agoHugging Face27Baction /cvetabular10K<n<100K0 likes6 downloads1y agoHugging Face28sulaimank /lug-one-cvgatedtabular100K<n<1M0 likes5 downloads2y agoHugging Face29CVPR2024 /CVPR2024-paper-statstabular1K<n<10K0 likes3 downloads2y agoHugging Face30BSYM-25 /indic-cv-validated Indic Mozilla Common Voice Validated Metadata This dataset contains validated metadata for various Indic languages from Mozilla Common Voice. Validation includes: SNR (Signal-to-Noise Ratio) Silence Ratio Clipping Detection Unicode Normalization Mixed Script Detection Current Languages Assamese: 200 rows (Validated) Validation Stats snr: Signal-to-noise ratio in dB. duration: Audio duration in seconds. validation_issues: Semicolon-separated list of… See the full description on the dataset page: https://huggingface.co/datasets/BSYM-25/indic-cv-validated.tabularn<1K0 likes3 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.