CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01unitedideas /practice-radar-behavioral-health-npi-sample New behavioral-health organization NPIs — weekly NPPES sample A 15-row public sample from a weekly, reproducible selection of newly enumerated Type 2 behavioral-health organizations in the U.S. Centers for Medicare & Medicaid Services National Plan and Provider Enumeration System (NPPES). Edition at a glance Measured period: July 6–12, 2026 New Type 2 organizations screened: 2,722 Behavioral-health organizations selected: 486 States and territories represented:… See the full description on the dataset page: https://huggingface.co/datasets/unitedideas/practice-radar-behavioral-health-npi-sample.tabularn<1K0 likes1.8k downloads2mo agoHugging Face02trader298 /sec-nport SEC Form N-PORT Data Sets Monthly portfolio holdings reported by registered investment companies and ETFs on Form N-PORT, published by the U.S. SEC as quarterly structured data sets and mirrored here as typed, partitioned Parquet — queryable directly from DuckDB. Source: SEC Form N-PORT Data Sets — public domain (U.S. Government work) https://www.sec.gov/data-research/sec-markets-data/form-n-port-data-sets Coverage: October 2019 onward, refreshed quarterly Format: one Parquet… See the full description on the dataset page: https://huggingface.co/datasets/trader298/sec-nport.tabular100M<n<1B0 likes1.4k downloads2mo agoHugging Face03pk1308 /digenai-nppe-datasettabular100K<n<1M0 likes1.2k downloads26d agoHugging Face04Vancheeswaran /digenai-nppe-datasettabularn<1K0 likes1.2k downloads26d agoHugging Face0522f2001542 /dlgenai-nppe2-datasettabularn<1K0 likes591 downloads3d agoHugging Face06Ouroboros-Research /llama-9b-bulk-npztabularn<1K0 likes540 downloads14d agoHugging Face07colabfit /OMat24_train_aimd_from_PBE_1000_npt Cite this dataset Barroso-Luque, L., Shuaibi, M., Fu, X., Wood, B. M., Dzamba, M., Gao, M., Rizvi, A., Zitnick, C. L., and Ulissi, Z. W. OMat24 train aimd from PBE 1000 npt. ColabFit, 2025. https://doi.org/10.60732/25f16f85 This dataset has been curated and formatted for the ColabFit Exchange This dataset is also available on the ColabFit Exchange: https://materials.colabfit.org/id/DS_jqrkc9e7cgmh_0 Visit the ColabFit Exchange to search… See the full description on the dataset page: https://huggingface.co/datasets/colabfit/OMat24_train_aimd_from_PBE_1000_npt.tabular10M<n<100M0 likes535 downloads11mo agoHugging Face08nphearum /gpt-5.5-agentThis dataset was generated using teich by TeichAI Prepare these datasets for supervised fine-tuning in just a few lines of code — see the Conversion section below. gpt 5.5 Agent Traces This directory contains raw agent trace files generated by teich. (I also dropped in some of my own personal traces) All assistant responses were generated by openai/gpt-5.5. JSONL files: 88 Training-ready tools A complete configured tools schema snapshot is embedded in the… See the full description on the dataset page: https://huggingface.co/datasets/nphearum/gpt-5.5-agent.tabularn<1K1 likes451 downloads4mo agoHugging Face09open-index /open-npm npm Registry - Complete Package Archive Every npm package with full metadata, versions, dependencies, and download stats What is it? This dataset contains a comprehensive snapshot of the npm registry, the default package manager for Node.js and the largest software package registry in the world. npm hosts millions of packages and serves billions of downloads every week. If you have ever run npm install, you have used the registry that this dataset mirrors. The archive… See the full description on the dataset page: https://huggingface.co/datasets/open-index/open-npm.tabulartext-classification10M<n<100M2 likes400 downloads5mo agoHugging Face10123Ashwani /dlgenai-nppe-datasettabularn<1K0 likes358 downloads29d agoHugging Face11colabfit /OMat24_train_aimd_from_PBE_3000_npt Cite this dataset Barroso-Luque, L., Shuaibi, M., Fu, X., Wood, B. M., Dzamba, M., Gao, M., Rizvi, A., Zitnick, C. L., and Ulissi, Z. W. OMat24 train aimd from PBE 3000 npt. ColabFit, 2025. https://doi.org/10.60732/edd12490 This dataset has been curated and formatted for the ColabFit Exchange This dataset is also available on the ColabFit Exchange: https://materials.colabfit.org/id/DS_6xvvh8yl7rfd_0 Visit the ColabFit Exchange to search… See the full description on the dataset page: https://huggingface.co/datasets/colabfit/OMat24_train_aimd_from_PBE_3000_npt.tabular1M<n<10M0 likes295 downloads11mo agoHugging Face12GeoMeterData /nphard_tsp2tabularn<1K0 likes258 downloads2y agoHugging Face13YanyiPU716 /ECtHR-NPD ECtHR-NPD — unified current release One current dataset, not separate v1.0/v1.1 releases. This author-approved distribution is available for manual review. It preserves the paper's 14,575-case cohort and original partitions, with documented amount corrections. It is not a byte-identical copy of the targets used for the original experiments. Data data/case_level.csv: 14,575 cases; 33 columns. data/applicant_level.csv: 44,581 applicant source units; 14 columns.… See the full description on the dataset page: https://huggingface.co/datasets/YanyiPU716/ECtHR-NPD.tabular10K<n<100K0 likes240 downloads6d agoHugging Face14COINjecture /NP_Solutions_v2 🔬 COINjecture NP Solutions Dataset v2 Institutional-Grade Blockchain Research Data A comprehensive, real-time dataset of NP-complete problem solutions generated through Proof-of-Useful-Work (PoUW) blockchain consensus Overview • Data Schema • Metrics Categories • Usage • Citation 📋 Overview This dataset contains institutional-grade metrics from the COINjecture Network B blockchain, which implements a novel Proof-of-Useful-Work (PoUW) consensus… See the full description on the dataset page: https://huggingface.co/datasets/COINjecture/NP_Solutions_v2.tabularother10K<n<100K0 likes211 downloads10mo agoHugging Face15ZipLime /fund-holdings-nport US Fund and ETF Holdings — Form N-PORT Every position of every US mutual fund and ETF, monthly, including the bonds, loans, asset-backed paper and derivatives that 13F does not report at all. 145 630 731 positions · 17 513 funds · 341 049 monthly reports · 2019-09-30 to 2026-05-31 The pipeline lives in recipe/ at the same revision as the data. See PIPELINE.md for the method. Why this and not 13F 13F is what everyone uses because it is what everyone knows about. It… See the full description on the dataset page: https://huggingface.co/datasets/ZipLime/fund-holdings-nport.tabulartabular-regression100M<n<1B0 likes208 downloads15d agoHugging Face16NP235 /finemath 📐 FineMath What is it? 📐 FineMath consists of 34B tokens (FineMath-3+) and 54B tokens (FineMath-3+ with InfiMM-WebMath-3+) of mathematical educational content filtered from CommonCrawl. To curate this dataset, we trained a mathematical content classifier using annotations generated by LLama-3.1-70B-Instruct. We used the classifier to retain only the most educational mathematics content, focusing on clear explanations and step-by-step problem solving rather than… See the full description on the dataset page: https://huggingface.co/datasets/NP235/finemath.tabular10M<n<100M0 likes159 downloads2mo agoHugging Face17COINjecture /NP_Solutions_v4 COINjecture NP Solutions v4 Dataset Description This dataset contains verified solutions to NP-hard computational problems from the COINjecture Network B blockchain. Version 4 Features ADZDB Storage: File-based Append-Delete-Zero Database for efficient block storage Unified Streaming: All problem types in one continuous dataset Real-time Updates: Solutions streamed as blocks are mined Problem Types TSP (Traveling Salesman Problem) 3SAT (Boolean… See the full description on the dataset page: https://huggingface.co/datasets/COINjecture/NP_Solutions_v4.tabularother10K<n<100K1 likes153 downloads10mo agoHugging Face18chloecodes /NPRVideoEmbeddingstabular10K<n<100K0 likes151 downloads3y agoHugging Face19ChipYTY /titans_NPC Titans - Pytorch Unofficial implementation of Titans in Pytorch. Will also contain some explorations into architectures beyond their simple 1-4 layer MLP for the neural memory module, if it works well to any degree. Paper review by Yannic Quick Colab Run Appreciation Eryk for sharing his early experimental results with me, positive for 2 layer MLP Install $ pip install titans-pytorch Usage import torch from titans_pytorch import… See the full description on the dataset page: https://huggingface.co/datasets/ChipYTY/titans_NPC.tabularn<1K0 likes138 downloads8mo agoHugging Face20COINjecture /NP_Solutions_v3 🔬 COINjecture NP Solutions Dataset v3 Institutional-Grade Blockchain Research Data A comprehensive, real-time dataset of NP-complete problem solutions generated through Proof-of-Useful-Work (PoUW) blockchain consensus Overview • Data Schema • Metrics • Pipeline• Usage • Citation 📋 Overview This dataset contains institutional-grade metrics from the COINjecture Network B blockchain, which implements a novel Proof-of-Useful-Work (PoUW) consensus mechanism.… See the full description on the dataset page: https://huggingface.co/datasets/COINjecture/NP_Solutions_v3.tabularother10K<n<100K0 likes136 downloads10mo agoHugging Face21npaleti2002 /World_Bank_GDP_by_Country_and_Continent_2000-2024 World Bank GDP by Country & Continent (2000–2025) A clean, visual, non-technical analysis of World Bank GDP (current US$) organized by country and summarized to a 7-continent view. All values are expressed in USD (billions) for readability. Why this repo exists To provide a single place to: Extract authoritative GDP data from the World Bank (2000–2025), Publish & share a tidy dataset on Kaggle, Analyze & present insights in a visual, plain-English notebook.… See the full description on the dataset page: https://huggingface.co/datasets/npaleti2002/World_Bank_GDP_by_Country_and_Continent_2000-2024.tabularn<1K1 likes134 downloads1y agoHugging Face22Madnesss /npy_file_hsrtabular1K<n<10K0 likes127 downloads2y agoHugging Face23nprak26 /remote-worker-productivity Key Features: Primary Research Focus: Age vs Productivity correlation Years of Experience impact on remote work effectiveness WFH Days per Week optimal balance analysis Multiple productivity metrics (not just one score) Dataset Highlights: 1,500 rows - Perfect size for analysis 30+ columns - Rich feature set Realistic correlations - Built-in meaningful relationships Clean data - No missing values, proper data types Multiple target variables - 5 different… See the full description on the dataset page: https://huggingface.co/datasets/nprak26/remote-worker-productivity.tabulartabular-classification1K<n<10K11 likes123 downloads1y agoHugging Face24shunanhe /NPM-Artifact-Explanation-Benchmark NPM-Artifact-Explanation-Benchmark English NPM-Artifact-Explanation-Benchmark is a cross-category multimodal corpus and benchmark resource for Chinese cultural artifact understanding and explanation. This release contains 28,826 cleaned artifact records derived from National Palace Museum source records' opendata (https://digitalarchive.npm.gov.tw/opendata/). Each record includes structured artifact metadata, image URLs, source record URLs, and human-written… See the full description on the dataset page: https://huggingface.co/datasets/shunanhe/NPM-Artifact-Explanation-Benchmark.tabularimage-to-text10K<n<100K1 likes121 downloads8d agoHugging Face25deepklarity /top-npm-packagesTop NPM Packages Dataset This dataset contains a snapshot of Top 5200+ popular node packages hosted on Node Package Manager The dataset was scraped in October-2024. We aim to use this dataset to perform analysis and identify trends and get a bird's eye view of nodejs ecosystem. Mantainers: Nishritha Damera tabular1K<n<10K6 likes100 downloads2y agoHugging Face26nph4rd /eleusis-calibrated-rules Eleusis Calibrated Rules — 100-turn reward calibration A calibrated rule dataset for the single-player Eleusis inductive-reasoning environment. It extends the 26-rule Hugging Face benchmark with controlled static, transition, conditional, periodic, chunk, higher-order history, global history, and compositional rule families. Source benchmark: Hugging Face Eleusis. Dataset version: v2.1-frontier-calibrated-100turn-20260812Protocol: eleusis-100-v11 The structural, GPT Sol… See the full description on the dataset page: https://huggingface.co/datasets/nph4rd/eleusis-calibrated-rules.imagereinforcement-learning1K<n<10K0 likes83 downloads1mo agoHugging Face27npetro6 /nevada-federal-contractors Nevada Federal Contractors - FedComp Index (September 2026) Classified dataset of 779 federal contractors registered in Nevada as of September 2026, covering five years of USASpending base contract award data. Each contractor is assigned a Posture Class (1-4) based on two axes: base contract volume and base contract frequency. This dataset is updated weekly as new award data becomes available. Source All data is sourced from USASpending.gov and SAM.gov.… See the full description on the dataset page: https://huggingface.co/datasets/npetro6/nevada-federal-contractors.tabulartabular-classification1K<n<10K0 likes82 downloads20d agoHugging Face28AdleBens /nptqlzdcsx nptqlzdcsx This dataset was generated using a phospho dev kit. This dataset contains a series of episodes recorded with a robot and multiple cameras. It can be directly used to train a policy using imitation learning. It's compatible with LeRobot and RLDS. tabularrobotics1K<n<10K0 likes70 downloads1y agoHugging Face29npaka /air-hockey-amdThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so101_follower", "total_episodes": 21, "total_frames": 36046, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:21" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/npaka/air-hockey-amd.tabularrobotics10K<n<100K0 likes67 downloads10mo agoHugging Face30nphamdinh /textq-german TextQ-German       TextQ investigates how people perceive the quality of machine-generated German text and how these subjective judgments can be modeled automatically. We identified task-specific quality dimensions, quantified them through user ratings, and developed models that predict perceived quality for new generated texts. TextQ-German is a dataset suite for studying the Quality of Experience (QoE) of machine-generated German text. It covers two Natural Language… See the full description on the dataset page: https://huggingface.co/datasets/nphamdinh/textq-german.tabularn<1K1 likes65 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.