CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01DannHiroaki /China-Building-Footprints-CMAB-Mirror Origin Data @misc{Zhang2025CMAB, author = {Zhang, Yecheng and Zhao, Huimin and Long, Ying}, title = {{CMAB-The World's First National-Scale Multi-Attribute Building Dataset}}, year = {2025}, month = apr, publisher = {figshare}, doi = {10.6084/m9.figshare.27992417}, url = {https://doi.org/10.6084/m9.figshare.27992417}, howpublished = {dataset} } Paper @article{Zhang2025SciData, author = {Zhang, Y. and… See the full description on the dataset page: https://huggingface.co/datasets/DannHiroaki/China-Building-Footprints-CMAB-Mirror.geospatialn<1K0 likes13k downloads8mo agoHugging Face02Jio7 /danbooru-tags-classified danbooru-tags-classified Danbooru tags split into categories, one CSV per category. Each CSV is tag,count sorted by count descending, where count is the tag's post frequency on Danbooru. file tags contents artist.csv 83355 artist names character.csv 57653 character names series.csv 12405 copyright / series names other.csv 15942 not yet assigned to a category attire.csv 9646 clothing and worn items object.csv 4227 objects feature.csv 2930 body and… See the full description on the dataset page: https://huggingface.co/datasets/Jio7/danbooru-tags-classified.texttext-classification100K<n<1M5 likes678 downloads15d agoHugging Face03Dan-Kos /arxivannotations Title Annotation PDF Latex Axion bremsstrahlung from collisions of global strings We calculate axion radiation emitted in the collision of two straight globalstrings. The strings are supposed to be in the unexcited ground state, to beinclined with respect to each other, and to move in parallel planes. Radiationarises when the point of minimal separation between the strings moves fasterthan light. This effect exhibits a typical Cerenkov nature. Surprisingly, itallows an alternative… See the full description on the dataset page: https://huggingface.co/datasets/Dan-Kos/arxivannotations.textsummarization100K<n<1M2 likes506 downloads3y agoHugging Face04danielharkin21 /futuressstabular100M<n<1B0 likes370 downloads5mo agoHugging Face05Danau5tin /terminal-taskstextn<1K7 likes316 downloads1y agoHugging Face06danielrosehill /GHG-Emissions-Data GHG Emissions Data Pipeline Description This repository contains a comprehensive pipeline for processing and analyzing greenhouse gas (GHG) emissions data. The pipeline integrates datasets from multiple sources, including Climate TRACE and Our World in Data, to provide insights into global emissions trends. It supports sustainability reporting, emissions tracking, and climate action planning. Dataset Details Sources and Methodologies The pipeline… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/GHG-Emissions-Data.tabularn<1K0 likes185 downloads2y agoHugging Face07Bl4ckSpaces /Danbooru-Top1000-Latents-NPZtext1K<n<10K0 likes181 downloads5mo agoHugging Face08Bl4ckSpaces /Danbooru-Top1000-Latents-SDXLtext1K<n<10K0 likes177 downloads5mo agoHugging Face09ame-la /danbooru-tags-data-zh danbooru-tags-data-zh Danbooru 标签的中文翻译数据库,收录截至 2026 年 8 月图片数量大于 50 的标签,译名以社区通用叫法为准。 同一份数据同时发布在 GitHub 仓库 与 HuggingFace 数据集,推送 GitHub 后由 Actions 自动同步。 覆盖范围 收录 2026 年 8 月时图片数量大于 50 的标签 按分类收录:画师(artist)、作品/版权(copyright)、角色(character)、通用(general)、元标签(meta) 当前收录情况: 分类 标签数 画师 artist 24881 作品/版权 copyright 8413 角色 character 35382 通用 general 30664 元标签 meta 585 特点 现有的同类数据多为较早期抓取、之后未再更新,且通常只提供单一译名,不含别名,也没有对标签含义的说明。本项目在以下方面做了补充:… See the full description on the dataset page: https://huggingface.co/datasets/ame-la/danbooru-tags-data-zh.tabular10K<n<100K0 likes173 downloads5d agoHugging Face10StoryAura /Danbooru-Dataset-csv Danbooru Dataset CSV 面向 Danbooru 标签管理 / 打标工具的公开元数据合集。这里只放整理后的 CSV,不含任何图片。后续还会继续补充 artist、copyright 等更多表;本页只做项目总览,各文件以仓库里的 CSV 为准。 标签与 wiki 来自 Danbooru。本仓库整理表使用 MIT 协议。原图版权仍归各自作者。 当前文件 文件 内容 截止日期 行数 danbooru_dataset_general_260820.csv general 通用标签(别名、层级、父子、分类、wiki) 2026-08-20 106,414 danbooru_character_tags.csv character 角色标签(别名、作品、父标签、投稿数) 2026-07-20 329,747 danbooru_artist_tags.csv artist 画师标签(译名、数据量) — 576,842 tag-near-synonym-relations4.csv… See the full description on the dataset page: https://huggingface.co/datasets/StoryAura/Danbooru-Dataset-csv.tabulartext-classification100K<n<1M1 likes141 downloads7d agoHugging Face11danielrosehill /storage-container-dimensions Industrial Storage Container Dimensions Planning-grade volumetric reference for the plastic containers that European and North American warehouses actually run on: Euroboxes (Euro stacking containers), attached-lid containers (ALCs), and VDA 4500 KLTs. Two tables: Config Rows What it is default → containers.csv 48 One typical row per nominal size. External and internal dimensions, usable capacity, the spread of what vendors publish, lid and nesting behaviour… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/storage-container-dimensions.documentn<1K0 likes133 downloads2mo agoHugging Face12danliu1226 /cross_species_benchmarking**Repository: https://d-script.readthedocs.io/en/stable/data.html **Reference: Sledzieski, S., Singh, R., Cowen, L. & Berger, B. D-SCRIPT translates genome to phenome with sequence-based, structure-aware, genome-scale predictions of protein-protein interactions. Cell Systems 12, 969-982.e6 (2021). text100K<n<1M2 likes126 downloads1y agoHugging Face13danilocorsi /LLMs-Sentiment-Augmented-Bitcoin-Dataset Leveraging LLMs for Informed Bitcoin Trading Decisions: Prompting with Social and News Data Reveals Promising Predictive Abilities The work was carried out by: Danilo Corsi Cesare Campagnano Description This project investigates the potential of leveraging Large Language Models (LLMs) to support Bitcoin traders. Specifically, we analyze the correlation between Bitcoin price movements and sentiment expressed in news headlines, posts, and comments on social media. We… See the full description on the dataset page: https://huggingface.co/datasets/danilocorsi/LLMs-Sentiment-Augmented-Bitcoin-Dataset.tabulartext-classification10K<n<100K8 likes111 downloads2y agoHugging Face14daniilakk /cheboksary_transport ChebTransport Daily GPS Dataset This dataset contains daily GPS and metadata records of public transport vehicles in Cheboksary, Russia, for the period from 2025-04-22 to 2025-07-03. Each file corresponds to a "transport day" (which may start and end at different times depending on the actual end of public transport service, not at midnight). Data Source The data was parsed from the website buscheb.ru, which aggregates public transport data for the city of Cheboksary as a… See the full description on the dataset page: https://huggingface.co/datasets/daniilakk/cheboksary_transport.tabular10M<n<100M1 likes79 downloads1y agoHugging Face15danielelvs /multilingual-islr-mediapipe Multilingual ISLR MediaPipe Landmarks Dataset Description This dataset combines frame-level MediaPipe Holistic landmarks derived from four isolated sign language recognition (ISLR) resources: INCLUDE-50, KSL, MINDS-Libras, and LIBRAS-UFOP. It provides a common tabular schema for research on landmark selection, temporal modeling, signer-independent evaluation, and multilingual transfer learning. The release contains landmarks rather than source RGB videos. Every… See the full description on the dataset page: https://huggingface.co/datasets/danielelvs/multilingual-islr-mediapipe.tabularvideo-classification1K<n<10K0 likes71 downloads6d agoHugging Face16dangvantuan /IEEE-118_overloadTODO tabular100K<n<1M1 likes66 downloads2y agoHugging Face17LEEEFA /danbooru-tags-classified danbooru-tags-classified Danbooru tags split into categories, one CSV per category. Each CSV is tag,count sorted by count descending, where count is the tag's post frequency on Danbooru. file tags contents artist.csv 83355 artist names character.csv 57653 character names series.csv 12405 copyright / series names other.csv 15942 not yet assigned to a category attire.csv 9646 clothing and worn items object.csv 4227 objects feature.csv 2930 body and… See the full description on the dataset page: https://huggingface.co/datasets/LEEEFA/danbooru-tags-classified.texttext-classification100K<n<1M0 likes64 downloads15d agoHugging Face18danliu1226 /Bernett_benchmarking Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Dataset Sources [optional] **Repository: https://doi.org/10.6084/m9.figshare.21591618.v3 **Reference: Bernett, J., Blumenthal, D. B. & List, M. Cracking the black box of deep sequence-based protein–protein interaction prediction. Briefings in Bioinformatics 25, bbae076… See the full description on the dataset page: https://huggingface.co/datasets/danliu1226/Bernett_benchmarking.text100K<n<1M0 likes63 downloads1y agoHugging Face19dant555 /flipfinder-usa 🎥 Project Walkthrough Video 🏠 FlipFinder USA Identifying Undervalued Real Estate Investment Opportunities Across the United States Author: Dan | HuggingFace: @dant555 📋 Project Overview This project transforms a general-purpose Zillow real estate dataset into a focused investment screening tool. Using Exploratory Data Analysis (EDA), I engineered a binary target variable (is_good_flip) to identify properties that are genuinely… See the full description on the dataset page: https://huggingface.co/datasets/dant555/flipfinder-usa.image10K<n<100K1 likes53 downloads6mo agoHugging Face20Dannyang /Nanobody_Sequence_DatasetRepresentative sequence dataset extracted after clustering of Integrated NANOBODY® Database for Immunoinformatics (INDI) by MMseqs2 program. For more information, please visit https://github.com/DynaX-C/EvoNB. text1M<n<10M1 likes51 downloads1y agoHugging Face21danliu1226 /Mutation_effect_dataset**Repository: https://ftp.ebi.ac.uk/pub/databases/intact/current/various/mutations.tsv **Reference: Kerrien, S. et al. The IntAct molecular interaction database in 2012. Nucleic Acids Research 40, D841–D846 (2012). text1K<n<10K0 likes49 downloads1y agoHugging Face22Danielbrdz /Barcenas-HumorNegroDataset en español con 500 chistes de humor negro y una explicación. Datos creados de manera sintética por Claude 3 Haiku y Llama 3 70B Instruct. El proceso para crear el dataset fue el recopilar de varias fuentes chistes de humor negro en español para luego ser utilizadas en los mejores modelos como Gemini 1.5 Pro, Claude 3, etc. Con eso genere cientos de chistes de humor negro en español para tener más datos y hacer un super recopilatorio de chistes de humor negro en español, aproximadamente… See the full description on the dataset page: https://huggingface.co/datasets/Danielbrdz/Barcenas-HumorNegro.texttext-classificationn<1K1 likes47 downloads2y agoHugging Face23danf0 /elizaThis repository contains synthetic ELIZA chatbot conversations. See https://github.com/princeton-nlp/ELIZA-Transformer for more details. tabular100K<n<1M0 likes46 downloads2y agoHugging Face24danliu1226 /STRING_V12_TrainingSet**Repository: https://stringdb-downloads.org/download/protein.physical.links.v12.0.txt.gz **Reference: Szklarczyk, D. et al. The STRING database in 2023: protein–protein association networks and functional enrichment analyses for any sequenced genome of interest. Nucleic Acids Research 51, D638–D646 (2023). text100K<n<1M0 likes44 downloads1y agoHugging Face25nhatnguyet /ngay-ky-dan-gian-2026-2035 Ngày kỵ dân gian quy về ngày dương, 2026 tới 2035 Folk inauspicious days mapped onto solar dates, 2026 to 2035 1. Mô tả · Description Mỗi ngày dương lịch trong mười năm, cho biết ngày ấy có rơi vào Tam Nương, Nguyệt Kỵ, Sát Chủ hay Thọ Tử không, kèm ngày âm và can chi để đối chiếu. Every solar day across ten years, showing whether it falls on Tam Nuong, Nguyet Ky, Sat Chu or Tho Tu, with the lunar date and stem-branch for checking. Số dòng · Rows: 3,652 Phiên bản… See the full description on the dataset page: https://huggingface.co/datasets/nhatnguyet/ngay-ky-dan-gian-2026-2035.tabular1K<n<10K1 likes43 downloads3d agoHugging Face26Shio-Koube /150k-anime-danbooru-fluxtext10K<n<100K0 likes42 downloads1y agoHugging Face27Danie1Arias /sentiment-analysis-catalan-reviews CSXSC: Classificador de Sentiments de Xarxes Socials en Català This repository contains the CSXSC (Classificador de Sentiments a Xarxes Socials en Català) dataset, a comprehensive corpus designed for sentiment analysis of Catalan-language content from social media. The dataset contains 23,788 text entries, each classified as positive, negative, or neutral. It was specifically constructed to address the significant class imbalance often found in user-generated content, resulting in a… See the full description on the dataset page: https://huggingface.co/datasets/Danie1Arias/sentiment-analysis-catalan-reviews.text10K<n<100K0 likes42 downloads1y agoHugging Face28daniilakk /Russia_Real_Estate_2018_2021 Context The dataset consists of lists of unique objects of popular portals for the sale of real estate in Russia. More than 540 thousand objects. The dataset contains 540000 real estate objects in Russia. Content The Russian real estate market has a relatively short history. In the Soviet era, all properties were state-owned; people only had the right to use them with apartments allocated based on one's place of work. As a result, options for moving were fairly limited.… See the full description on the dataset page: https://huggingface.co/datasets/daniilakk/Russia_Real_Estate_2018_2021.tabular1M<n<10M3 likes40 downloads4y agoHugging Face29dangvantuan /IEEE_14_datasettext10K<n<100K0 likes39 downloads2y agoHugging Face30kyosukeyamamoto63 /dangerous-relation-685db9 dangerous-relation-685db9 Synthetic products test data: 39 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at… See the full description on the dataset page: https://huggingface.co/datasets/kyosukeyamamoto63/dangerous-relation-685db9.tabularn<1K0 likes39 downloads13d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.