CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01banned-historical-archives /banned-historical-archives 和谐历史档案馆数据集 - Banned Historical Archives Datasets 和谐历史档案馆数据集包含已录入 https://banned-historical-archives.github.io 和暂未未录入的原始文件。 目录结构 banned-historical-archives.github.io # 已录入该网站的原始数据,不定期从 github 仓库中同步 raw # 原始文件 config # 配置文件 todo # 存放暂未录入网站的文件 部分报纸和图片资料存放在单独的仓库: 名称 地址 状态 参考消息 https://huggingface.co/datasets/banned-historical-archives/ckxx 未录入 人民日报 https://huggingface.co/datasets/banned-historical-archives/rmrb 已精选重要的文章录入 文汇报… See the full description on the dataset page: https://huggingface.co/datasets/banned-historical-archives/banned-historical-archives.imagen<1K92 likes1.7m downloads11mo agoHugging Face02picbreeder-vlm /picbreeder-vlm-archive Picbreeder-VLM Archive Every image evolved by the swarm of vision-language-model "breeders" in In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models (GECCO 2026), together with the CPPN genomes that produced them, the agents' reasoning transcripts, the lineage graphs, and the analysis artifacts behind the paper and blog. The original Picbreeder (Secretan et al., 2008) let crowds of humans collaboratively evolve images from CPPN… See the full description on the dataset page: https://huggingface.co/datasets/picbreeder-vlm/picbreeder-vlm-archive.imageimage-to-text100K<n<1M14 likes68k downloads3mo agoHugging Face03Dragonegg2026 /banned-historical-archives 和谐历史档案馆数据集 - Banned Historical Archives Datasets 和谐历史档案馆数据集包含已录入 https://banned-historical-archives.github.io 和暂未未录入的原始文件。 目录结构 banned-historical-archives.github.io # 已录入该网站的原始数据,不定期从 github 仓库中同步 raw # 原始文件 config # 配置文件 todo # 存放暂未录入网站的文件 部分报纸和图片资料存放在单独的仓库: 名称 地址 状态 参考消息 https://huggingface.co/datasets/banned-historical-archives/ckxx 未录入 人民日报 https://huggingface.co/datasets/banned-historical-archives/rmrb 已精选重要的文章录入 文汇报… See the full description on the dataset page: https://huggingface.co/datasets/Dragonegg2026/banned-historical-archives.imagen<1K1 likes12k downloads6mo agoHugging Face04OmniAICreator /ASMR-Archive-Processed ASMR-Archive-Processed (WIP) Update (2026-04-03): This dataset has reached the Hugging Face Public Storage Limit. After contacting support, we were informed that the only option is to pay for a storage expansion. Consequently, updates to this dataset are now suspended. Work in Progress — expect breaking changes while the pipeline and data layout stabilize. This dataset contains ASMR audio data sourced from DeliberatorArchiver/asmr-archive-data-01 and… See the full description on the dataset page: https://huggingface.co/datasets/OmniAICreator/ASMR-Archive-Processed.imageautomatic-speech-recognition97 likes8.7k downloads6mo agoHugging Face05diwash-barla /meta-archiveimage1K<n<10K2 likes2.6k downloads9d agoHugging Face06eastbrush /eastbrush_archive Eastbrush Archive Official Website (Full Archive System): https://www.eastbrush.com This dataset contains high-resolution images and structured tags for AI training. The full archive system — including chapter exhibitions, structural context, and extended records — is available on the official website. What is my true self? The Eastbrush Archive is a long-term, evolving system that documents the visual language ofJang Byeong Eun (Eastbrush / 張炳彥) — a painter whose… See the full description on the dataset page: https://huggingface.co/datasets/eastbrush/eastbrush_archive.imagen<1K2 likes2.4k downloads5d agoHugging Face07Augmentiv /ArchiveDatadocumentn<1K0 likes1.9k downloads6mo agoHugging Face08pnsk-lab /scratch-archiveimage3 likes1.9k downloads7mo agoHugging Face09Vasy7777 /cs2-demo-archive CS2 Demo to Dataset — Pipeline Output Samples Sample archives produced by the open-source cs2-demo-to-dataset pipeline, which converts a single CS2 .dem replay file into per-round, per-player first-person video aligned to tick-level state, input and event tables. This release is not a dataset contribution. The point of the upload is to demonstrate that the pipeline produces a coherent, reproducible archive format. Please see the GitHub repository for the recorder code… See the full description on the dataset page: https://huggingface.co/datasets/Vasy7777/cs2-demo-archive.imageother1K<n<10K1 likes1.6k downloads3mo agoHugging Face10aditya487 /cbi-archive-raw Central Bank of Ireland Archive: original source files 6,309 original files, 6.56 GB. Every PDF, spreadsheet, Word document and archive gathered from the Central Bank of Ireland's public website, stored by content hash so that a search result can be turned back into the document a human would actually read. This is the raw tier. If you want the text, you almost certainly want aditya487/cbi-archive-corpus instead: 5,568 documents and 89,242 page or pseudo-page rows as Parquet… See the full description on the dataset page: https://huggingface.co/datasets/aditya487/cbi-archive-raw.document1K<n<10K0 likes1.3k downloads23d agoHugging Face11banned-historical-archives /hkgongshangwanbaoimage1K<n<10K0 likes1.1k downloads2y agoHugging Face12hmar-heritage-org /corpus-archivegated corpus-archive [!WARNING] Experimental Dataset Architecture: The repository structure, metadata tiers, category taxonomies, and catalog indexing formats are currently under active design and evaluation. All specifications, metadata keys, and JSON schemas detailed below represent representational examples and intended targets. This repository serves as a structured digital textual archive preserving Hmar literature, historical accounts, school textbooks, dictionaries, parallel… See the full description on the dataset page: https://huggingface.co/datasets/hmar-heritage-org/corpus-archive.imagetext-classificationn<1K4 likes1.1k downloads9d agoHugging Face13banned-historical-archives /dagongbaoimage1K<n<10K0 likes1k downloads2y agoHugging Face14banned-historical-archives /hkgongshangribaoimage1K<n<10K0 likes1k downloads2y agoHugging Face15banned-historical-archives /huaqiaoribaoimage1K<n<10K0 likes643 downloads2y agoHugging Face16zenless-archive /danbooru2023 [Mirror]Danbooru2023: A Large-Scale Crowdsourced and Tagged Anime Illustration Dataset Danbooru2023 is an extension of Danbooru2021, featuring over 6.8 million anime-style images, totaling more than 8.3 TB. Each image is accompanied by community-contributed tags that provide detailed descriptions of its content, including characters, artists, copyright information, concepts, and attire. This makes it a crucial resource for stylized computer vision tasks and transfer learning.… See the full description on the dataset page: https://huggingface.co/datasets/zenless-archive/danbooru2023.imageimage-to-text1M<n<10M0 likes630 downloads1y agoHugging Face17maximilian-franz /basil-instances-archive-3 maximilian-franz/basil-instances Per-plant-instance segmented crops derived from ['maximilian-franz/basil', 'maximilian-franz/basil-2'], one row per (original frame × confirmed plant instance). Layout ImageFolder-style dataset: metadata.csv at the repo root, with a file_name column pointing to each row's masked crop under images/<plant_instance_id>/<NNNN>.png, and a bbox_file_name column pointing to the same row's plain rectangular crop under… See the full description on the dataset page: https://huggingface.co/datasets/maximilian-franz/basil-instances-archive-3.image1K<n<10K0 likes460 downloads2mo agoHugging Face18maximilian-franz /basil-instances-archive-2 maximilian-franz/basil-instances Per-plant-instance segmented crops derived from ['maximilian-franz/basil', 'maximilian-franz/basil-2'], one row per (original frame × confirmed plant instance). Layout ImageFolder-style dataset: metadata.csv at the repo root, with a file_name column pointing to each row's image under images/<plant_instance_id>/<NNNN>.png. Load with: from datasets import load_dataset ds = load_dataset("maximilian-franz/basil-instances")… See the full description on the dataset page: https://huggingface.co/datasets/maximilian-franz/basil-instances-archive-2.image1K<n<10K0 likes437 downloads2mo agoHugging Face19NullVoider /Memory-Archive-Paradigm Memory Archive: A Memory-Grounded Training Paradigm for Computer Use Agents Kartik A. · Independent Researcher · Project Dockyard 📄 Read the Paper (PDF) &nbsp;·&nbsp; 🗂️ The Corpus (Hugging Face) &nbsp;·&nbsp; 💻 Memory Archive Tool Publication note: As an independent researcher, this architecture is published as an open-science preprint via Zenodo (CERN) to establish formal prior art, with a permanent, globally recognised DOI. This is the design specification and… See the full description on the dataset page: https://huggingface.co/datasets/NullVoider/Memory-Archive-Paradigm.image1K<n<10K0 likes411 downloads1mo agoHugging Face20aarhus-city-archives /historical-danish-handwriting Dataset Description Dataset Summary The Historical Danish handwriting dataset is a Danish-language dataset containing more than 11.000 pages of transcribed and proofread handwritten text. The dataset currently consists of the published minutes from a number of City and Parish Council meetings, all dated between 1841 and 1939. Languages All the text is in Danish. The BCP-47 code for Danish is da. Dataset Structure Data Instances Each data… See the full description on the dataset page: https://huggingface.co/datasets/aarhus-city-archives/historical-danish-handwriting.imageimage-to-text1K<n<10K4 likes407 downloads1y agoHugging Face21banned-historical-archives /wenhuibao 文汇报(扫描图像,不完整) 作为文汇报光盘的补充材料 https://huggingface.co/datasets/banned-historical-archives/wenhuibao_disk imagen<1K0 likes386 downloads2y agoHugging Face22TobanDjan /isic_archiveimage10K<n<100K0 likes379 downloads2y agoHugging Face23Arthur12137 /SoftVTBench-archive SoftVTBench — archive Frozen snapshot of everything that lived in Arthur12137/SoftVTBench before the 2026-08-13 re-release: the evaluation USD assets, the soft-body assets, and the first partial data drops. This repo is not maintained. The current dataset is at Arthur12137/SoftVTBench. Contents: eval-assets/, soft-assets/, object-rigid/, object-soft/, spatial-rigid/, spatial-soft/ (9319 files, 2.3 GB). Original dataset card (kept verbatim) SoftVTBench dataset… See the full description on the dataset page: https://huggingface.co/datasets/Arthur12137/SoftVTBench-archive.imageroboticsn<1K0 likes366 downloads1mo agoHugging Face24pwc-archive /datasets [!CAUTION] This dataset will not be updated. It corresponds to the last available public snapshot of the data, retrieved on July 28th, 2025. image10K<n<100K5 likes365 downloads1y agoHugging Face25aaaad1 /banned-historical-archives 和谐历史档案馆数据集 - Banned Historical Archives Datasets 和谐历史档案馆数据集包含已录入 https://banned-historical-archives.github.io 和暂未未录入的原始文件。 目录结构 banned-historical-archives.github.io # 已录入该网站的原始数据,不定期从 github 仓库中同步 raw # 原始文件 config # 配置文件 todo # 存放暂未录入网站的文件 部分报纸和图片资料存放在单独的仓库: 名称 地址 状态 参考消息 https://huggingface.co/datasets/banned-historical-archives/ckxx 未录入 人民日报 https://huggingface.co/datasets/banned-historical-archives/rmrb 已精选重要的文章录入 文汇报… See the full description on the dataset page: https://huggingface.co/datasets/aaaad1/banned-historical-archives.imagen<1K0 likes353 downloads5mo agoHugging Face26taechasith /kala-uap-archivedocumentn<1K0 likes340 downloads4mo agoHugging Face27previtus /jpl_trace_gases_archivegeospatialn<1K0 likes267 downloads8d agoHugging Face28karo2w /archiveimagen<1K0 likes262 downloads1y agoHugging Face29Hutao0514 /schale-archive-mirror Schale Archive Overview This repository contains Blue Archive game assets. Known Missing Characters Spine NPC GSC President Yume Nao Nagusa Nyanten-maru Pei Reizyo A.R.O.N.A How to Properly Import Character to Spine Editor First, you need Spine Editor at least v3.8.x and later. Create a new project. Delete the default skeleton in Hierarchy on the right. Click Spine logo on top left, then select Import Data. Select… See the full description on the dataset page: https://huggingface.co/datasets/Hutao0514/schale-archive-mirror.audio1K<n<10K0 likes223 downloads2mo agoHugging Face30QTE-Technologies /industrial-technical-archive 🚀 Latest Updates (July, 2026) Version: v07.2026 (Verified) Status: Integrated with 1,000,000+ records. New Files: product-E-26-07-2026.csv & product-V-26-07-2026.csv. QTE Technologies: Industrial & Scientific Knowledge Base Wikidata Entity: Q138411149 IPFS CID: bafybeibogxxuhmzfrsuhcfd4qr4tmc4okhmrcwhp3266hq47ccuyjnjxoq Official Neural Hub: qtetech.github.io This is the permanent technical archive for QTE Technologies, ensuring long-term accessibility of… See the full description on the dataset page: https://huggingface.co/datasets/QTE-Technologies/industrial-technical-archive.image10K<n<100K0 likes212 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.