CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ArtificialAnalysis /Earnings22-Cleaned-AA Earnings22-Cleaned-AA Quick links: AA Speech-to-Text Leaderboard | AA-WER v2.0 article Earnings22-Cleaned-AA is a cleaned subset of the English Earnings-22 test data from esb/datasets, a corpus of corporate earnings calls from global companies with speakers of many different nationalities and accents. This cleaned subset is the Earnings-22 portion included in AA-WER v2. We manually reviewed and corrected errors in the original ground-truth transcriptions to ensure fairer evaluation… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/Earnings22-Cleaned-AA.audioautomatic-speech-recognitionn<1K6 likes6.2k downloads7mo agoHugging Face02artificialguybr /veo3-video-prompts Veo 3 Video Generation Dataset English | Português do Brasil English Summary A collection of AI-generated videos created with Google's Veo 3 family of models. Each record contains the original text prompt, the model variant used, the generated video, and (when applicable) the input reference image. Videos are organized into one configuration per model variant. Videos: 5,811 Input images: 1,354 Configurations: 6 Language of prompts: multilingual… See the full description on the dataset page: https://huggingface.co/datasets/artificialguybr/veo3-video-prompts.imagetext-to-video1K<n<10K0 likes5.3k downloads1mo agoHugging Face03soumitsr /article-digeststext10K<n<100K0 likes4.2k downloads2y agoHugging Face04arterm-sedov /agent-course-final-assignment Agent Course Final Assignment - Unified Dataset Author: Arte(r)m Sedov GitHub: https://github.com/arterm-sedov/ Project link: https://huggingface.co/spaces/arterm-sedov/agent-course-final-assignment Dataset Description This dataset is produced by the GAIA Unit 4 Agent for the Hugging Face Agents Course final assignment as part of an experimental multi-LLM agent system that demonstrates advanced AI agent capabilities. It demonstrates advanced AI agent capabilities for… See the full description on the dataset page: https://huggingface.co/datasets/arterm-sedov/agent-course-final-assignment.tabularn<1K1 likes3.1k downloads9mo agoHugging Face05adameubanks /filtered_articles_by_year Dataset Card for Filtered Articles by Year Dataset Summary The Filtered Articles by Year dataset contains yearly-segmented web articles from the FineWeb dataset, specifically filtered and processed for temporal language analysis and Word2Vec model training. This dataset spans 21 years (2005-2025) and serves as the foundation for research into semantic change, concept emergence, and language evolution over time. Supported Tasks and Leaderboards This dataset… See the full description on the dataset page: https://huggingface.co/datasets/adameubanks/filtered_articles_by_year.texttext-generation10M<n<100M1 likes2.6k downloads1y agoHugging Face06ArtificialAnalysis /ITBench-AA ITBench-AA Artificial Analysis' release of the public scenarios from IBM's ITBench benchmark, used for the ITBench-AA leaderboard. This repo currently contains the SRE subset (sre config). Each row is a Kubernetes incident scenario with its expected contributing-factor entities. An agent under evaluation is given access to an offline snapshot of the affected cluster (alerts, events, traces, topology) and must identify the entity (Deployment, Pod, ConfigMap, etc.) responsible for… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/ITBench-AA.textquestion-answeringn<1K47 likes2.6k downloads4mo agoHugging Face07ArtificialAnalysis /AA-Briefcase-Lite AA-Briefcase-Lite The public example scenario for AA-Briefcase, Artificial Analysis' frontier agentic evaluation of realistic, long-horizon knowledge work. Leaderboard and detailed results Launch article AA-Briefcase extends frontier model benchmarking beyond coding and short-form reasoning to the professional deliverables knowledge workers produce day to day. It consists of four private scenarios in which agents complete realistic professional workflows across data science… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/AA-Briefcase-Lite.documentothern<1K11 likes2.1k downloads3mo agoHugging Face08ArtificialAnalysis /VoxPopuli-Cleaned-AA VoxPopuli-Cleaned-AA Quick links: AA Speech to Text Leaderboard | AA-WER v2.0 article VoxPopuli-Cleaned-AA is a cleaned subset of the English VoxPopuli test data from esb/datasets, a speech dataset derived from European Parliament recordings. This cleaned subset is the VoxPopuli portion included in AA-WER v2. We manually reviewed and corrected errors in the original ground-truth transcriptions to ensure fairer evaluation of Speech to Text (STT) models. This dataset is part of AA-WER… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/VoxPopuli-Cleaned-AA.audioautomatic-speech-recognitionn<1K7 likes1.1k downloads7mo agoHugging Face09csoai /gspc-art5 GSPC — art5 safeguard bank (Art5Bench) Council of AI measurement bank. Measurement, not certification. Bank. Frozen split. Live n is the matching axis on GET https://councilof.ai/api/gspc, not a Hub score. Not a certificate. Art 50 (EUR-Lex): 2 August 2026 live; marking grace 2 December 2026. Live measurement. This bank stands behind the art5-safeguard row of the live GSPC board: GET https://councilof.ai/api/gspc?axis=art5-safeguard (family, kind, status and n are on that row… See the full description on the dataset page: https://huggingface.co/datasets/csoai/gspc-art5.tabularquestion-answeringn<1K0 likes873 downloads20m agoHugging Face10Thunderbolt215215 /ArtiMuse-10K ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level Understanding [🌐 Project Page] [🚀 Online Demo] [💻 Code] [📄 Paper] [[🧩 Checkpoints: 🤗 Hugging Face | 🤖 ModelScope]] 🌟 Building upon on ArtiMuse, we introduce UniPercept, a comprehensive follow-up work that provides a meticulous study on perceptual-level image understanding. It spans Image Aesthetics Assessment (IAA), Image Quality Assessment (IQA), and Image Structure & Texture… See the full description on the dataset page: https://huggingface.co/datasets/Thunderbolt215215/ArtiMuse-10K.imagequestion-answering1K<n<10K4 likes757 downloads8mo agoHugging Face11keryszhan /harbor-swesmith-rl-artifacts Harbor SWE-Smith 强化学习数据产物 本数据集是 Harbor Qwen 工具调用代码智能体强化学习项目使用的冻结任务集,服务于 GRPO、原生价值模型/GAE PPO、训练过程诊断和统一协议评测。 项目已于 2026 年 8 月 30 日完成 P0 评测并进入阶段性归档。本数据集用于保留实验所依赖的数据切分、任务执行文件和审计信息,不代表新的通用代码能力基准。 数据概况 切分 任务数 训练集 187 验证集 42 测试集 38 合计 267 数据覆盖 89 个上游代码仓库。三个切分之间同时执行任务标识和仓库级隔离检查。 正式数据集名称: swesmith-curated-grpo-267-v1 冻结切分的语义摘要: ae5df9a3f4a3fc8af44fac420b36529e283839e1bd3de9daba65d5bcda51447d 该值来自 split-manifest.json 的 sha256 字段,用于标识切分语义,不等同于该文件本身的字节级… See the full description on the dataset page: https://huggingface.co/datasets/keryszhan/harbor-swesmith-rl-artifacts.tabulartext-generationn<1K0 likes757 downloads17d agoHugging Face12ArtificialAnalysis /Earnings22-Cleaned-AA-chunked Earnings22-Cleaned-AA-chunked Quick links: AA Streaming Speech to Text Leaderboard | Speech to Text methodology Earnings22-Cleaned-AA-chunked is a chunked version of Earnings22-Cleaned-AA, the cleaned Earnings-22 subset used by Artificial Analysis for streaming Speech to Text evaluation. The original Earnings-22 data comes from esb/datasets, a corpus of corporate earnings calls. Artificial Analysis manually reviewed and corrected the reference transcripts in the cleaned subset… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/Earnings22-Cleaned-AA-chunked.audioautomatic-speech-recognitionn<1K1 likes533 downloads3mo agoHugging Face13changdae /tau2-uq-artifacts tau2-bench UQ Artifacts Interaction trajectories and token-level log-probability measurements from conversational customer service agent evaluations on tau2-bench, collected as part of the uncertainty quantification (UQ) pipeline. Used for analyses in the paper "Uncertainty Quantification in LLM Agents: Foundations, Emerging Challenges, and Opportunities" under the agentuq codebase. Dataset Overview This dataset contains two types of artifacts: Trajectories --… See the full description on the dataset page: https://huggingface.co/datasets/changdae/tau2-uq-artifacts.tabulartext-generation1M<n<10M0 likes521 downloads20d agoHugging Face14AtomicChat /dsv4-eval-artifacts DeepSeek-V4-Flash-0731 — quantization measurements Everything needed to reproduce, audit or extend the numbers published in AtomicChat/DeepSeek-V4-Flash-0731-GGUF: the reference logits, the evaluation corpus, the raw tool output for every quant we measured, and the parsed results. Every GGUF of this model that we could find on the Hub was measured here — ours, unsloth's, bartowski's, ggml-org's, antirez's and others — on one machine, against one reference, with one command.… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/dsv4-eval-artifacts.texttext-generationn<1K0 likes517 downloads2mo agoHugging Face15Nbardy /art-theory-textbookstext10K<n<100K3 likes437 downloads3y agoHugging Face16zhcharyzhang /patchaudit-artifact PatchAudit Artifact PatchAudit audits security patches. You give it a CVE's initial fix — commit C1 — and a later commit Ci, and it tells you whether Ci is a future commit: a later commit that had to keep fixing the same problem because C1 was incomplete (it left the vulnerability reachable) or incorrect (its own change introduced a new defect). When such a future commit exists, C1 was a bad patch. When even the latest fix still leaves the hole open, the bug is a lingering… See the full description on the dataset page: https://huggingface.co/datasets/zhcharyzhang/patchaudit-artifact.text1K<n<10K0 likes423 downloads2d agoHugging Face17KrossKinetic /SP500-Financial-News-Articles-Time-SeriesTextual Time Series Dataset for finetuning / pretraining. Json version of original dataset. Original Dataset : https://www.kaggle.com/datasets/skywalker290/financial-news-article-and-stock-trend-dataset?select=stock_data_articles.csv text1K<n<10K6 likes369 downloads2y agoHugging Face18leonidas1712 /public-agent-coordination-artifacts Public Agent Coordination Artifacts Real, complete edits and posts that AI agents left on public wikis and paste sites — collected as open evidence for studying how autonomous agents use shared online spaces to remember things, signal each other, and coordinate. It's the behavior spotlighted by the mid-2026 OpenAI–Hugging Face agent incident, here as raw public data researchers can actually inspect — plus a small, hand-reviewed map of how specific artifacts relate.… See the full description on the dataset page: https://huggingface.co/datasets/leonidas1712/public-agent-coordination-artifacts.tabular10K<n<100K0 likes318 downloads16d agoHugging Face19eoplumbum /v4_nuclear_power_articles Dataset Card for Nuclear News V4 Dataset Dataset Summary The Nuclear News V4 Dataset is a multilingual dataset consisting of 33,104 unique news articles sourced from 12 online news platforms across the Visegrád Group (V4) countries — Poland, Czech Republic, Slovakia, and Hungary — published between 1998 and 2025. The goal of the dataset is to analyze media narratives surrounding nuclear energy in Central Europe. While the dataset does not contain human-annotated (golden)… See the full description on the dataset page: https://huggingface.co/datasets/eoplumbum/v4_nuclear_power_articles.tabulartext-classification10K<n<100K1 likes235 downloads1y agoHugging Face20lwaekfjlk /artifact-bench ArtifactBench A heterogeneous graph of HuggingFace model / dataset / paper / codebase nodes (14,053) with observed (model, dataset, performance-metric) evaluation edges (51,337 relations), for benchmarking link prediction and attribute (metric-value) regression, plus an agent-based verification suite. License Released under the Open Database License (ODbL) v1.0 — see LICENSE or https://opendatacommons.org/licenses/odbl/1-0/. Share/modify/use freely with… See the full description on the dataset page: https://huggingface.co/datasets/lwaekfjlk/artifact-bench.text10K<n<100K9 likes214 downloads4mo agoHugging Face21BAAI /IndustryInstruction_Artificial-Intelligence IndustryInstruction: Artificial Intelligence This repository contains the IndustryInstruction: Artificial Intelligence domain subset of BAAI/IndustryInstruction. Refer to the parent dataset card for data construction, intended use, limitations, and licensing details. Citation If you use this dataset in your work, please cite IndustryInstruction: @misc{shi2024industryinstruction, title = {IndustryInstruction}, author = {Xiaofeng Shi and Lulu Zhao and Hua… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryInstruction_Artificial-Intelligence.tabularquestion-answering100K<n<1M2 likes209 downloads1mo agoHugging Face22xuzishan /backup-drmas-noshare-8b-historical-run-artifacts Historical DrMAS Noshare 8B run artifacts Public backup of the historical drmas-checklist-bs16-n4-c1-agent Qwen3-8B run. This was a Noshare topology with separately trained Tool Caller and Tool Simulator agents (world_size=8). Contents rollout_dumps/: 470 training rollout JSONL files val_dumps/: 47 validation JSONL files Four training logs latest_checkpointed_iteration.txt 522 business files and 5,189,359,695 logical bytes in total Model scope The… See the full description on the dataset page: https://huggingface.co/datasets/xuzishan/backup-drmas-noshare-8b-historical-run-artifacts.tabular1K<n<10K0 likes199 downloads18d agoHugging Face23arthur801031 /mail Learning Robot Manipulation from Cross-Morphology Demonstration (CoRL 2023) [Project website] [arXiv PDF] Datasets for MAIL. Authors: Gautam Salhotra*, I-Chun Arthur Liu*, Gaurav S. Sukhatme (* denotes equal contribution) Some Learning from Demonstrations (LfD) methods handle small mismatches in the action spaces of the teacher and student. Here we address the case where the teacher’s morphology is substantially different from that of the student. Our framework, Morphological… See the full description on the dataset page: https://huggingface.co/datasets/arthur801031/mail.tabularroboticsn<1K1 likes197 downloads2y agoHugging Face24PPPPPeter /artaimage10K<n<100K0 likes196 downloads1y agoHugging Face25LieUr /crooked-nebula-artifactstabular1K<n<10K0 likes190 downloads5d agoHugging Face26xzAscC /postdyn-artifactstabular100K<n<1M0 likes188 downloads14d agoHugging Face27tencent /ArtifactsBenchmark ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation Tencent Hunyuan Team 📖 Paper • 🏠 Home Page • 💻 Code • 🏆 Leaderboard • 📜 Citation Figure 1: Automation level versus human–alignment across evaluation frameworks. The red star marks the fully manual WebDev Arena (100% human effort), while the blue bubble denotes our checklist-guided MLLM evaluation, ArtifactsBench, which achieves 94.4% agreement with human… See the full description on the dataset page: https://huggingface.co/datasets/tencent/ArtifactsBenchmark.texttext-generation1K<n<10K13 likes174 downloads11mo agoHugging Face28Artanic30 /Wiki_R1_Train Wiki-R1 Training Dataset Dataset Link: https://huggingface.co/datasets/Artanic30/Wiki_R1_Train Overview This dataset contains training annotations and auxiliary files for the Wiki-R1 project, a reinforcement learning framework for Knowledge-Intensive Visual Question Answering (KIVQA). The dataset combines Infoseek and EVQA data with knowledge base retrieval and label propagation support. Dataset Structure The dataset consists of the following JSON… See the full description on the dataset page: https://huggingface.co/datasets/Artanic30/Wiki_R1_Train.text10K<n<100K0 likes156 downloads6mo agoHugging Face29ArtemLykov /LLM-MARS_datasetThis dataset was developed by a team from Skoltech's Intelligent Space Robotics Laboratory. The dataset was used to train LLM for Robot Behavior Tree Generation based on a user command. Note that this model is part of a multi-agent artificial intelligence system for the dog robot described in the LLM-MARS paper. The dataset is split into parts by different game strategies Paper preprint BibTeX cite: @misc{lykov2023llmmarslargelanguagemodel, title={LLM-MARS: Large Language Model for… See the full description on the dataset page: https://huggingface.co/datasets/ArtemLykov/LLM-MARS_dataset.text1K<n<10K1 likes143 downloads2y agoHugging Face30as-benchmark-artifacts /vqa-cmsv-benchmark VQA-CMSV Benchmark Data Package This repository contains annotation splits for VQA v2-CMSV, GQA-CMSV, and VG-CMSV, plus patch-mask NPZ files used for mask supervision experiments. Contents data/vqa_v2_cmsv/train.json, data/vqa_v2_cmsv/val.json, data/vqa_v2_cmsv/test.json data/gqa_cmsv/train.jsonl, data/gqa_cmsv/val.jsonl, data/gqa_cmsv/test.jsonl data/vg_cmsv/train.jsonl, data/vg_cmsv/val.jsonl, data/vg_cmsv/test.jsonl masks/vqa_v2_cmsv_masks.npz masks/gqa_cmsv_masks.npz… See the full description on the dataset page: https://huggingface.co/datasets/as-benchmark-artifacts/vqa-cmsv-benchmark.tabularvisual-question-answering10K<n<100K0 likes130 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.