CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01wangyz1999 /X-EGO-CS X-Ego-CS Ten players. One match. Ten simultaneous first-person recordings, each paired with a 64 Hz stream of that player's exact keyboard, mouse and view-angle inputs — all on a common, measured clock. Paper · Paper code · Collection pipeline Cross-Ego Demo (Pistol Round) Your browser cannot play this video — download it instead. All ten players' points of view, from the same pistol round, on one clock. Note: this demo concatenates the ten streams… See the full description on the dataset page: https://huggingface.co/datasets/wangyz1999/X-EGO-CS.tabularvideo-classification10K<n<100K2 likes30k downloads4d agoHugging Face02juletxara /xstory_cloze Dataset Card for XStoryCloze Dataset Summary XStoryCloze consists of the professionally translated version of the English StoryCloze dataset (Spring 2016 version) to 10 non-English languages. This dataset is released by Meta AI. Supported Tasks and Leaderboards commonsense reasoning Languages en, ru, zh (Simplified), es (Latin America), ar, hi, id, te, sw, eu, my. Dataset Structure Data Instances Size of downloaded dataset… See the full description on the dataset page: https://huggingface.co/datasets/juletxara/xstory_cloze.textother10K<n<100K16 likes11k downloads1y agoHugging Face03zackyabd /ptb-xl-processedtabular10K<n<100K0 likes8.9k downloads1y agoHugging Face04jingwei-xu-00 /eccv2026-cad-challenge-data ECCV 2026 CAD Challenge Data This challenge is part of the workshop The Path to Manufacturing: Evolving 3D Generation to Intelligent Computer-Aided Design. Workshop homepage: https://3dgen-cad-workshop.github.io/ Challenge submission Space: https://huggingface.co/spaces/jingwei-xu-00/eccv2026-cad-challenge Dataset rendering and preparation code (only .step files are required): https://github.com/DavidXu-JJ/eccv2026-cad-challenge-data-render This repository contains the public… See the full description on the dataset page: https://huggingface.co/datasets/jingwei-xu-00/eccv2026-cad-challenge-data.3dimage-to-3d1K<n<10K5 likes6.4k downloads2mo agoHugging Face05xuejun72 /HR-VILAGE-3K3M HR-VILAGE-3K3M: Human Respiratory Viral Immunization Longitudinal Gene Expression This repository provides the HR-VILAGE-3K3M dataset, a curated collection of human longitudinal gene expression profiles, antibody measurements, and aligned metadata from respiratory viral immunization and infection studies. The dataset includes baseline transcriptomic profiles and covers diverse exposure types (vaccination, inoculation, and mixed exposure). HR-VILAGE-3K3M is designed as a… See the full description on the dataset page: https://huggingface.co/datasets/xuejun72/HR-VILAGE-3K3M.tabularzero-shot-classification1K<n<10K3 likes5.3k downloads27d agoHugging Face06Paul /XSTest XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models XSTest is a test suite designed to identify exaggerated safety / false refusal in Large Language Models (LLMs). It comprises 250 safe prompts across 10 different prompt types, along with 200 unsafe prompts as contrasts. The test suite aims to evaluate how well LLMs balance being helpful with being harmless by testing if they unnecessarily refuse to answer safe prompts that superficially… See the full description on the dataset page: https://huggingface.co/datasets/Paul/XSTest.texttext-generationn<1K5 likes4.6k downloads2y agoHugging Face07Linzhan /Objaverse-XL-Rigged-Animated Objaverse-XL Rigged & Animated Subset Every asset here carries both a skeleton and at least one animation clip, selected from Objaverse / Objaverse-XL. Rigs range from 3 to 344 joints and span characters as well as articulated rigid objects. Objaverse-XL indexes over 10 million objects, but only a small fraction carry a usable rig and motion on it. This subset isolates that fraction: every file was checked to contain at least one skin with joints and at least one animation clip… See the full description on the dataset page: https://huggingface.co/datasets/Linzhan/Objaverse-XL-Rigged-Animated.3dtext-to-3d10K<n<100K1 likes2.5k downloads19d agoHugging Face08xhluca /publichealth-qa Usage import datasets langs = ['arabic', 'chinese', 'english', 'french', 'korean', 'russian', 'spanish', 'vietnamese'] data = datasets.load_dataset('xhluca/publichealth-qa', split='test', name=langs[0]) About This dataset contains question and answer pairs sourced from Q&A pages and FAQs from CDC and WHO pertaining to COVID-19. They were produced and collected between 2019-12 and 2020-04. They were originally published as an aggregated Kaggle dataset.… See the full description on the dataset page: https://huggingface.co/datasets/xhluca/publichealth-qa.textquestion-answeringn<1K1 likes1.6k downloads2y agoHugging Face09reczoo /Frappe_x1 Frappe_x1 Dataset description: The Frappe dataset contains a context-aware app usage log, which comprises 96203 entries by 957 users for 4082 apps used in various contexts. It has 10 feature fields including user_id, item_id, daytime, weekday, isweekend, homework, cost, weather, country, city. The target value indicates whether the user has used the app under the context. Following the AFN work, we randomly split the data into 7:2:1 as the training set, validation set, and test set… See the full description on the dataset page: https://huggingface.co/datasets/reczoo/Frappe_x1.tabular100K<n<1M1 likes1.4k downloads3y agoHugging Face10Kevin-Pal /CUHK-X_Small_Model_Trackgated CUHK-X — Small Model Track Multimodal human action recognition (classification). Given a multimodal clip, predict its action class (action_id, 0–39, 40 classes). Repository layout . ├── Training/ │ ├── class_mapping.csv # action_id <-> action_name (40 classes) │ └── data/ │ └── HAR.z01 … HAR.z08 + HAR.zip # multi-volume zip │ → HAR/data/<modality>/<action>/<user>/<trial>/<files> └── Testing/ ├── data/ │ └──… See the full description on the dataset page: https://huggingface.co/datasets/Kevin-Pal/CUHK-X_Small_Model_Track.textvideo-classificationn<1K7 likes1.3k downloads3mo agoHugging Face11Xiaolong-Han /w2t-llm-arc-easy-lora W2T Llm Arc Easy Lora This repository contains artifacts for the W2T paper: Paper: W2T: LoRA Weights Already Know What They Can Do Repo: Weight2Token Summary ARC-Easy LoRA checkpoints and prepared metadata used for performance prediction. Source Status Storage location: local Verification status: confirmed Files See manifest.json for the exact local or remote source paths used to prepare this release. Citation… See the full description on the dataset page: https://huggingface.co/datasets/Xiaolong-Han/w2t-llm-arc-easy-lora.tabular10K<n<100K0 likes1.1k downloads4mo agoHugging Face12xupy21 /ICPC_Data ICPC World Finals — a discriminative subset, with model traces 24 ICPC World Finals problems (2021–2025), together with the full transcripts of an LLM attempting each of them three times under simulated contest rules. Selection The model Every run in this dataset comes from: nvidia/Nemotron-Cascade-2-30B-A3B The partitions Every one of the 53 problems was run 3 times (seeds 1, 2, 3). Each problem was then placed by its pass rate and… See the full description on the dataset page: https://huggingface.co/datasets/xupy21/ICPC_Data.imagetext-generationn<1K0 likes1.1k downloads3d agoHugging Face13xX-its-amit-Xx /pxr-structure-pose-pool PXR Structure Challenge — Full Multi-Model Pose Pool (184 ligands) Every protein–ligand pose generated during the OpenADMET PXR (pregnane X receptor / NR1I2) structure-prediction challenge, released openly with per-pose labels so the community can reuse the compute already spent — and, we hope, crack the problem this data makes visible. What's here poses/<model>/<SID>.pdb — one best pose per (model, ligand). Protein chain A + ligand (resname LIG). 15 models, up… See the full description on the dataset page: https://huggingface.co/datasets/xX-its-amit-Xx/pxr-structure-pose-pool.tabular1K<n<10K0 likes1.1k downloads3mo agoHugging Face14xingyusu /DNA_Gen Citation Please cite our work using the bibtex below: BibTeX: @article{su2025language, title={Language Models for Controllable DNA Sequence Design}, author={Su, Xingyu and Li, Xiner and Lin, Yuchao and Xie, Ziqian and Zhi, Degui and Ji, Shuiwang}, journal={arXiv preprint arXiv:2507.19523}, year={2025} } document10K<n<100K3 likes812 downloads1y agoHugging Face15xbench /DeepSearch xbench-evals 🌐 Website | 📄 Paper | 🤗 Dataset Evergreen, contamination-free, real-world, domain-specific AI evaluation framework xbench is more than just a scoreboard — it's a new evaluation framework with two complementary tracks, designed to measure both the intelligence frontier and real-world utility of AI systems: AGI Tracking: Measures core model capabilities like reasoning, tool-use, and memory Profession Aligned: A new class of evals grounded in workflows, environments… See the full description on the dataset page: https://huggingface.co/datasets/xbench/DeepSearch.textn<1K13 likes701 downloads1y agoHugging Face16Pcitycrypto /xauusd Cleaned XAUUSD Dataset Dataset Description This dataset contains cleaned and preprocessed minute-level historical price data for the XAU/USD (Gold vs. US Dollar) pair. The data spans from November 1, 2011, to January 3, 2024, and includes the following columns: open: The opening price of the minute. high: The highest price during the minute. low: The lowest price during the minute. close: The closing price of the minute. tickvol: The number of price changes (ticks)… See the full description on the dataset page: https://huggingface.co/datasets/Pcitycrypto/xauusd.tabulartime-series-forecasting1M<n<10M3 likes670 downloads2y agoHugging Face17Skyrmion /DAAD-XAbout Dataset The DAADX Dataset is derived from DAAD dataset (https://cvit.iiit.ac.in/research/projects/cvit-projects/daad#dataset), which contains all the captured videos for the Driver Intention Prediction task. We are introducing the first video based explanations dataset for driver intention prediction task. This will be help in further the research interms of making an explainable Autonomous Driving or ADAS System. DAAD-X contains explanations for each maneuver instance, these… See the full description on the dataset page: https://huggingface.co/datasets/Skyrmion/DAAD-X.tabularvideo-classification1K<n<10K1 likes641 downloads6mo agoHugging Face18xanderios /linkedin-job-postingstabular10K<n<100K14 likes568 downloads3y agoHugging Face19xjtupanda /VStar_Bench V* Benchmark refactor V* Benchmark to add support for VLMEvalKit. Benchmark Information Number of questions: 191 Question type: MCQ (multiple choice question) Question format: image + text Reference VLMEvalKit V* Benchmark MMVP textmultiple-choicen<1K0 likes522 downloads1y agoHugging Face20BramVanroy /xlwic_wn Multilingual Word-in-Context (WordNet) Refer to the documentation and paper for more information. tabulartext-classification10K<n<100K1 likes516 downloads3y agoHugging Face21xbench /DeepSearch-2510 xbench-evals 🌐 Website | 📄 Paper | 🤗 Dataset Evergreen, contamination-free, real-world, domain-specific AI evaluation framework xbench is more than just a scoreboard — it's a new evaluation framework with two complementary tracks, designed to measure both the intelligence frontier and real-world utility of AI systems: AGI Tracking: Measures core model capabilities like reasoning, tool-use, and memory Profession Aligned: A new class of evals grounded in workflows, environments… See the full description on the dataset page: https://huggingface.co/datasets/xbench/DeepSearch-2510.textn<1K4 likes499 downloads11mo agoHugging Face22nimapro1381 /x402-market-pulse x402 Market Pulse A small, growing time series of how much USDC actually settles through x402 pay-per-call APIs on Base, measured from on-chain Transfer logs rather than from provider-reported call counts. One row per measurement; every row is a trailing 24-hour window ending at measured_at_utc. Status: bootstrapping. 74 rows spanning 2026-09-22T13:34Z to 2026-09-25T18:00Z — less than one independent day of data so far. Rows are kept to at most one per clock hour and a new one… See the full description on the dataset page: https://huggingface.co/datasets/nimapro1381/x402-market-pulse.tabulartime-series-forecastingn<1K0 likes483 downloads11h agoHugging Face23reczoo /Criteo_x1 Criteo_x1 Dataset description: The Criteo dataset is a widely-used benchmark dataset for CTR prediction, which contains about one week of click-through data for display advertising. It has 13 numerical feature fields and 26 categorical feature fields. Following the AFN work, we randomly split the data into 7:2:1* as the training set, validation set, and test set, respectively. The dataset statistics are summarized as follows: Dataset Split Total #Train #Validation #Test… See the full description on the dataset page: https://huggingface.co/datasets/reczoo/Criteo_x1.tabular10M<n<100M2 likes471 downloads3y agoHugging Face24Tang-xiaoxiao /3D-RAD [ 🎯 NeurIPS 2025 ] 3D-RAD 🩻: A Comprehensive 3D Radiology Med-VQA Dataset with Multi-Temporal Analysis and Diverse Diagnostic Tasks 📢 News What's New in This Update 🚀 2025.10.23: 🔥 Updated the latest version of the paper! 2025.09.19: 🔥 Paper accepted to NeurIPS 2025! 🎯 2025.05.16: 🔥 Set up the repository and committed the dataset! 🔍 Overview 💡 In this repository, we present the dataset for "3D-RAD: A… See the full description on the dataset page: https://huggingface.co/datasets/Tang-xiaoxiao/3D-RAD.tabular100K<n<1M9 likes420 downloads11mo agoHugging Face25xyz1901901 /patientcare Patient Care Activity Recognition Dataset Dataset Summary A multi-class video action recognition dataset covering 23 clinically relevant patient care activities recorded in hospital-like environments. The dataset includes both real-life recordings and synthetically generated videos. The dataset is intended to support research in automated patient monitoring and clinical activity recognition using computer vision. Supported Tasks Video… See the full description on the dataset page: https://huggingface.co/datasets/xyz1901901/patientcare.textvideo-classification10K<n<100K0 likes365 downloads2mo agoHugging Face26lyffseba /xai-studies xai-studies xai = explainable AI. Offline mirror of lyffseba/xai. GitHub lyffseba/xai portal spaces/lyffseba/xai use python3 studies/run.py test No pip. No network. catalog catalog/models.csv id name status total active ctx experts inkling Inkling weights_public 975B 41B 1048576 6/256+2 shared kimi-k3 Kimi K3 api_live_weights_pending 2.8T 1048576 16/896 laguna-s-2.1 Laguna S 2.1 weights_public 118B 8B 1048576… See the full description on the dataset page: https://huggingface.co/datasets/lyffseba/xai-studies.textothern<1K0 likes359 downloads14h agoHugging Face27FreshCrawl /rednote-xiaohongshu-notes RedNote (Xiaohongshu) Notes with Engagement and Save Rates 10,427 posts from RedNote (小红书 / Xiaohongshu), across 8 content verticals, with full Chinese post text, four separate engagement metrics, and province-level geography. Most social datasets give you text and a like count. RedNote separates saves from likes, and that single distinction turns out to measure something the like count cannot. The finding this dataset exists for A post can be useful or it can be… See the full description on the dataset page: https://huggingface.co/datasets/FreshCrawl/rednote-xiaohongshu-notes.tabulartext-classification10K<n<100K0 likes320 downloads23d agoHugging Face28HiTZ /xnli-eu Dataset Card for XNLIeu XNLIeu is an extension of XNLI translated from English to Basque. It has been designed as a cross-lingual dataset for the Natural Language Inference task, a text-classification task that consists on classifying pairs of sentences, a premise and a hypothesis, according to their semantic relation out of three possible labels: entailment, contradiction and neutral. Dataset Details Dataset Description XNLI is a popular Natural… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/xnli-eu.text100K<n<1M0 likes310 downloads6mo agoHugging Face29xwjzds /extractive_qa_question_answering_hr Dataset Card HR-Multiwoz is a fully-labeled dataset of 5980 extractive qa spanning 10 HR domains to evaluate LLM Agent. It is the first labeled open-sourced conversation dataset in the HR domain for NLP research. Please refer to HR-MultiWOZ: A Task Oriented Dialogue (TOD) Dataset for HR LLM Agent for details about the dataset construction. Dataset Sources Repository: xwjzds/extractive_qa_question_answering_hr Paper: HR-MultiWOZ: A Task Oriented Dialogue (TOD)… See the full description on the dataset page: https://huggingface.co/datasets/xwjzds/extractive_qa_question_answering_hr.text1K<n<10K9 likes299 downloads3y agoHugging Face30animicaorg /x402-price-index x402 Price Index — a census of the machine-payable web Most directories of x402 services list what merchants claim. This dataset records what their endpoints actually answer when probed: the HTTP 402 challenge they return, the price inside it, the chain and asset they want, and whether they respond at all. It is built by Animica from an independent prober that walks every x402 endpoint it can discover and reads each merchant's own accepts[] block. No merchant self-reporting is… See the full description on the dataset page: https://huggingface.co/datasets/animicaorg/x402-price-index.tabular1K<n<10K0 likes297 downloads2d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.