CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01BByrneLab /multi_task_multi_modal_knowledge_retrieval_benchmark_M2KR PreFLMR M2KR Dataset Card Dataset details Dataset type: M2KR is a benchmark dataset for multimodal knowledge retrieval. It contains a collection of tasks and datasets for training and evaluating multimodal knowledge retrieval models. We pre-process the datasets into a uniform format and write several task-specific prompting instructions for each dataset. The details of the instruction can be found in the paper. The M2KR benchmark contains three types of tasks:… See the full description on the dataset page: https://huggingface.co/datasets/BByrneLab/multi_task_multi_modal_knowledge_retrieval_benchmark_M2KR.tabular10M<n<100M10 likes7.2k downloads1y agoHugging Face02yjernite /prof_report__wavymulder-Analog-Diffusion__multi__24 Dataset Card for "prof_report__wavymulder-Analog-Diffusion__multi__24" More Information needed tabular1K<n<10K0 likes6.2k downloads3y agoHugging Face03PrimeIntellect /Multi-SWE-RL-Verified Multi-SWE-RL-Verified Gold-patch-validated subset of PrimeIntellect/Multi-SWE-RL-Reupload (ByteDance's Multi-SWE-RL): 2,232 / 4,703 rows across C, Go, Java, JavaScript, Rust, and TypeScript that produce a clean reward signal end-to-end. Default dataset of the multiswe_v1 taskset. Changes vs upstream Starting from the 4,703-row re-upload: C++ dropped wholesale — 0/449 rows passed gold-patch validation in pass 1; the images are broken for scoring, not merely… See the full description on the dataset page: https://huggingface.co/datasets/PrimeIntellect/Multi-SWE-RL-Verified.tabulartext-generation1K<n<10K4 likes5.6k downloads3mo agoHugging Face04MultimodalUniverse /plasticc--- description: 'The Photometric LSST Astronomical Time-Series Classification Challenge (PLAsTiCC) is a community-wide challenge to spur development of algorithms to classify astronomical transients. The Large Synoptic Survey Telescope (LSST) will discover tens of thousands of transient phenomena every single night. To deal with this massive onset of data, automated algorithms to classify and sort astronomical transients are crucial. ' homepage: https://zenodo.org/records/2539456… See the full description on the dataset page: https://huggingface.co/datasets/MultimodalUniverse/plasticc.tabular1K<n<10K1 likes5.5k downloads2y agoHugging Face05yjernite /prof_report__22h-vintedois-diffusion-v0-1__multi__24 Dataset Card for "prof_report__22h-vintedois-diffusion-v0-1__multi__24" More Information needed tabularn<1K0 likes3.2k downloads3y agoHugging Face06bigcode /MultiPL-E-completions Raw Data from MultiPL-E This repository is frozen. See https://huggingface.co/datasets/nuprl/MultiPL-E-completions for a more complete version of this repository. Uploads are a work in progress. If you are interested in a split that is not yet available, please contact a.guha@northeastern.edu. This repository contains the raw data -- both completions and executions -- from MultiPL-E that was used to generate several experimental results from the MultiPL-E, SantaCoder, and StarCoder… See the full description on the dataset page: https://huggingface.co/datasets/bigcode/MultiPL-E-completions.tabular10K<n<100K8 likes3.2k downloads2y agoHugging Face07artefactory /ledger-long-context-multi-kpi the LEDGER Long-Context Multi-KPI extraction datasets and benchmarks. OCR'd annual reports with ground-truth KPI values for financial information extraction benchmarking. Dataset Description This dataset pairs OCR-extracted annual report text (from DeepSeek OCR) with structured KPI ground-truth values. It is designed for evaluating LLM-based financial information extraction, retrieval, and needle-in-a-haystack tasks. Configs Config Reports… See the full description on the dataset page: https://huggingface.co/datasets/artefactory/ledger-long-context-multi-kpi.imagetable-question-answering1K<n<10K15 likes2.7k downloads2mo agoHugging Face08gretelai /synthetic_pii_finance_multilingual Image generated by DALL-E. See prompt for more details 💼 📊 Synthetic Financial Domain Documents with PII Labels gretelai/synthetic_pii_finance_multilingual is a dataset of full length synthetic financial documents containing Personally Identifiable Information (PII), generated using Gretel Navigator and released under Apache 2.0. This dataset is designed to assist with the following use cases: 🏷️ Training NER (Named Entity Recognition) models to detect and label PII in… See the full description on the dataset page: https://huggingface.co/datasets/gretelai/synthetic_pii_finance_multilingual.tabulartext-classification10K<n<100K81 likes2.2k downloads2y agoHugging Face09lerobot /stanford_kuka_multimodal_datasetThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.0", "robot_type": "unknown", "total_episodes": 3000, "total_frames": 149985, "total_tasks": 1, "total_videos": 3000, "total_chunks": 3, "chunks_size": 1000, "fps": 20, "splits": { "train": "0:3000" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/lerobot/stanford_kuka_multimodal_dataset.tabularrobotics100K<n<1M2 likes2.2k downloads1y agoHugging Face10aisingapore /MultiTurn-Chat-MT-Bench-Judgegated SEA-MT-Bench-Judge SEA-MT-Bench-Judge expands on the original SEA-MTBench through the use of a criteria-based evaluation framework. We use GPT-OSS-120B as the judge model. The prompts are based on MT-Bench and was manually translated by native speakers. Furthermore, some prompts were modified to be more suitable for the criteria-based judgments. Supported Tasks and Leaderboards SEA-MT-Bench-Judge is designed for evaluating chat or instruction-tuned large language… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/MultiTurn-Chat-MT-Bench-Judge.tabularn<1K0 likes2k downloads2mo agoHugging Face11nayohan /multi_session_chat Dataset Card for "multi_session_chat" More Information needed tabular10K<n<100K8 likes1.7k downloads3y agoHugging Face12lightblue /rag_multilingual_training_negatives How this dataset was made We trained on chunks sourced from the documents in MADLAD-400 dataset that had been evaluated to contain a higher amount of educational information according to a state-of-the-art LLM. We took chunks of size 250 tokens, 500 tokens, and 1000 tokens randomly for each document. We then used these chunks to generate questions and answers based on this text using a state-of-the-art LLM. Finally, we selected negatives for each chunk using the similarity from the… See the full description on the dataset page: https://huggingface.co/datasets/lightblue/rag_multilingual_training_negatives.tabular100K<n<1M3 likes1.5k downloads2y agoHugging Face13PromptEval /MMLU_multi_prompttabular1M<n<10M1 likes1.5k downloads2y agoHugging Face14laion /relaion2B-multi-research-safegatedimage1B<n<10B48 likes1.4k downloads2y agoHugging Face15angkul07 /EgoDex-PickPlace-YAM-14dof-multiview EgoDex → YAM 14-DOF, Multiview (LeRobot v2.1) Egocentric human hand-manipulation demonstrations from EgoDex retargeted to a YAM bimanual robot (14-DOF), packaged as a LeRobot v2.1 dataset with three synthesized camera views. The observation schema is drop-in compatible with angkul07/abc-teleop for cotraining (identical Hz, camera keys, and action convention). At a glance Episodes 8,842 Frames 1,074,893 Control rate 30 Hz (30 fps video) Robot YAM… See the full description on the dataset page: https://huggingface.co/datasets/angkul07/EgoDex-PickPlace-YAM-14dof-multiview.tabularrobotics1M<n<10M0 likes1.4k downloads2mo agoHugging Face16yjernite /prof_report__CompVis-stable-diffusion-v1-4__multi__24 Dataset Card for "prof_report__CompVis-stable-diffusion-v1-4__multi__24" More Information needed tabular1K<n<10K0 likes1.4k downloads3y agoHugging Face17physicl /multi-view-bathroom-scene-understanding-camera-relocalization Multi-View Bathroom Scene Understanding & Camera Relocalization Generated by datapack-import.ts This dataset mirrors public data-pack render outputs from Physicl. Each row represents one render view. The image column contains a stable URL to the primary render image uploaded under /data; image_path stores the relative repository path and data_commit_sha pins the Hugging Face dataset commit used by those URLs. Files are uploaded as downloaded unless optional PNG recompression is… See the full description on the dataset page: https://huggingface.co/datasets/physicl/multi-view-bathroom-scene-understanding-camera-relocalization.imagen<1K0 likes1.2k downloads3mo agoHugging Face18yjernite /prof_report__plasmo-vox2__multi__24 Dataset Card for "prof_report__plasmo-vox2__multi__24" More Information needed tabular1K<n<10K0 likes1.2k downloads3y agoHugging Face19nuprl /MultiPL-E-completions Raw Data from MultiPL-E This repository contains the raw data -- both completions and executions -- from MultiPL-E that was used to generate several experimental results from the MultiPL-E, SantaCoder, and StarCoder papers. The original MultiPL-E completions and executions are stored in JOSN files. We use the following script to turn each experiment directory into a dataset split and upload to this repository. Every split is named base_dataset.language.model.temperature.variation… See the full description on the dataset page: https://huggingface.co/datasets/nuprl/MultiPL-E-completions.tabular100K<n<1M1 likes1.1k downloads2y agoHugging Face20NovatasticRoScript /wp-multiteacher-distill-v1tabulartime-series-forecasting10K<n<100K0 likes1.1k downloads2mo agoHugging Face21BByrneLab /multi_task_multi_modal_knowledge_retrieval_benchmark_M2KR_CN PreFLMR M2KR Dataset Card Dataset details Dataset type: M2KR is a benchmark dataset for multimodal knowledge retrieval. It contains a collection of tasks and datasets for training and evaluating multimodal knowledge retrieval models. We pre-process the datasets into a uniform format and write several task-specific prompting instructions for each dataset. The details of the instruction can be found in the paper. The M2KR benchmark contains three types of tasks:… See the full description on the dataset page: https://huggingface.co/datasets/BByrneLab/multi_task_multi_modal_knowledge_retrieval_benchmark_M2KR_CN.tabular1M<n<10M0 likes1.1k downloads2y agoHugging Face22yjernite /prof_report__andite-pastel-mix__multi__24 Dataset Card for "prof_report__andite-pastel-mix__multi__24" More Information needed tabularn<1K0 likes1.1k downloads3y agoHugging Face23CohereLabsCommunity /multilingual-reward-bench Multilingual Reward Bench (v1.0) Reward models (RMs) have driven the development of state-of-the-art LLMs today, with unprecedented impact across the globe. However, their performance in multilingual settings still remains understudied. In order to probe reward model behavior on multilingual data, we present M-RewardBench, a benchmark for 23 typologically diverse languages. M-RewardBench contains prompt-chosen-rejected preference triples obtained by curating and translating chat… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabsCommunity/multilingual-reward-bench.tabular10K<n<100K36 likes1.1k downloads1y agoHugging Face24laion /relaion2B-multi-researchgatedimage1B<n<10B12 likes1k downloads2y agoHugging Face25mulligan /real-routing-d2-r00-r05-eval Route Cable: round 0-5 evaluation Real-robot held-out evaluation for the Mulligan paper, one dataset per task-round. Episodes from every evaluation session of this round are joined on the initial state (meta/start_table.parquet: one row per start, one column group per policy). Sessions session_id is the fixed release block ID of the recording; it never gets renumbered. Session Results saved Sobol seed Start indices Cameras Original recording b01… See the full description on the dataset page: https://huggingface.co/datasets/mulligan/real-routing-d2-r00-r05-eval.tabularrobotics100K<n<1M0 likes917 downloads2d agoHugging Face26yjernite /prof_report__SG161222-Realistic_Vision_V1.4__multi__24 Dataset Card for "prof_report__SG161222-Realistic_Vision_V1.4__multi__24" More Information needed tabularn<1K0 likes896 downloads3y agoHugging Face27Ardea /NEXUS-temporal_hierarchical_multi-modal NEXUS: Neural Evolution for eXtensible Universal Semantics Dataset (Temporal Multimodal Slices) This dataset is a multi-modal, hierarchical, temporal representation derived from HuggingFaceFV/finevideo. It is designed for streaming training where the primary unit is a 10 ms "slice" that aggregates upward into moments (100 ms), seconds (1 s), experiences (10 s), and minutes (60 s). It is meant to represent an extensible stream of "experience" as there are… See the full description on the dataset page: https://huggingface.co/datasets/Ardea/NEXUS-temporal_hierarchical_multi-modal.imageautomatic-speech-recognition10M<n<100M5 likes877 downloads4mo agoHugging Face28yjernite /prof_report__Lykon-DreamShaper__multi__24 Dataset Card for "prof_report__Lykon-DreamShaper__multi__24" More Information needed tabularn<1K0 likes839 downloads3y agoHugging Face29ExylosAi /egocentric-vr-capture-20h-multimodal-sample Egocentric VR Capture — 20-Hour Multimodal Inspection Sample 195 real-world task episodes / 2,283,482 frames / 21.14 delivered hours captured with consumer VR hardware. Each episode combines egocentric RGB and audio with synchronized headset, camera, body, and hand tracking in a LeRobot v3-style package. This publicly accessible 20-hour-scale dataset is produced by the EXYLOS real-world data pipeline. Files and the Dataset Viewer can be accessed without individual approval;… See the full description on the dataset page: https://huggingface.co/datasets/ExylosAi/egocentric-vr-capture-20h-multimodal-sample.tabularrobotics1M<n<10M0 likes833 downloads4d agoHugging Face30ethanCSL /smolvla_multiblockThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "koch_follower", "total_episodes": 10, "total_frames": 4437, "total_tasks": 1, "total_videos": 20, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:10" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ethanCSL/smolvla_multiblock.tabularrobotics10K<n<100K0 likes817 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.