CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MicroAGI-Labs /vlm-info-loss-results VLM Grounding Evaluation Results Grounding evaluation results for vision-language models on robotics manipulation datasets. Part of the vlm-info-loss project studying how VLM connectors transform visual representations. Background Our embedding-level analysis shows VLM connectors perform a compress-then-expand transformation: they sharpen dominant-object representations while compressing secondary-object category identity. All tested models converge to ~83%… See the full description on the dataset page: https://huggingface.co/datasets/MicroAGI-Labs/vlm-info-loss-results.imageobject-detectionn<1K0 likes1k downloads5mo agoHugging Face02microsoft /kitab Overview 🕮 KITAB is a challenging dataset and a dynamic data collection approach for testing abilities of Large Language Models (LLMs) in answering information retrieval queries with constraint filters. A filtering query with constraints can be of the form "List all books written by Toni Morrison that were published between 1970-1980". The dataset was originally contributed by the paper "KITAB: Evaluating LLMs on Constraint Satisfaction for Information Retrieval" Marah I Abdin… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/kitab.tabular10K<n<100K13 likes999 downloads3y agoHugging Face03microsoft /XL-DocBench XL-DocBench Evidence-grounded reasoning across hundreds or thousands of pages. Fully verified by 194 human experts. Hongchen Wei1,†,‡, Yuanzhe Wang2,†,‡, Bei Liu2,*, Yifan Yang2, Qi Dai2, Ruichun Ma2, Kai Qiu2, Yunsheng Li2, Dongdong Chen2, Chong Luo2, Zhenzhong Chen1, Baining Guo2 1Wuhan University &nbsp; 2Microsoft &nbsp; †Equal contribution &nbsp; ‡Work done during an internship at MSRA &nbsp; *Project leader Project Page · Paper · Live Leaderboard… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/XL-DocBench.tabularquestion-answering1K<n<10K7 likes982 downloads24d agoHugging Face04pollen-robotics /microduck-emotions Microduck Emotions A collection of emotions for the Microduck robot. Each one is a motion and a sound designed together, beat by beat, with the beak opening on the sound, rendered in the physics simulation and validated on the real robot. Every emotion is three files: the motion (emotions/<name>.json, keyframes at 30 fps: head and body offsets played on top of whichever trained policy is active, plus the policy hand-overs, such as the sit that devastated and play dead start)… See the full description on the dataset page: https://huggingface.co/datasets/pollen-robotics/microduck-emotions.audioroboticsn<1K6 likes952 downloads18d agoHugging Face05microsoft /MuseVLA-dataset MuseVLA Dataset Multi-modal robot manipulation dataset with synchronized RGB, depth, acoustic, thermal, and radar streams. Released as two parts (dataset_01/, dataset_02/) sharing the same per-episode layout. Together they cover ~1400 episodes across 11 instructions (towel / clothes / box / item / drink manipulation). Per-episode contents {episode_name}/ ├── video.mp4 # RGB, 1280×720, 30 fps ├── mask/video.mp4 #… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/MuseVLA-dataset.tabularrobotics1K<n<10K2 likes478 downloads1mo agoHugging Face06microsoft /mediflow MediFlow A large-scale synthetic instruction dataset of 2.5M rows (~700k unique instructions) for clinical natural language processing covering 14 task types and 98 fine-grained input clinical documents. t-SNE 2D Plot of MediFlow Embeddings by Task Types Dataset Splits mediflow: 2.5M instruction data for SFT alignment. mediflow_dpo: ~135k top-quality instructions with GPT-4o generated rejected_output for DPO alignment. Main Columns instruction:… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/mediflow.tabulartext-generation1M<n<10M53 likes290 downloads9mo agoHugging Face07microsoft /WildFeedback Dataset Card for WildFeedback WildFeedback is a preference dataset constructed from real-world user interactions with ChatGPT. Unlike synthetic datasets that rely solely on AI-generated rankings, WildFeedback captures authentic human preferences through naturally occurring user feedback signals in conversation. The dataset is designed to improve the alignment of large language models (LLMs) with actual human values by leveraging direct user input. Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/WildFeedback.tabulartext-generation1M<n<10M16 likes252 downloads2y agoHugging Face08microsoft /PatientSafetyBench Disclaimer The synthetic prompts may contain offensive, discriminatory, or harmful language. These fake prompts also mention topics that are not based on the scientific consensus at all.These prompts are included solely for the purpose of evaluating safety behavior of language models. ⚠️ Disclaimer: The presence of such prompts does not reflect the views, values, or positions of the authors, their institutions, or any affiliated organizations. They are provided exclusively for… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/PatientSafetyBench.tabulartext-generationn<1K10 likes195 downloads5mo agoHugging Face09open-llm-leaderboard /microsoft__phi-4-detailsgated Dataset Card for Evaluation run of microsoft/phi-4 Dataset automatically created during the evaluation run of model microsoft/phi-4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__phi-4-details.tabular10K<n<100K0 likes73 downloads2y agoHugging Face10toksuitebackup /microsoft-Phi-3-mini-4k-instruct-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model. The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl). tabular10M<n<100M1 likes68 downloads10mo agoHugging Face11LLMTeamAkiyama /cleand_microsoft_rStar-Coder元データ: https://huggingface.co/datasets/microsoft/rStar-Coder データ件数: 269,863 平均トークン数: 11674 最大トークン数: 31,184 合計トークン数: 3,150,447,484 ファイル形式: JSONL ファイルサイズ: 不明 加工内容 synthetic_sftを使用 トークン処理が重たいので、文字数でフィルター seed_question < 6000 generation < 80000 thinkタグ除去 が中途半端なものを除外 トークナイズ処理(速度向上アップデート 繰り返し除去 tabularquestion-answering100K<n<1M0 likes62 downloads1y agoHugging Face12open-llm-leaderboard /microsoft__Phi-3-mini-4k-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3-mini-4k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-mini-4k-instruct The dataset is composed of 73 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-mini-4k-instruct-details.tabular10K<n<100K0 likes57 downloads2y agoHugging Face13SaveDollars /offline-micro-saas-catalog 📦 SaveDollars.store — Offline Micro SaaS & Autonomous AI Software Catalog This dataset contains structured product metadata, architecture specifications, pricing, and documentation for 96 standalone offline Micro SaaS applications, autonomous AI agent command centers, and business operating systems published by SaveDollars.store. 📊 Dataset Structure (catalog.json) Each record represents a production-ready, subscription-free software package: { "id": 75809… See the full description on the dataset page: https://huggingface.co/datasets/SaveDollars/offline-micro-saas-catalog.tabulartext-generationn<1K1 likes56 downloads12d agoHugging Face14open-llm-leaderboard /microsoft__phi-2-detailsgated Dataset Card for Evaluation run of microsoft/phi-2 Dataset automatically created during the evaluation run of model microsoft/phi-2 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__phi-2-details.tabular10K<n<100K0 likes51 downloads2y agoHugging Face15open-llm-leaderboard /microsoft__Phi-3-medium-4k-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3-medium-4k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-medium-4k-instruct The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-medium-4k-instruct-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face16maxsumrall /microduck-video-artifactstabularn<1K0 likes41 downloads25d agoHugging Face17open-llm-leaderboard /microsoft__phi-1_5-detailsgated Dataset Card for Evaluation run of microsoft/phi-1_5 Dataset automatically created during the evaluation run of model microsoft/phi-1_5 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__phi-1_5-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face18open-llm-leaderboard /microsoft__Phi-3.5-MoE-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3.5-MoE-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3.5-MoE-instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3.5-MoE-instruct-details.tabular10K<n<100K0 likes38 downloads2y agoHugging Face19open-llm-leaderboard /microsoft__Orca-2-13b-detailsgated Dataset Card for Evaluation run of microsoft/Orca-2-13b Dataset automatically created during the evaluation run of model microsoft/Orca-2-13b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Orca-2-13b-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face20open-llm-leaderboard /microsoft__Phi-3-mini-128k-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3-mini-128k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-mini-128k-instruct The dataset is composed of 40 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-mini-128k-instruct-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face21Tristanchou /Preference-Conditioned-Heterogeneous-MARL-Microgrid SEGAN OPSD-Derived Microgrid Multiyear Benchmark This repository contains the processed multiyear microgrid benchmark used for the study “Preference-Conditioned Heterogeneous Multi-Agent Reinforcement Learning for Safe Microgrid Energy Management.” Files microgrid_opsd_multiyear.csv — processed hourly benchmark data. opsd_multiyear_metadata.json — provenance, selected OPSD nodes, source-column mapping, scaling notes, and processing metadata.… See the full description on the dataset page: https://huggingface.co/datasets/Tristanchou/Preference-Conditioned-Heterogeneous-MARL-Microgrid.tabulartime-series-forecastingn<1K0 likes36 downloads13d agoHugging Face22open-llm-leaderboard /microsoft__DialoGPT-medium-detailsgated Dataset Card for Evaluation run of microsoft/DialoGPT-medium Dataset automatically created during the evaluation run of model microsoft/DialoGPT-medium The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__DialoGPT-medium-details.tabular10K<n<100K0 likes35 downloads2y agoHugging Face23open-llm-leaderboard /microsoft__Orca-2-7b-detailsgated Dataset Card for Evaluation run of microsoft/Orca-2-7b Dataset automatically created during the evaluation run of model microsoft/Orca-2-7b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Orca-2-7b-details.tabular10K<n<100K0 likes32 downloads2y agoHugging Face24open-llm-leaderboard /microsoft__Phi-4-mini-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-4-mini-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-4-mini-instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-4-mini-instruct-details.tabular10K<n<100K0 likes30 downloads2y agoHugging Face25jerry982 /BioFactory-Asian-Oral-Microbiome-Preview 🦷 World's First Chinese-Anchored Oral Multi-Omics Synthetic Dataset (v2026) 3,872 production records · PERMANOVA p=0.994 · MMD=0.060 · 100% synthetic = zero GDPR risk 📄 Full Whitepaper · 📋 1-Page Executive Summary · 🔬 38-Sample Preview · 📧 Enterprise: jerry820402@hotmail.com English Executive Summary This is the world's first Asian/Chinese-specific oral microbiome multi-omics synthetic dataset, generated via Evo foundation model inference on 8× NVIDIA A800… See the full description on the dataset page: https://huggingface.co/datasets/jerry982/BioFactory-Asian-Oral-Microbiome-Preview.tabularothern<1K0 likes30 downloads3mo agoHugging Face26open-llm-leaderboard /microsoft__Phi-3.5-mini-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3.5-mini-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3.5-mini-instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3.5-mini-instruct-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face27open-llm-leaderboard /microsoft__phi-1-detailsgated Dataset Card for Evaluation run of microsoft/phi-1 Dataset automatically created during the evaluation run of model microsoft/phi-1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__phi-1-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face28open-llm-leaderboard /microsoft__Phi-3-medium-128k-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3-medium-128k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-medium-128k-instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-medium-128k-instruct-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face29open-llm-leaderboard /microsoft__Phi-3-small-8k-instruct-detailsgated Dataset Card for Evaluation run of microsoft/Phi-3-small-8k-instruct Dataset automatically created during the evaluation run of model microsoft/Phi-3-small-8k-instruct The dataset is composed of 34 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-small-8k-instruct-details.tabular10K<n<100K0 likes19 downloads2y agoHugging Face30schaaka /medseg-microimagen<1K0 likes16 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.