datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SWE-Sharp-Bench
SWE-Sharp-Bench
SWE-Sharp-Bench is a comprehensive benchmark suite for evaluating software engineering capabilities of AI agents and models on C# and .NET codebases. This benchmark extends the SWE-Bench framework to the C# ecosystem, providing real-world software engineering tasks from popular open-source repositories.
Code - https://github.com/microsoft/prose/tree/main/misc/SWE-Sharp-Bench
Research Paper Draft & Benchmark Analysis: https://aka.ms/swesharparxiv
Contact… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/SWE-Sharp-Bench.PEACE
PEACE: Empowering Geologic Map Holistic Understanding with MLLMs
[Code] [Paper] [Data]
Introduction
We construct a geologic map benchmark, GeoMap-Bench, to evaluate the performance of MLLMs on geologic map understanding across different abilities, the overview of it is as shown in below Table.
Property
Description
Source
USGS(English)
CGS(Chinese)
Content
Image-question pair… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/PEACE.delulu-fim-benchmarkDelulu — Fill-in-the-Middle Code Hallucination Benchmark
A verified multilingual benchmark for code-completion hallucinations.
Every golden completion compiles. Every hallucination provably doesn't.
📄 Read the preprint on arXiv →
Every Delulu sample ships as a self-contained Docker image. The viewer above lets you browse the dataset, pull a sample's verifier, and re-run verify golden / verify hallucinated / verify patch <your-completion> with one… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/delulu-fim-benchmark.WorkflowPerturb
WorkflowPerturb — Dataset Artifact
Companion data for the EMNLP 2026 Industry Track paper
“WorkflowPerturb: Calibrated Stress Tests for Evaluating Multi-Agent Workflow Metrics.”
Canonical location: https://huggingface.co/datasets/microsoft/WorkflowPerturbPaper: https://arxiv.org/abs/2602.17990
This release is the complete WorkflowPerturb benchmark plus documentation. It is
self-contained: the CSVs carry every golden workflow, every perturbed variant, and all
shipped pre-computed… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/WorkflowPerturb.Microbenchhnm-search-data
HnM Search Dataset Created from Recommendations Dataset
This synthetic data-set is created using the recommendations dataset:
https://huggingface.co/datasets/einrafh/hnm-fashion-recommendations-data (Use of this dataset is subject to the terms and conditions set forth on the original distribution page. This dataset is intended for non-commercial and research use.)
https://www.kaggle.com/competitions/h-and-m-personalized-fashion-recommendations/data (DATA ACCESS AND USE:… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/hnm-search-data.microduck-trajectory-dataset
🦆 MicroDuck Bipedal Robot Simulation Trajectory Dataset
This dataset contains multi-modal state-action trajectories collected from the MicroDuck 14-DOF bipedal robot simulation in MuJoCo.
📂 Formats Available
.csv files: Human-readable tabular format containing timestamps, base pose (pos & quat), linear/angular velocities, 14 motor angles, velocities, applied torques (ctrl), and user velocity commands.
.npz files: Compressed NumPy tensor arrays (obs, actions… See the full description on the dataset page: https://huggingface.co/datasets/allen73/microduck-trajectory-dataset.claimify-dataset
Dataset Card
Dataset Overview
This dataset is associated with the paper Towards Effective Extraction and Evaluation of Factual Claims by Dasha Metropolitansky and Jonathan Larson, accepted to the ACL 2025 Main Conference. See also our video and blog post.
The dataset contains 6,490 sentences, each annotated with a binary label indicating whether it contains a verifiable factual claim. These sentences were extracted from the 396 answers in the BingCheck dataset (Li et al.… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/claimify-dataset.microduck-locomotion-dataset-v1
🦆 MicroDuck Locomotion Dataset V1.0: Zero-Slip Flat Ground Baseline
Publisher: devorah-ai-2026 / devorahai2026License: Commercial / CC-BY-SA 4.0Task: Bipedal Locomotion, Reinforcement Learning, Sim-to-RealHardware Platform: Antigravity 14-DOF MicroDuck Biped Robot
📌 Overview (개요)
본 데이터셋은 **14-DOF 소형 2족 보행 로봇(MicroDuck)**의 완벽한 평면 보행(Flat-ground walking) 궤적 및 제어 로그를 담은 고품질 합성 데이터(Synthetic Data)입니다. 기존 강화학습 기반 2족 보행 에이전트들이 흔히 겪는 발바닥 미끄러짐(Ice-Skating) 및… See the full description on the dataset page: https://huggingface.co/datasets/devorah-ai-2016/microduck-locomotion-dataset-v1.MicroG-HAR-train-ready
MicroG-HAR-train-ready dataset.
This dataset is converted from the MicroG-4M dataset.
For more information please:
Refer to our paper
Visit our GitHub
Datasets formatted in this way can be used directly as input data directories for the PySlowFast_for_HAR framework without additional preprocessing, enabling seamless model training and evaluation.
This dataset follows the organizational format of the AVA dataset. The only difference is that the original CSV header… See the full description on the dataset page: https://huggingface.co/datasets/lei-qi-233/MicroG-HAR-train-ready.africa-synth-financial-inclusion-microfinance-access-africa-all
Africa Synth Financial Inclusion Microfinance Access Africa All | Africa (Electric Sheep Africa metadata inventory)
Size category: 10K<n<100K - Formats: csv - Sector: economics_finance - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-synth-financial-inclusion-microfinance-access-africa-all.Burmese-Microbiology-1K
Burmese-Microbiology-1K
Min Si Thu, min@globalmagicko.com
Microbiology 1K QA pairs in Burmese Language
Purpose
Before this Burmese Clinical Microbiology 1K dataset, the open-source resources to train the Burmese Large Language Model in Medical fields were rare.
Thus, the high-quality dataset needs to be curated to cover medical knowledge for the development of LLM in the Burmese language
Motivation
I found an old notebook in my box. The book was… See the full description on the dataset page: https://huggingface.co/datasets/jojo-ai-mst/Burmese-Microbiology-1K.reddit_finance_posts_apple-tesla-microsoft
Reddit Finance Posts Dataset (Apple, Tesla, Microsoft)
This dataset contains 12046 Reddit Posts collected from 20 finance-related subreddits via the Reddit API using the keywords Apple, Tesla, and Microsoft.
The included subreddits are:
stocks, wallstreetbets, investing, StockMarket, options, RobinHood,
pennystocks, SecurityAnalysis, personalfinance, Dividends, CryptoCurrency,
CryptoMarkets, ETFs, FinancialIndependence, ValueInvesting, quant,
algotrading, forex, economy, Superstonk… See the full description on the dataset page: https://huggingface.co/datasets/emilpartow/reddit_finance_posts_apple-tesla-microsoft.Human_Gut_Microbiome_Data
Human Gut Microbiome Dataset (40-class)
Pre-processed, train/val/test-split human gut microbiome dataset used to train and evaluate
MicrobiomeFM, a Transformer-based foundation model for multi-class disease classification
from shotgun metagenomic data.
This release contains the 40-class filtered version of the dataset that the published
results were obtained on (test accuracy ≈ 96.05%, weighted F1 ≈ 0.950, macro F1 ≈ 0.691).
It is derived from the curatedMetagenomicData (cMD)… See the full description on the dataset page: https://huggingface.co/datasets/mohitraiyani27/Human_Gut_Microbiome_Data.vaccine-hesitancy-for-covid-19-public-use-microdat
Vaccine Hesitancy for COVID-19: Public Use Microdata Areas (PUMAs)
Description
Due to the change in the survey instrument regarding intention to vaccinate, our estimates for “hesitant or unsure” or “hesitant” derived from April 14-26, 2021, are not directly comparable with prior Household Pulse Survey data and should not be used to examine trends in hesitancy.
To support state and local communication and outreach efforts, ASPE developed state, county, and sub-state level… See the full description on the dataset page: https://huggingface.co/datasets/HHS-Official/vaccine-hesitancy-for-covid-19-public-use-microdat.microsoft-phi2-mental-healthclinical-rtt-micro-intervention-trajectory-nudging-v0.1What this dataset tests
Whether a model can recommend small, precise actionsthat shift recovery into a more stable basin.
It ties actions to early signals.It penalizes overload.
Typical failures
pushing intensity in stalled states
ignoring sleep or adherence signals
recommending generic advice without target
Suggested prompt wrapper
System
You propose micro-interventions to nudge recovery.
User
Predicted Topology{predicted_topology}
Early Signals{early_signals}
Risk… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-rtt-micro-intervention-trajectory-nudging-v0.1.hsclassify-micro-dataset
Dataset Card for HSClassify Micro Training Dataset
Dataset Summary
This dataset supports multilingual HS code classification for customs and trade workflows.
It combines:
HS nomenclature records (6-digit level and hierarchy context)
Synthetic product descriptions mapped to HS codes
Human-readable chapter/category labels for UI and latent-space analysis
Included Files
training_data_indexed.csv: training rows with text, HS code, chapter metadata, and language.… See the full description on the dataset page: https://huggingface.co/datasets/Mead0w1ark/hsclassify-micro-dataset.air-pollution-mean-annual-exposure-micrograms-per-cubic-meter-africa
Air Pollution Mean Annual Exposure Micrograms Per Cubic Meter Africa | Africa (World Bank)
Size category: n<1K - Formats: csv - Sector: climate_environment - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This Dataset Covers
Public… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/air-pollution-mean-annual-exposure-micrograms-per-cubic-meter-africa.microfluidic-pressure-drop-benchmark
Microfluidic Straight-Tube Pressure-Drop Benchmark
This small, deterministic engineering dataset contains 420 straight-tube pressure-drop cases spanning tube internal diameter, length, volumetric flow rate, and dynamic viscosity. It is intended for equation verification, unit-conversion tests, engineering education, tabular regression experiments, and first-pass fluid-path screening.
Why this dataset exists
Small changes in tube internal diameter can dominate a… See the full description on the dataset page: https://huggingface.co/datasets/AlexHu2026/microfluidic-pressure-drop-benchmark.new_epa_micro_all民生公共物聯網,微型感測器資料集,
整合 pm2.5 ,溫度,濕度資料,
每微型感測器每小時一筆,
先取樣 2023.3月 與 2023.4月 資料。
microsoft_malwaremicrolangMicrolang was designed to test text generation architectures
It consists of 16 tokens:
special (implicit):
<>
noun:
bob, tom, bike, speech
transitive verb:
take, use
intransitive verb:
talk, go
adjective:
good, active
adverb:
not
conjunction:
and, then, but
punctuation:
.
The tokenizer can be found on umarzein/microlang-utils and can be loaded this:
import transformers
tokenizer = transformers.PreTrainedTokenizerFast.from_pretrained("umarzein/microlang-utils")
cases_microbiomeThis dataset contains 105 patient descriptions that are looking for clinical trials investigating interventions targeting their microbiome.
The dataset was intentionally distributed across 18 disease categories of various disease settings, locations spread accross 17 countries, and broad spectrum of age ranges.
Disease Categories covered by the dataset:
Cardiovascular Diseases
Musculoskeletal Issues
Type 2 Diabetes
Skin Diseases
Gastrointestinal Disorders
Digestive Disorders
Mental Health… See the full description on the dataset page: https://huggingface.co/datasets/Elysr/cases_microbiome.A_dataset_collected_from_a_microgridmicrobiome-disease-dataset
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/elizah521/microbiome-disease-dataset.Microbench_mini
xuxuxuxuxu/Microbench_mini
This dataset includes a TSV file uploaded directly.
ABX-HT-006_microbiome_diversity_collapse-v0.1ABX-HT-006 Microbiome Diversity Collapse
Purpose
Detect early microbiome collapse that predicts C. difficile risk.
Core pattern
host_stress_index high
abx_exposure_index high
microbiome_coherence_index drops
shannon_drop_vs_expected stays high
domination_rise_vs_expected stays high
later_cdiff_flag appears
Files
data/train.csv
data/test.csv
scorer.py
Schema
Each row is one timepoint in a within series microbiome time course.
Required columns
row_id
series_id
timepoint_d
host_model
drug… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ABX-HT-006_microbiome_diversity_collapse-v0.1.microwaveQAMicroVQA_VLM
