datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
BioFactory-Asian-Oral-Microbiome-Preview
🦷 World's First Chinese-Anchored Oral Multi-Omics Synthetic Dataset (v2026)
3,872 production records · PERMANOVA p=0.994 · MMD=0.060 · 100% synthetic = zero GDPR risk
📄 Full Whitepaper ·
📋 1-Page Executive Summary ·
🔬 38-Sample Preview ·
📧 Enterprise: jerry820402@hotmail.com
English Executive Summary
This is the world's first Asian/Chinese-specific oral microbiome multi-omics synthetic dataset, generated via Evo foundation model inference on 8× NVIDIA A800… See the full description on the dataset page: https://huggingface.co/datasets/jerry982/BioFactory-Asian-Oral-Microbiome-Preview.crazy-asian-gfs
