pbm
Datasets
All datasets matching “pbm”PBMC_ATLAS10M_PBMCsplatter-cube-pbmc3k
Splatter cube: controlled scRNA-seq variants with known cluster structure
180 simulated datasets: 60 parameter points x 3 seeds, each 2,000 cells x 16,085 genes with a known number of clusters.
Generated with splatter, with baseline
parameters estimated from real reference data rather than chosen by hand.
Files
data/pP_sS.h5ad -- counts (CSR). obs["Group"] holds the ground-truth cluster label.
metrics.csv -- one row per simulation: parameters, realised sparsity… See the full description on the dataset page: https://huggingface.co/datasets/btraven/splatter-cube-pbmc3k.aida-asian-pbmc-cell-sentence-top2000
AIDA Asian PBMC Cell Sentences (Top 2000 Genes)
Dataset Description
This dataset contains 1,265,624 single cells from peripheral blood mononuclear cells (PBMCs) of 619 healthy donors across 5 Asian countries, transformed into "cell sentences" - space-separated gene symbols ordered by expression level.
Each cell is represented as a sequence of the top 2,000 most highly expressed genes, enabling language model-style analysis of single-cell transcriptomics data.… See the full description on the dataset page: https://huggingface.co/datasets/transhumanist-already-exists/aida-asian-pbmc-cell-sentence-top2000.10x-pbmc-5kpb-mutants-v1
