vandijklab/immune-c2s
Overview Cell2Sentence is a novel method for adapting large language models to single-cell transcriptomics. We transform single-cell RNA sequencing data into sequences of gene names ordered by expression level, termed "cell sentences". This dataset was constructed from the immune tissue dataset in Domínguez et al., and it was used to train the Pythia-160m model capable of generating complete cells described in our paper. Details about the Cell2Sentence transformation and… See the full description on the dataset page: https://huggingface.co/datasets/vandijklab/immune-c2s.
updated readme
delete dataset_dict.json
deleted arrow val
deleted arrow train
deleted arrow test
Upload README.md with huggingface_hub
Upload data/val-00000-of-00001-87795001142a0e9a.parquet with huggingface_hub
Upload data/test-00000-of-00001-73deed04f25e2fa7.parquet with huggingface_hub
Upload data/train-00004-of-00005-e56fb5645860b585.parquet with huggingface_hub
Upload data/train-00003-of-00005-51f4bfadf690916b.parquet with huggingface_hub
Upload data/train-00002-of-00005-60ef01b633f99095.parquet with huggingface_hub
Upload data/train-00001-of-00005-e5c578d2a250c963.parquet with huggingface_hub
Upload data/train-00000-of-00005-28b282e43646543d.parquet with huggingface_hub
uploaded correct files
upload data
initial commit
