datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
declref-01-declref_01_symbols-2B
declref-01-declref_01_symbols-2B
Procedurally generated decl-ref-01 documents — the scoped declare/reference language of declref, with noisy references, sparse part-transition chains, interleaved openings, and periodic topic shifts tuned so that a model trained on it matches natural-language / code entropy dynamics (positional entropy profile, its fluctuation texture, and the entropy-quantile distribution), and holds that match as training doubles. Each document is a random… See the full description on the dataset page: https://huggingface.co/datasets/alexkstern/declref-01-declref_01_symbols-2B.gene-symbols-v1
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
gene_symbols
This dataset consists of a collection of human gene symbols, including well-known entries like VEGFA, TNF, and BRCA1. Each sample represents a single gene identifier formatted as a standard uppercase text string. The data appears to be a curated list of significant genes often associated with cancer research or cellular signaling pathways.
Dataset size
There are 70 data… See the full description on the dataset page: https://huggingface.co/datasets/joduor/gene-symbols-v1.gene-symbols
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
gene_symbols
This dataset consists of a collection of human gene symbols, including well-known entries like VEGFA, TNF, and BRCA1. Each sample represents a single gene identifier formatted as a standard uppercase text string. The data appears to be a curated list of significant genes often associated with cancer research or cellular signaling pathways.
Dataset size
There are 70 data… See the full description on the dataset page: https://huggingface.co/datasets/joduor/gene-symbols.loremipsum-29k-symbolschanged_symbolsnazy-symbols-classification-openclip-encoded-image-data
