datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
tcga-brca-titan-idc-ilc
tcga-brca-titan-idc-ilc
1. Tổng quan
[CẦN ĐIỀN THỦ CÔNG: mục đích, ngữ cảnh tạo dataset]
Tổng số bản ghi (cộng tất cả manifest phát hiện được): 4228
Số manifest phát hiện được trong bộ nhớ: 3 (df, brca_df, full_df)
Repo HuggingFace: okbro1234/tcga-brca-titan-idc-ilc
2. Cấu trúc lưu trữ tại đích
/ # suy từ hàm `HfApi`
file.txt # suy từ hàm `HfApi`
lfs.bin # suy từ hàm `HfApi`
shard_{i}_of_5.bin # suy từ hàm `HfApi`
remote/file/path.h5 #… See the full description on the dataset page: https://huggingface.co/datasets/okbro1234/tcga-brca-titan-idc-ilc.biolatent-brca-tcgaYear: 2025License: TCGA/GDC Data Use PoliciesAuthor: Sepideh Moafi
BioLatent-BRCA-TCGA Dataset
Dataset Summary
A processed transcriptomic dataset derived from TCGA-BRCA RNA-seq data, developed as part of the OmniLatent research project for representation learning on high-dimensional gene-expression data.
The dataset contains 1,231 samples × 23,375 genes with log1p(TPM) transformation and gene filtering applied.
Source and Provenance
Source: NCI… See the full description on the dataset page: https://huggingface.co/datasets/Sepideh2027/biolatent-brca-tcga.TCGA_BRCA_DatasetTCGA-BRCA-Details-OpenQATCGA-BRCA-Details-CloseQA
