CoolFace
Datasetpublic

damlab/uniprot

Dataset Description Dataset Summary This dataset is a mirror of the Uniprot/SwissProt database. It contains the names and sequences of >500K proteins. This dataset was parsed from the FASTA file at https://ftp.uniprot.org/pub/databases/uniprot/current_release/knowledgebase/complete/uniprot_sprot.fasta.gz. Supported Tasks and Leaderboards: None Languages: English Dataset Structure Data Instances Data Fields: id, description… See the full description on the dataset page: https://huggingface.co/datasets/damlab/uniprot.

sourceHugging Faceupdated 5y agoView on Hugging Face
4likes184downloads
../
filetrain-00000-of-00001.parquet148.2 MBdownload

damlab/uniprot · main · files are served by the source, never re-hosted here