CoolFace
Datasetpublic

macwiatrak/operon-identification-long-read-rna-sequencing-protein-sequences

Dataset for operon identification from long-read RNA sequencing A dataset of annotated operons across 5 distinct bacterial strains. The operons were annotated by running and analysing long-read RNA sequencing and identifying genes located on the same transcripts. The genome protein sequences have been extracted from GenBank. Each row contains whole bacterial genome represented by an ordered list of protein sequences. Usage For a complete example on how to read and… See the full description on the dataset page: https://huggingface.co/datasets/macwiatrak/operon-identification-long-read-rna-sequencing-protein-sequences.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
0likes34downloads
3 commits on main
d77df901y ago

Update README.md

macwiatrak
99e036a1y ago

Upload dataset

macwiatrak
f3250071y ago

initial commit

macwiatrak