CoolFace
Datasetpublic

as-cle-bert/genetics-arxiv-wiki

Dataset Card for Dataset Name Small genetics-related text dataset based on 23200 ArXiv abstact records and 111 Wikipedia pages. Dataset Details Dataset Description Dataset was produced using the python scripts you will find in this GitHub repository. It represents a collection of genetics-related text data taken from ArXiv abstracts dataset and Wikipedia. Dataset holds a total of 23311 text records, 23200 of which belonging to categories q-bio.BM… See the full description on the dataset page: https://huggingface.co/datasets/as-cle-bert/genetics-arxiv-wiki.

sourceHugging Faceccupdated 3y agoView on Hugging Face
2likes14downloads
5 commits on main
01589613y ago

Upload genetics-arxiv-wiki.jsonl

as-cle-bert
49e9f613y ago

Delete genetics-arxiv-wiki.jsonl

as-cle-bert
b9bc5643y ago

Update README.md

as-cle-bert
b8794303y ago

Upload genetics-arxiv-wiki.jsonl

as-cle-bert
f56d8393y ago

initial commit

as-cle-bert