CoolFace
Datasetpublic

timaeus/pile-pubmed_central

Dataset Creation Process These subsets were created by streaming over the rows from monology/pile-uncopyrighted and filtering by the meta column. Each subset is generally limited to the first 100,000 qualifying rows encountered.

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes70downloads

timaeus/pile-pubmed_central · main · files are served by the source, never re-hosted here