CoolFace
Datasetpublic

qingdu-giter/HuatuoGPT2-Pretraining-Instruction

HuatuoGPT2-Pretraining-Instruction-5200K Here are the pre-training instructions for HuatuoGPT-II, developed with 5.2 million medical corpus using ChatGPT. This dataset is used to incorporate extensive medical knowledge and enable a one-stage medical adaptation. All our data have been made publicly accessible. Data Volume The following table details the volume and distribution of pre-training data for HuatuoGPT2: Data Source Data Volume… See the full description on the dataset page: https://huggingface.co/datasets/qingdu-giter/HuatuoGPT2-Pretraining-Instruction.

sourceHugging Faceapache-2.0updated 6mo agoView on Hugging Face
0likes13downloads
1 commits on main
4ebee346mo ago

Duplicate from FreedomIntelligence/HuatuoGPT2-Pretraining-Instruction

qingdu-giter, jymcc