2imi9/Alpaca_ShareGPT_10G
Dataset Description This dataset consists of 10GB of open-source bilingual data (Chinese and English), sourced from platforms such as Hugging Face. The data covers a wide range of topics, with an emphasis on multi-round conversational logic and reasoning. It includes both general and technical question-answer pairs, making it ideal for training AI models that need to handle extended conversations and maintain context across multiple exchanges. The dataset is designed to improve… See the full description on the dataset page: https://huggingface.co/datasets/2imi9/Alpaca_ShareGPT_10G.
This repository belongs to 2imi9 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
