2imi9/Alpaca_ShareGPT_10G
Dataset Description This dataset consists of 10GB of open-source bilingual data (Chinese and English), sourced from platforms such as Hugging Face. The data covers a wide range of topics, with an emphasis on multi-round conversational logic and reasoning. It includes both general and technical question-answer pairs, making it ideal for training AI models that need to handle extended conversations and maintain context across multiple exchanges. The dataset is designed to improve… See the full description on the dataset page: https://huggingface.co/datasets/2imi9/Alpaca_ShareGPT_10G.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face