datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vi-self-chat-sharegpt-format
🇻🇳 Vietnamese Self-Chat Dataset
This dataset is designed to enhance the model's ability to engage in multi-turn conversations with humans.
To construct this dataset, we follow a two-step process:
Step 1: Instruction Generation
We employ the methodology outlined in the Self-Instruct paper to craft a diverse set of instructions. This paper serves as a guide for aligning pretrained language models with specific instructions, providing a structured foundation for subsequent dialogue… See the full description on the dataset page: https://huggingface.co/datasets/bkai-foundation-models/vi-self-chat-sharegpt-format.Self-J-score-wo-ref-skywork-pref-model-yi-1.5-16k-chat-thre-1-10000Self-J-score-wo-ref-skywork-pref-model-yi-1.5-16k-chat-thre-1Self-J-score-w-ref-skywork-pref-ref-lla31-70b-inst-model-yi-1.5-16k-chat-thre-1mlx-chat-modelsllama2-chat-raft-software-life-cycle-modelsquotaclimat-model-finetune-desinformation-chatSelf-J-score-w-ref-skywork-pref-ref-lla31-70b-inst-model-yi-1.5-16k-chat-thre-1-10000PythonForllama-2-7b-chat-hf-modelgenshin-impact-role-chat-model
