mistral-nemo
m_low_res_sae_wiki_big_tokenized_Mistral-Nemo-Base-2407mistral-nvidia-Llama-Nemotron-Post-Training-Dataset-sftmistral_nemo_base-mmlu-valdetails_cognitivecomputations__dolphin-2.9.3-mistral-nemo-12b
Dataset Card for Evaluation run of cognitivecomputations/dolphin-2.9.3-mistral-nemo-12b
Dataset automatically created during the evaluation run of model cognitivecomputations/dolphin-2.9.3-mistral-nemo-12b.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_cognitivecomputations__dolphin-2.9.3-mistral-nemo-12b.output_Mistral-Nemo-Base-2407-simpleqa-0_1000-m_generation-n_32-t_1.0-k_40-p_0.9-l_128based-chat-v0.1-Mistral-Nemo-Base-2407
Based-Chat v0.1 (Mistral Nemo Base 2407)
This dataset was developed as part of an exploration into understanding the necessity of supervised datasets for fine-tuning base LLMs into conversational models.
It's a synthetic dataset created with Mistral-Nemo-Base-2407, and used to fine-tune that model, producing relay-v0.1-Mistral-Nemo-2407.
Methodology
This synthetic dataset is generated using the following as conversation starters:
facebook/empathetic_dialogues… See the full description on the dataset page: https://huggingface.co/datasets/danlou/based-chat-v0.1-Mistral-Nemo-Base-2407.
