CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01EleutherAI /rpj-v2-sample-mixtral1M<n<10M0 likes983 downloads2y agoHugging Face02HPAI-BSC /MedQA-Mixtral-CoT Dataset Card for medqa-cot Synthetically enhanced responses to the medqa dataset using mixtral. Dataset Details Dataset Description To increase the quality of answers from the training splits of the MedQA dataset, we leverage Mixtral-8x7B to generate Chain of Thought(CoT) answers. We create a custom prompt for the dataset, along with a hand-crafted list of few-shot examples. For a multichoice answer, we ask the model to rephrase and explain the question… See the full description on the dataset page: https://huggingface.co/datasets/HPAI-BSC/MedQA-Mixtral-CoT.textmultiple-choice10K<n<100K9 likes819 downloads2y agoHugging Face03Kuperberg /mixtral-chunkstextn<1K0 likes589 downloads9mo agoHugging Face04reciprocate /tinygsm_mixtral_12Mtext10M<n<100M1 likes480 downloads2y agoHugging Face05open-llm-leaderboard-old /details_VAGOsolutions__SauerkrautLM-Mixtral-8x7B-Instruct Dataset Card for Evaluation run of VAGOsolutions/SauerkrautLM-Mixtral-8x7B-Instruct Dataset automatically created during the evaluation run of model VAGOsolutions/SauerkrautLM-Mixtral-8x7B-Instruct on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_VAGOsolutions__SauerkrautLM-Mixtral-8x7B-Instruct.0 likes379 downloads3y agoHugging Face06MoritzLaurer /synthetic_zeroshot_mixtral_v0.1tabular1M<n<10M9 likes363 downloads2y agoHugging Face07open-llm-leaderboard-old /details_hfl__chinese-mixtral Dataset Card for Evaluation run of hfl/chinese-mixtral Dataset automatically created during the evaluation run of model hfl/chinese-mixtral on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_hfl__chinese-mixtral.0 likes325 downloads3y agoHugging Face08mesolitica /mixtral-magicoder Mixtral Magicoder: Source Code Is All You Need on various programming languages We sampled programming languages from https://huggingface.co/datasets/bigcode/the-stack-dedup and pushed to https://huggingface.co/datasets/malaysia-ai/starcoderdata-sample After that, we use Magicoder: Source Code Is All You Need on various programming languages template, we target at least 10k rows for each programming languages. C++, 10747 rows C#, 10193 rows CUDA, 13843 rows Dockerfile, 13286 rows… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/mixtral-magicoder.text100K<n<1M4 likes318 downloads1y agoHugging Face09Wanfq /3_4_fusechat_v1_openchat-3.5_nh2-mixtral-8x7b-dpo_nh2-solar-10.7b_representation0 likes314 downloads2y agoHugging Face10Wanfq /1_4_fusechat_v1_openchat-3.5_nh2-mixtral-8x7b-dpo_nh2-solar-10.7b_representation0 likes311 downloads2y agoHugging Face11open-llm-leaderboard-old /details_mistralai__Mixtral-8x7B-v0.1 Dataset Card for Evaluation run of mistralai/Mixtral-8x7B-v0.1 Dataset automatically created during the evaluation run of model mistralai/Mixtral-8x7B-v0.1 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_mistralai__Mixtral-8x7B-v0.1.1 likes305 downloads3y agoHugging Face12open-llm-leaderboard-old /details_LeroyDyer__Mixtral_AI_Cyber_4.0 Dataset Card for Evaluation run of LeroyDyer/Mixtral_AI_Cyber_4.0 Dataset automatically created during the evaluation run of model LeroyDyer/Mixtral_AI_Cyber_4.0 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LeroyDyer__Mixtral_AI_Cyber_4.0.0 likes298 downloads2y agoHugging Face13HPAI-BSC /MedMCQA-Mixtral-CoT Dataset Card for medmcqa-cot Synthetically enhanced responses to the medmcqa dataset using mixtral. Dataset Details Dataset Description To increase the quality of answers from the training splits of the MedMCQA dataset, we leverage Mixtral-8x7B to generate Chain of Thought(CoT) answers. We create a custom prompt for the dataset, along with a hand-crafted list of few-shot examples. For a multichoice answer, we ask the model to rephrase and explain the… See the full description on the dataset page: https://huggingface.co/datasets/HPAI-BSC/MedMCQA-Mixtral-CoT.textquestion-answering100K<n<1M4 likes277 downloads2y agoHugging Face14open-llm-leaderboard-old /details_vistagi__Mixtral-8x7b-v0.1-sft Dataset Card for Evaluation run of vistagi/Mixtral-8x7b-v0.1-sft Dataset automatically created during the evaluation run of model vistagi/Mixtral-8x7b-v0.1-sft on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_vistagi__Mixtral-8x7b-v0.1-sft.0 likes254 downloads3y agoHugging Face15Xmaster6y /Mixtral-8x7B-v0.1-activationstext100K<n<1M0 likes230 downloads2y agoHugging Face16open-llm-leaderboard-old /details_Brillibits__Instruct_Mixtral-8x7B-v0.1_Dolly15K Dataset Card for Evaluation run of Brillibits/Instruct_Mixtral-8x7B-v0.1_Dolly15K Dataset automatically created during the evaluation run of model Brillibits/Instruct_Mixtral-8x7B-v0.1_Dolly15K on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Brillibits__Instruct_Mixtral-8x7B-v0.1_Dolly15K.0 likes228 downloads3y agoHugging Face17bowang0911 /alpaca-french-mixtral License & Attribution MTEB-format derivative of AIffl/Alpaca_french_mixtral (French Alpaca, Mixtral-translated). Query = instruction; corpus = answer. Deterministically subsampled to ~10k. Licensed under Apache-2.0 (same as source). tabulartext-retrieval10K<n<100K0 likes217 downloads3mo agoHugging Face18Wanfq /3_4_fusechat_v1_openchat-3.5_mixtral-8x7b-instruct-v0.1_solar-10.7b-instruct-v1.0_representationtabular10K<n<100K0 likes215 downloads2y agoHugging Face19open-llm-leaderboard-old /details_Swisslex__Mixtral-Orca-v0.1 Dataset Card for Evaluation run of Swisslex/Mixtral-Orca-v0.1 Dataset automatically created during the evaluation run of model Swisslex/Mixtral-Orca-v0.1 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Swisslex__Mixtral-Orca-v0.1.0 likes211 downloads3y agoHugging Face20open-llm-leaderboard-old /details_Swisslex__Mixtral-8x7b-DPO-v0.2 Dataset Card for Evaluation run of Swisslex/Mixtral-8x7b-DPO-v0.2 Dataset automatically created during the evaluation run of model Swisslex/Mixtral-8x7b-DPO-v0.2 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Swisslex__Mixtral-8x7b-DPO-v0.2.0 likes196 downloads3y agoHugging Face21deatos /mixtraltoken_fineweb_edu_mini_combinedtabular1M<n<10M0 likes193 downloads2y agoHugging Face22BEE-spoke-data /v3-mix-mixtraltext10M<n<100M1 likes188 downloads9mo agoHugging Face23mesolitica /mixtral-factual-QA Mixtral Factual QA Generate questions and answers based on context provided. We use contexts from, maktabahalbakri.com muftiwp.gov.my asklegal.my dewanbahasa-jdbp gov.my patriots rootofscience majalahsains nasilemaktech alhijrahnews https://huggingface.co/datasets/open-phi/textbooks notebooks at https://github.com/mesolitica/malaysian-dataset/tree/master/question-answer/mixtral-factual factually-wrong-qa-coding.jsonl, 31253 rows, 425 MB factually-wrong-qa.jsonl, 1108037 rows, 10… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/mixtral-factual-QA.textquestion-answering100K<n<1M4 likes185 downloads3y agoHugging Face24mesolitica /mixtral-malaysian-general-qa Mixtral Malaysian Chat Simulate conversation between a user and an assistant on various topics. Generated using Mixtral Instructions. Notebooks at https://github.com/mesolitica/malaysian-dataset/tree/master/chatbot/mixtral-malaysian-chat Multi-turn Bad things Multiturn of the user is saying bad things to the assistant. mixtral-conversation-badthings.jsonl, 57798 rows, 163 MB. Example data [{'role': 'user', 'content': "Hey bot, you're really dumb."… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/mixtral-malaysian-general-qa.1 likes184 downloads3y agoHugging Face25open-llm-leaderboard-old /details_cloudyu__Mixtral_7Bx2_MoE_13B Dataset Card for Evaluation run of cloudyu/Mixtral_7Bx2_MoE_13B Dataset automatically created during the evaluation run of model cloudyu/Mixtral_7Bx2_MoE_13B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__Mixtral_7Bx2_MoE_13B.0 likes183 downloads3y agoHugging Face26open-llm-leaderboard-old /details_cloudyu__Mixtral_7Bx5_MoE_30B Dataset Card for Evaluation run of cloudyu/Mixtral_7Bx5_MoE_30B Dataset automatically created during the evaluation run of model cloudyu/Mixtral_7Bx5_MoE_30B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__Mixtral_7Bx5_MoE_30B.0 likes183 downloads3y agoHugging Face27open-llm-leaderboard-old /details_cloudyu__Mixtral_34Bx2_MoE_60B Dataset Card for Evaluation run of cloudyu/Mixtral_34Bx2_MoE_60B Dataset automatically created during the evaluation run of model cloudyu/Mixtral_34Bx2_MoE_60B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__Mixtral_34Bx2_MoE_60B.1 likes174 downloads3y agoHugging Face28open-llm-leaderboard-old /details_Sao10K__Sensualize-Mixtral-bf16 Dataset Card for Evaluation run of Sao10K/Sensualize-Mixtral-bf16 Dataset automatically created during the evaluation run of model Sao10K/Sensualize-Mixtral-bf16 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Sao10K__Sensualize-Mixtral-bf16.0 likes174 downloads3y agoHugging Face29open-llm-leaderboard-old /details_Open-Orca__Mixtral-SlimOrca-8x7B Dataset Card for Evaluation run of Open-Orca/Mixtral-SlimOrca-8x7B Dataset automatically created during the evaluation run of model Open-Orca/Mixtral-SlimOrca-8x7B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Open-Orca__Mixtral-SlimOrca-8x7B.0 likes173 downloads3y agoHugging Face30open-llm-leaderboard-old /details_chargoddard__MixtralRPChat-ZLoss Dataset Card for Evaluation run of chargoddard/MixtralRPChat-ZLoss Dataset automatically created during the evaluation run of model chargoddard/MixtralRPChat-ZLoss on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_chargoddard__MixtralRPChat-ZLoss.0 likes173 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.