prodigy
Datasets
All datasets matching “prodigy”details_ChaoticNeutrals__Prodigy_7B
Dataset Card for Evaluation run of ChaoticNeutrals/Prodigy_7B
Dataset automatically created during the evaluation run of model ChaoticNeutrals/Prodigy_7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ChaoticNeutrals__Prodigy_7B.PRODIGY-LAB_SARA
Dataset Card for PRODIGY-LAB_CLEANED
Repository: https://github.com/aadhithyaravi
Created by: Aadhithya
Contact: aadhithyaxll@gmail.com
Instagram: @aadhi.arc
LinkedIn: www.linkedin.com/in/aadhithya-ravi-135019289
Dataset Description
PRODIGY-SARA-MODEL is a refined and enhanced dataset designed for instruction-based fine-tuning of large language models (LLMs).It combines multiple high-quality sources, including cleaned and normalized instructions, to improve… See the full description on the dataset page: https://huggingface.co/datasets/Apex-X/PRODIGY-LAB_SARA.bunnycore__Llama-3.2-3B-ProdigyPlusPlus-details
Dataset Card for Evaluation run of bunnycore/Llama-3.2-3B-ProdigyPlusPlus
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.2-3B-ProdigyPlusPlus
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.2-3B-ProdigyPlusPlus-details.bunnycore__Llama-3.2-3B-ProdigyPlus-details
Dataset Card for Evaluation run of bunnycore/Llama-3.2-3B-ProdigyPlus
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.2-3B-ProdigyPlus
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.2-3B-ProdigyPlus-details.prodigyprodigy-cleaned
Dataset Card for Alpaca-Cleaned
Repository: https://github.com/gururise/AlpacaDataCleaned
Dataset Description
This is a cleaned version of the original Alpaca Dataset released by Stanford. The following issues have been identified in the original release and fixed in this dataset:
Hallucinations: Many instructions in the original dataset had instructions referencing data on the internet, which just caused GPT3 to hallucinate an answer.
"instruction":"Summarize the… See the full description on the dataset page: https://huggingface.co/datasets/Apex-X/prodigy-cleaned.
