datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
JP-AlpaCare-MedInstruct-52k
JP-AlpaCare-MedInstruct-52k
This dataset is a Japanese-translated and aligned version of AlpaCare-MedInstruct-52k.
The translation was performed automatically using gpt-4o-2024-05-13, preserving alignment between English and Japanese instructions, inputs, and outputs. Total data size is 51992.
Dataset Details
Original Dataset: AlpaCare-MedInstruct-52k
Translation Model: GPT-4o (gpt-4o-2024-05-13)
Fields:
id (ID)
instruction_ja, input_ja, output_ja (Japanese)
id_en… See the full description on the dataset page: https://huggingface.co/datasets/li-lab/JP-AlpaCare-MedInstruct-52k.godlikehhd__alpaca_data_score_max_0.1_2600-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_score_max_0.1_2600
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_score_max_0.1_2600
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_score_max_0.1_2600-details.alpaca-cleaned-gemini-hun-ratingsEz az adathalmaz úgy keletkezett, hogy a Bazsalanszky/alpaca-cleaned-gemini-hun-n lefuttattam egy llm által támogatott értékelést.
Az értékelő modell a gemini-pro (az ingyenes) volt. A használt kód az alpagasus módosítása: https://github.com/boapps/alpagasus-hu
EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details
Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details.EpistemeAI__Alpaca-Llama3.1-8B-details
Dataset Card for Evaluation run of EpistemeAI/Alpaca-Llama3.1-8B
Dataset automatically created during the evaluation run of model EpistemeAI/Alpaca-Llama3.1-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Alpaca-Llama3.1-8B-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.06-8B-Philos-dpo-details.EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200-details
Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200-details.EpistemeAI2__Fireball-Alpaca-Llama3.1-8B-Philos-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1-8B-Philos
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1-8B-Philos
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1-8B-Philos-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-details.EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2-details
Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R2-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.03-8B-Philos-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.03-8B-Philos
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.03-8B-Philos
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.03-8B-Philos-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.04-8B-Philos-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.04-8B-Philos
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.04-8B-Philos
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.04-8B-Philos-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection-details.EpistemeAI__Athene-codegemma-2-7b-it-alpaca-v1.3-details
Dataset Card for Evaluation run of EpistemeAI/Athene-codegemma-2-7b-it-alpaca-v1.3
Dataset automatically created during the evaluation run of model EpistemeAI/Athene-codegemma-2-7b-it-alpaca-v1.3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Athene-codegemma-2-7b-it-alpaca-v1.3-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-details.arabic_alpaca_modelEpistemeAI2__Athene-codegemma-2-7b-it-alpaca-v1.2-details
Dataset Card for Evaluation run of EpistemeAI2/Athene-codegemma-2-7b-it-alpaca-v1.2
Dataset automatically created during the evaluation run of model EpistemeAI2/Athene-codegemma-2-7b-it-alpaca-v1.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Athene-codegemma-2-7b-it-alpaca-v1.2-details.EpistemeAI2__Fireball-Alpaca-Llama3.1.01-8B-Philos-details
Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.01-8B-Philos
Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.01-8B-Philos
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.01-8B-Philos-details.alpaca_hu_2k-ratingsEz az adathalmaz úgy keletkezett, hogy az NYTK/alpaca_hu_2k-n lefuttattam egy llm által támogatott értékelést.
Az értékelő modell a gemini-pro (az ingyenes) volt. A használt kód az alpagasus módosítása: https://github.com/boapps/alpagasus-hu
alpaca-prompts-annotated
Alpaca Annotated Dataset
This dataset includes prompts taken from yahma/alpaca-cleaned that have been annotated using the nvidia/prompt-task-and-complexity-classifier.
Each entry separates the instruction and input fields with two newline characters (\n\n).
The annotations describe the type of task and its complexity, as determined by NVIDIA’s classifier.
To know more about what each annotation means, see the classifier’s page on Hugging Face.
The prompts have been randomly… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/alpaca-prompts-annotated.alpaca-hu-v2-ratingsEz az adathalmaz úgy keletkezett, hogy az alpaca-hu-v2-n lefuttattam egy llm által támogatott értékelést.
Ez első ránézésre elég királyul kiszűri (0-ás ratinget ad) a Gemini random halandzsáira.
Az értékelő modell a gemini-pro (az ingyenes) volt. A használt kód az alpagasus módosítása: https://github.com/boapps/alpagasus-hu
godlikehhd__alpaca_data_ifd_max_2600_3B-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_ifd_max_2600_3B
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_ifd_max_2600_3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_ifd_max_2600_3B-details.godlikehhd__alpaca_data_ifd_max_2600-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_ifd_max_2600
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_ifd_max_2600
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_ifd_max_2600-details.alpaca_hu_mt-ratingsgodlikehhd__alpaca_data_sampled_ifd_5200-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_sampled_ifd_5200
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_sampled_ifd_5200
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_sampled_ifd_5200-details.godlikehhd__alpaca_data_ins_max_5200-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_ins_max_5200
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_ins_max_5200
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_ins_max_5200-details.godlikehhd__alpaca_data_ifd_min_2600-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_ifd_min_2600
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_ifd_min_2600
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_ifd_min_2600-details.SaisExperiments__Not-So-Small-Alpaca-24B-details
Dataset Card for Evaluation run of SaisExperiments/Not-So-Small-Alpaca-24B
Dataset automatically created during the evaluation run of model SaisExperiments/Not-So-Small-Alpaca-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SaisExperiments__Not-So-Small-Alpaca-24B-details.godlikehhd__alpaca_data_full_2-details
Dataset Card for Evaluation run of godlikehhd/alpaca_data_full_2
Dataset automatically created during the evaluation run of model godlikehhd/alpaca_data_full_2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/godlikehhd__alpaca_data_full_2-details.ADG-Qwen2.5-Alpaca-GPT4
