datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
toy-models-of-sft-data
Toy Models of SFT Data
This is a public-clean candidate data package for the Toy Models of SFT project.
It is built for researcher inspection first.
The package answers two questions:
What were the models trained on?
How did the models actually behave under evaluation?
The package includes training data, eval inputs, model rollouts, judge scores,
parsed GPQA outputs, aggregate tables, paper figures, frozen plot data, and
provenance records. It deliberately includes some… See the full description on the dataset page: https://huggingface.co/datasets/matonski/toy-models-of-sft-data.coderTraining dataset for finetuning for human-eval.
This dataset has been created from the following datasets:
sahil2801/CodeAlpaca-20k
sahil2801/code_instructions_120k
mhhmm/leetcode-solutions-python
teknium1/GPTeacher
Script for generating dataset: create_dataset.py.
my-first-dataset
My First Dataset
panduan-fengxian-matoumatouLeLoup__ECE-PRYMMAL-0.5B-FT-EnhancedMUSREnsembleV3-details
Dataset Card for Evaluation run of matouLeLoup/ECE-PRYMMAL-0.5B-FT-EnhancedMUSREnsembleV3
Dataset automatically created during the evaluation run of model matouLeLoup/ECE-PRYMMAL-0.5B-FT-EnhancedMUSREnsembleV3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/matouLeLoup__ECE-PRYMMAL-0.5B-FT-EnhancedMUSREnsembleV3-details.matouLeLoup__ECE-PRYMMAL-0.5B-FT-V5-MUSR-Mathis-details
Dataset Card for Evaluation run of matouLeLoup/ECE-PRYMMAL-0.5B-FT-V5-MUSR-Mathis
Dataset automatically created during the evaluation run of model matouLeLoup/ECE-PRYMMAL-0.5B-FT-V5-MUSR-Mathis
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/matouLeLoup__ECE-PRYMMAL-0.5B-FT-V5-MUSR-Mathis-details.matouLeLoup__ECE-PRYMMAL-0.5B-FT-V4-MUSR-Mathis-details
Dataset Card for Evaluation run of matouLeLoup/ECE-PRYMMAL-0.5B-FT-V4-MUSR-Mathis
Dataset automatically created during the evaluation run of model matouLeLoup/ECE-PRYMMAL-0.5B-FT-V4-MUSR-Mathis
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/matouLeLoup__ECE-PRYMMAL-0.5B-FT-V4-MUSR-Mathis-details.matouLeLoup__ECE-PRYMMAL-0.5B-FT-V4-MUSR-ENSEMBLE-Mathis-details
Dataset Card for Evaluation run of matouLeLoup/ECE-PRYMMAL-0.5B-FT-V4-MUSR-ENSEMBLE-Mathis
Dataset automatically created during the evaluation run of model matouLeLoup/ECE-PRYMMAL-0.5B-FT-V4-MUSR-ENSEMBLE-Mathis
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/matouLeLoup__ECE-PRYMMAL-0.5B-FT-V4-MUSR-ENSEMBLE-Mathis-details.matoverflow_scrape_1matouLeLoup__ECE-PRYMMAL-0.5B-FT-MUSR-ENSEMBLE-V2Mathis-details
Dataset Card for Evaluation run of matouLeLoup/ECE-PRYMMAL-0.5B-FT-MUSR-ENSEMBLE-V2Mathis
Dataset automatically created during the evaluation run of model matouLeLoup/ECE-PRYMMAL-0.5B-FT-MUSR-ENSEMBLE-V2Mathis
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/matouLeLoup__ECE-PRYMMAL-0.5B-FT-MUSR-ENSEMBLE-V2Mathis-details.SexBreak
