datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
code_contests_instruct
Dataset Card for "code_contests_instruct"
The deepmind/code_contests dataset formatted as markdown-instruct for text generation training.
There are several different configs. Look at them. Comments:
flesch_reading_ease is computed on the description col via textstat
hq means that python2 (aka PYTHON in language column) is dropped, and keeps only rows with flesch_reading_ease 75 or greater
min-cols drops all cols except language and text
possible values for language are {'CPP'… See the full description on the dataset page: https://huggingface.co/datasets/BEE-spoke-data/code_contests_instruct.code_contest_instruct_cpppython-code-instructions-18k-alpaca-standardized
Dataset Card for "python-code-instructions-18k-alpaca-standardized"
More Information needed
evol_instruct_code_filtered_39k
Dataset Card for "evol_instruct_code_filtered_38k"
Filtered version of nickrosh/Evol-Instruct-Code-80k-v1, with manual filtering, and automatic filtering based on quality and learning value classifiers.
instruction_filtering_askllm_seed_data_code_w_openthoughtsa1_code_star_coder_instruct_eval_636d
mlfoundations-dev/a1_code_star_coder_instruct_eval_636d
Precomputed model outputs for evaluation.
Evaluation Results
Summary
Metric
AIME24
AMC23
MATH500
MMLUPro
JEEBench
GPQADiamond
LiveCodeBench
CodeElo
CodeForces
Accuracy
15.0
51.0
72.2
28.2
34.2
35.9
29.0
6.9
5.3
AIME24
Average Accuracy: 15.00% ± 0.85%
Number of Runs: 10
Run
Accuracy
Questions Solved
Total Questions
1
13.33%
4
30
2
13.33%
4
30
3
16.67%
5
30
4… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/a1_code_star_coder_instruct_eval_636d.simct-author-code-10k-qwen25-7b-instruct
SimCT author-code baseline
Teacher Qwen2.5-7B-Instruct. 10000 raw prompts, 80000 candidates, 8705 author-selected targets. Author scripts pinned to cf0f33a0e6c967d4b74ea32b2dba12be01b73b9e.
This follows the released code, not a claim of exact paper replication or author data identity. Code responses receive format-only checks in the original verifier, not sandbox execution. Math uses the original custom checks. Selection may retain fewer than10000 prompts; no automatic… See the full description on the dataset page: https://huggingface.co/datasets/codemaivanngu/simct-author-code-10k-qwen25-7b-instruct.Evol-Instruct-Code-80k-v1-standardized
Dataset Card for "Evol-Instruct-Code-80k-v1-standardized"
More Information needed
code_instructions_standardized_cluster_19_std
Dataset Card for "code_instructions_standardized_cluster_19_std"
More Information needed
python-code-instructions-85k-mypo-qaqc
joshuasundance/python-code-instructions-85k-mypo QA/QC artifact
This dataset repo is a QA/QC derivative generated by myponline.
What is included
Root-level train.parquet / validation.parquet / test.parquet with full QA/QC annotations.
filtered_basic/ with rows that pass structural QA/QC checks.
filtered_strict/ with rows whose chosen side passes structural QA/QC plus standalone ruff and mypy --strict.
summary.json with aggregate counts and provenance.… See the full description on the dataset page: https://huggingface.co/datasets/joshuasundance/python-code-instructions-85k-mypo-qaqc.python_code_instructions_18k_alpaca-standardized
Dataset Card for "python_code_instructions_18k_alpaca-standardized"
More Information needed
code_instructions_filtered_7k
Dataset Card for "code_instructions_filtered_7k"
Filtered version of sahil2801/code_instructions_120k based on manual, quality, and learning value filters.
code_instructions_standardized_cluster_10_std
Dataset Card for "code_instructions_standardized_cluster_10_std"
More Information needed
python-code-instructions-18k-alpaca-standardized_cluster_2
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_2"
More Information needed
code_instructions_standardized_cluster_11_std
Dataset Card for "code_instructions_standardized_cluster_11_std"
More Information needed
code_instructions_standardized
Dataset Card for "code_instructions_standardized"
More Information needed
instruction_filtering_gemini_length_codecode-alpaca-eval-v0-deepseek-coder-7b-instruct-v1.5-annotationscode_instructions_122k_alpaca_style_standardizedcode_instructions_standardized_cluster_15_std
Dataset Card for "code_instructions_standardized_cluster_15_std"
More Information needed
python-code-instructions-18k-alpaca-standardized_cluster_2_std
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_2_std"
More Information needed
python-code-instructions-18k-alpaca-standardized_cluster_7_std
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_7_std"
More Information needed
instruction_filtering_scale_up_code_base_embedding_filter_meancode_instructions_122k_alpaca_style_standardizedpython-code-instructions-18k-alpaca-standardized_cluster_5_std
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_5_std"
More Information needed
python-code-instructions-18k-alpaca-standardized_cluster_8
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_8"
More Information needed
RM-Bench-code-Mistral-7B-Instruct-v0.1-scorespython-code-instructions-18k-alpaca-standardized_cluster_3_std
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_3_std"
More Information needed
python-code-instructions-18k-alpaca-standardized_cluster_9
Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_9"
More Information needed
code_instructions_standardized_cluster_1_std
Dataset Card for "code_instructions_standardized_cluster_1_std"
More Information needed
