datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192
Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192.
The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 16 run(s). Each run can be found as a specific split in each configuration, the split… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192.details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192
Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192.
The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 14 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192.20260722_3D_test_epoch1details_Lansechen__Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384
Dataset Card for Evaluation run of Lansechen/Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384.
The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 9 run(s). Each run can be found as a specific split in each configuration, the split being named using… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384.mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_160k_ratiodetails_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192
Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192.
The dataset is composed of 2 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192.mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_multisample_2.5kdetails_Lansechen__Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192
Dataset Card for Evaluation run of Lansechen/Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192.
The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192.mp_mistral7bv3_sft_dpo_beta1e-1_epoch1_40k_n16mp_mistral7bv3_sft_dpo_beta2e-2_epoch1_20k_n8mpg27_mistral7bv3_sft_dpo_beta5e-2_epoch1_40k_multisample_ratiomp_mistral7bv3_sft_dpo_beta5e-2_epoch1_40kmp_mistral7bv3_sft_dpo_beta2e-2_epoch1_160k_ratiollama3_rewritert0_v3_10k_vsm0_epoch1-002mp_mistral7bv3_sft_dpo_beta1e-1_epoch1_20k_n8mp_mistral7bv3_sft_dpo_beta2e-2_epoch1_40k_multisample_n2_mp_mistral7bv3_sftmp_gemma9b_sft_dpo_beta2e-2_epoch1_10k_n8details_Lansechen__Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-8192
Dataset Card for Evaluation run of Lansechen/Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-8192
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-8192.
The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 9 run(s). Each run can be found as a specific split in each configuration, the split being named using… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-8192.mp_mistral7bv3_sft_dpo_beta2e-2_epoch1_40k_n16mpg27_mistral7bv3_sft_dpo_beta5e-2_epoch1_40k_n16mp_mistral7b_sft1_b512_lr5e-6_self-penalty-ogd_rms_epoch1_greedymp_gemma9b_sft_dpo_beta2e-2_epoch1_multisample_2.5kmp_gemma9b_sft_dpo_beta2e-2_epoch1_40k_multisample_n2mpg27_mistral7bv3_sft_dpo_beta2e-2_epoch1_multisample_2.5kot3_100k_ckpt-epoch1_eval_27e9
mlfoundations-dev/ot3_100k_ckpt-epoch1_eval_27e9
Precomputed model outputs for evaluation.
Evaluation Results
LiveCodeBench
Average Accuracy: 33.27% ± 0.97%
Number of Runs: 3
Run
Accuracy
Questions Solved
Total Questions
1
31.51%
161
511
2
33.46%
171
511
3
34.83%
178
511
mp_mistral7bv3_sft_dpo_beta1e-1_epoch1_160k_ratiomp_mistral7bv3_sft_dpo_beta1e-1_epoch1_logratio_annotmodeify_rubric_epoch1mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_40k_multisample_ratiodetails_Lansechen__Qwen2.5-1.5B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384
Dataset Card for Evaluation run of Lansechen/Qwen2.5-1.5B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384
Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-1.5B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384.
The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-1.5B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384.
