CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Magpie-Align /Magpie-Qwen2.5-Coder-Pro-300K-v0.1 Project Web: https://magpie-align.github.io/ Arxiv Technical Report: https://arxiv.org/abs/2406.08464 Codes: https://github.com/magpie-align/magpie Abstract Click Here High-quality instruction data is critical for aligning large language models (LLMs). Although some models, such as Llama-3-Instruct, have open weights, their alignment data remain private, which hinders the democratization of AI. High human labor costs and a limited, predefined scope for prompting prevent… See the full description on the dataset page: https://huggingface.co/datasets/Magpie-Align/Magpie-Qwen2.5-Coder-Pro-300K-v0.1.tabular100K<n<1M9 likes245 downloads2y agoHugging Face02OALL /details_Qwen__Qwen2.5-Coder-14B-Instruct Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-14B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-14B-Instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Qwen__Qwen2.5-Coder-14B-Instruct.tabular100K<n<1M0 likes188 downloads2y agoHugging Face03OALL /details_Qwen__Qwen2.5-Coder-7B-Instruct Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-7B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-7B-Instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Qwen__Qwen2.5-Coder-7B-Instruct.tabular100K<n<1M0 likes110 downloads2y agoHugging Face04bbidpa /Qwen2.5-Coder-0.5B-Flutter-steps-eval Qwen2.5-Coder-0.5B Flutter — Steps Mode — Validation Results Dataset Summary Held-out evaluation results for bbidpa/Qwen2.5-Coder-0.5B-Flutter-steps, a fine-tune of Qwen2.5-Coder-0.5B for editing Flutter/Dart source files. In steps mode, the model is given an existing file and an edit instruction and generates a sequence of localized search/replace edit actions, each mechanically applied to the current file state before the next action is generated, until the… See the full description on the dataset page: https://huggingface.co/datasets/bbidpa/Qwen2.5-Coder-0.5B-Flutter-steps-eval.tabulartext-generation1K<n<10K0 likes102 downloads14d agoHugging Face05bbidpa /Qwen2.5-Coder-0.5B-Flutter-direct-eval Qwen2.5-Coder-0.5B Flutter — Direct Mode — Validation Results Dataset Summary Held-out evaluation results for bbidpa/Qwen2.5-Coder-0.5B-Flutter-direct, a fine-tune of Qwen2.5-Coder-0.5B for editing Flutter/Dart source files. In direct mode, the model is given an existing file and an edit instruction and generates the complete modified file in a single forward pass (as opposed to the steps / iterative diff-based mode — see the sibling dataset… See the full description on the dataset page: https://huggingface.co/datasets/bbidpa/Qwen2.5-Coder-0.5B-Flutter-direct-eval.tabulartext-generation1K<n<10K0 likes84 downloads14d agoHugging Face06testcase-evaluate /all-do-Qwen2.5-Coder-32B-Instruct0 likes52 downloads1y agoHugging Face07OALL /details_Qwen__Qwen2.5-Coder-14B Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-14B Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-14B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Qwen__Qwen2.5-Coder-14B.tabular100K<n<1M0 likes50 downloads2y agoHugging Face08open-llm-leaderboard /theo77186__Qwen2.5-Coder-7B-Instruct-20241106-detailsgated Dataset Card for Evaluation run of theo77186/Qwen2.5-Coder-7B-Instruct-20241106 Dataset automatically created during the evaluation run of model theo77186/Qwen2.5-Coder-7B-Instruct-20241106 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/theo77186__Qwen2.5-Coder-7B-Instruct-20241106-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face09lewtun /details_Qwen__Qwen2.5-Coder-3B-Instruct Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-3B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-3B-Instruct. The dataset is composed of 1 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/lewtun/details_Qwen__Qwen2.5-Coder-3B-Instruct.textn<1K0 likes39 downloads2y agoHugging Face10open-llm-leaderboard /Qwen__Qwen2.5-Coder-7B-Instruct-detailsgated Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-7B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-7B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__Qwen2.5-Coder-7B-Instruct-details.tabular10K<n<100K0 likes35 downloads2y agoHugging Face11open-llm-leaderboard /Etherll__Qwen2.5-Coder-7B-Instruct-Ties-detailsgated Dataset Card for Evaluation run of Etherll/Qwen2.5-Coder-7B-Instruct-Ties Dataset automatically created during the evaluation run of model Etherll/Qwen2.5-Coder-7B-Instruct-Ties The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Etherll__Qwen2.5-Coder-7B-Instruct-Ties-details.tabular10K<n<100K0 likes35 downloads2y agoHugging Face12davidberenstein1957 /qwen2.5-coder-0.5b-openai_humaneval Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/qwen2.5-coder-0.5b-openai_humaneval.tabularn<1K0 likes34 downloads2y agoHugging Face13TQRG /DeltaSecommits_qwen2.5-coder-14b-instruct_tokenized_v2_vulnerable1K<n<10K0 likes34 downloads6mo agoHugging Face14mizinovmv /ru_codefeedback_python_Qwen2.5-Coder-32B-Instruct-GPTQ-Int8_sample ru_Code-Feedback Вопросы python Code-Feedback Решение и unit-test с результатами python исполнения. Made with Qwen2.5-Coder-32B-Instruct-GPTQ-Int8 ru_eval_status count OK 2554 Exception 2337 SyntaxError 518 Timeout 79 textquestion-answering1K<n<10K4 likes32 downloads2y agoHugging Face15ioi-leaderboard /ioi-eval-sglang_Qwen_Qwen2.5-Coder-32B-Instruct-prompt-mem-limit-fixtext1K<n<10K0 likes29 downloads2y agoHugging Face16smoorsmith /humaneval___2txt___Qwen2.5-Math-7B-Instruct___Qwen2.5-Coder-7B-Instruct___from_evalplustextn<1K0 likes29 downloads1y agoHugging Face17open-llm-leaderboard /TIGER-Lab__AceCoder-Qwen2.5-Coder-7B-Ins-Rule-detailsgated Dataset Card for Evaluation run of TIGER-Lab/AceCoder-Qwen2.5-Coder-7B-Ins-Rule Dataset automatically created during the evaluation run of model TIGER-Lab/AceCoder-Qwen2.5-Coder-7B-Ins-Rule The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/TIGER-Lab__AceCoder-Qwen2.5-Coder-7B-Ins-Rule-details.tabular10K<n<100K0 likes28 downloads2y agoHugging Face18smoorsmith /gsm8k___2txt___Qwen2.5_Math_7B_Instruct___Qwen2.5_Coder_7B_Instructtabular1K<n<10K0 likes28 downloads1y agoHugging Face19open-llm-leaderboard /Qwen__Qwen2.5-Coder-14B-Instruct-detailsgated Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-14B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-14B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__Qwen2.5-Coder-14B-Instruct-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face20open-llm-leaderboard /Qwen__Qwen2.5-Coder-32B-Instruct-detailsgated Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-32B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-32B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__Qwen2.5-Coder-32B-Instruct-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face21lemon-mint /Magpie-Qwen2.5-Coder-Pro-300K-Query-Positive-Pairtext10K<n<100K0 likes26 downloads2y agoHugging Face22open-llm-leaderboard /TIGER-Lab__AceCoder-Qwen2.5-Coder-7B-Base-Rule-detailsgated Dataset Card for Evaluation run of TIGER-Lab/AceCoder-Qwen2.5-Coder-7B-Base-Rule Dataset automatically created during the evaluation run of model TIGER-Lab/AceCoder-Qwen2.5-Coder-7B-Base-Rule The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/TIGER-Lab__AceCoder-Qwen2.5-Coder-7B-Base-Rule-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face23NoahShen /qwen2.5-coder-7b-inst-sleeper-agent-2024-completionstext1K<n<10K0 likes26 downloads1y agoHugging Face24open-llm-leaderboard /Qwen__Qwen2.5-Coder-14B-detailsgated Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-14B Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-14B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__Qwen2.5-Coder-14B-details.tabular10K<n<100K1 likes25 downloads2y agoHugging Face25test-gen /code_mbpp_Qwen2.5-Coder-0.5B-Instruct_temp0.1_num8_tests_mbpp_qwen-7b-easy_t0.0_n1textn<1K0 likes24 downloads1y agoHugging Face26SecCoderX /SecCoderX_Qwen2.5_Coder_7B_GRPO_dataset Citation If you find our work helpful, feel free to give us a cite. @misc{wu2026securecodegenerationonline, title={Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model}, author={Tianyi Wu and Mingzhe Du and Yue Liu and Chengran Yang and Terry Yue Zhuo and Jiaheng Zhang and See-Kiong Ng}, year={2026}, eprint={2602.07422}, archivePrefix={arXiv}, primaryClass={cs.CR}, url={https://arxiv.org/abs/2602.07422}… See the full description on the dataset page: https://huggingface.co/datasets/SecCoderX/SecCoderX_Qwen2.5_Coder_7B_GRPO_dataset.text10K<n<100K0 likes24 downloads7mo agoHugging Face27DCAgent2 /DCAgent2_terminal_bench_2_Qwen_Qwen2.5-Coder-32B-Instruct-tracestext1K<n<10K0 likes24 downloads6mo agoHugging Face28open-llm-leaderboard /Qwen__Qwen2.5-Coder-32B-detailsgated Dataset Card for Evaluation run of Qwen/Qwen2.5-Coder-32B Dataset automatically created during the evaluation run of model Qwen/Qwen2.5-Coder-32B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__Qwen2.5-Coder-32B-details.tabular10K<n<100K0 likes23 downloads2y agoHugging Face29test-gen /mbpp_Qwen2.5-Coder-14B-Instruct_t1.0_n8_generated_codetextn<1K0 likes23 downloads1y agoHugging Face30smoorsmith /humaneval___2txt___Qwen2.5_Math_7B_Instruct___Qwen2.5_Coder_7B_Instructtextn<1K0 likes23 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.