CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01BEE-spoke-data /code_contests_instruct Dataset Card for "code_contests_instruct" The deepmind/code_contests dataset formatted as markdown-instruct for text generation training. There are several different configs. Look at them. Comments: flesch_reading_ease is computed on the description col via textstat hq means that python2 (aka PYTHON in language column) is dropped, and keeps only rows with flesch_reading_ease 75 or greater min-cols drops all cols except language and text possible values for language are {'CPP'… See the full description on the dataset page: https://huggingface.co/datasets/BEE-spoke-data/code_contests_instruct.tabulartext-generation10M<n<100M7 likes1.4k downloads9mo agoHugging Face02wttw /code_contest_instruct_cpptabulartext-generation1M<n<10M3 likes255 downloads2y agoHugging Face03HydraLM /python-code-instructions-18k-alpaca-standardized Dataset Card for "python-code-instructions-18k-alpaca-standardized" More Information needed tabular10K<n<100K1 likes123 downloads3y agoHugging Face04vikp /evol_instruct_code_filtered_39k Dataset Card for "evol_instruct_code_filtered_38k" Filtered version of nickrosh/Evol-Instruct-Code-80k-v1, with manual filtering, and automatic filtering based on quality and learning value classifiers. tabular10K<n<100K3 likes119 downloads3y agoHugging Face05mlfoundations-dev /instruction_filtering_askllm_seed_data_code_w_openthoughtstabular100K<n<1M0 likes73 downloads2y agoHugging Face06mlfoundations-dev /a1_code_star_coder_instruct_eval_636d mlfoundations-dev/a1_code_star_coder_instruct_eval_636d Precomputed model outputs for evaluation. Evaluation Results Summary Metric AIME24 AMC23 MATH500 MMLUPro JEEBench GPQADiamond LiveCodeBench CodeElo CodeForces Accuracy 15.0 51.0 72.2 28.2 34.2 35.9 29.0 6.9 5.3 AIME24 Average Accuracy: 15.00% ± 0.85% Number of Runs: 10 Run Accuracy Questions Solved Total Questions 1 13.33% 4 30 2 13.33% 4 30 3 16.67% 5 30 4… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/a1_code_star_coder_instruct_eval_636d.tabular1K<n<10K1 likes55 downloads1y agoHugging Face07codemaivanngu /simct-author-code-10k-qwen25-7b-instruct SimCT author-code baseline Teacher Qwen2.5-7B-Instruct. 10000 raw prompts, 80000 candidates, 8705 author-selected targets. Author scripts pinned to cf0f33a0e6c967d4b74ea32b2dba12be01b73b9e. This follows the released code, not a claim of exact paper replication or author data identity. Code responses receive format-only checks in the original verifier, not sandbox execution. Math uses the original custom checks. Selection may retain fewer than10000 prompts; no automatic… See the full description on the dataset page: https://huggingface.co/datasets/codemaivanngu/simct-author-code-10k-qwen25-7b-instruct.tabular10K<n<100K0 likes43 downloads15d agoHugging Face08HydraLM /Evol-Instruct-Code-80k-v1-standardized Dataset Card for "Evol-Instruct-Code-80k-v1-standardized" More Information needed tabular100K<n<1M2 likes41 downloads3y agoHugging Face09AdapterOcean /code_instructions_standardized_cluster_19_std Dataset Card for "code_instructions_standardized_cluster_19_std" More Information needed tabular10K<n<100K0 likes33 downloads3y agoHugging Face10joshuasundance /python-code-instructions-85k-mypo-qaqc joshuasundance/python-code-instructions-85k-mypo QA/QC artifact This dataset repo is a QA/QC derivative generated by myponline. What is included Root-level train.parquet / validation.parquet / test.parquet with full QA/QC annotations. filtered_basic/ with rows that pass structural QA/QC checks. filtered_strict/ with rows whose chosen side passes structural QA/QC plus standalone ruff and mypy --strict. summary.json with aggregate counts and provenance.… See the full description on the dataset page: https://huggingface.co/datasets/joshuasundance/python-code-instructions-85k-mypo-qaqc.tabular10K<n<100K0 likes29 downloads5mo agoHugging Face11khalidalt /python_code_instructions_18k_alpaca-standardized Dataset Card for "python_code_instructions_18k_alpaca-standardized" More Information needed tabular10K<n<100K1 likes27 downloads3y agoHugging Face12vikp /code_instructions_filtered_7k Dataset Card for "code_instructions_filtered_7k" Filtered version of sahil2801/code_instructions_120k based on manual, quality, and learning value filters. tabular1K<n<10K3 likes25 downloads3y agoHugging Face13AdapterOcean /code_instructions_standardized_cluster_10_std Dataset Card for "code_instructions_standardized_cluster_10_std" More Information needed tabular10K<n<100K0 likes23 downloads3y agoHugging Face14AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_2 Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_2" More Information needed tabular1K<n<10K1 likes20 downloads3y agoHugging Face15AdapterOcean /code_instructions_standardized_cluster_11_std Dataset Card for "code_instructions_standardized_cluster_11_std" More Information needed tabular10K<n<100K0 likes20 downloads3y agoHugging Face16HydraLM /code_instructions_standardized Dataset Card for "code_instructions_standardized" More Information needed tabular100K<n<1M1 likes18 downloads3y agoHugging Face17mlfoundations-dev /instruction_filtering_gemini_length_codetabular1K<n<10K0 likes18 downloads2y agoHugging Face18AlekseyKorshuk /code-alpaca-eval-v0-deepseek-coder-7b-instruct-v1.5-annotationstabularn<1K0 likes16 downloads2y agoHugging Face19kranthigv /code_instructions_122k_alpaca_style_standardizedtabular100K<n<1M0 likes15 downloads3y agoHugging Face20AdapterOcean /code_instructions_standardized_cluster_15_std Dataset Card for "code_instructions_standardized_cluster_15_std" More Information needed tabular1K<n<10K0 likes15 downloads3y agoHugging Face21AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_2_std Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_2_std" More Information needed tabular1K<n<10K1 likes14 downloads3y agoHugging Face22AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_7_std Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_7_std" More Information needed tabular1K<n<10K0 likes14 downloads3y agoHugging Face23mlfoundations-dev /instruction_filtering_scale_up_code_base_embedding_filter_meantabular10K<n<100K0 likes14 downloads2y agoHugging Face24HydraLM /code_instructions_122k_alpaca_style_standardizedtabular100K<n<1M1 likes13 downloads3y agoHugging Face25AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_5_std Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_5_std" More Information needed tabular1K<n<10K0 likes13 downloads3y agoHugging Face26AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_8 Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_8" More Information needed tabular1K<n<10K0 likes13 downloads3y agoHugging Face27Ayush-Singh /RM-Bench-code-Mistral-7B-Instruct-v0.1-scorestabularn<1K0 likes13 downloads2y agoHugging Face28AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_3_std Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_3_std" More Information needed tabular1K<n<10K0 likes12 downloads3y agoHugging Face29AdapterOcean /python-code-instructions-18k-alpaca-standardized_cluster_9 Dataset Card for "python-code-instructions-18k-alpaca-standardized_cluster_9" More Information needed tabular1K<n<10K0 likes12 downloads3y agoHugging Face30AdapterOcean /code_instructions_standardized_cluster_1_std Dataset Card for "code_instructions_standardized_cluster_1_std" More Information needed tabular10K<n<100K0 likes12 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.