CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Infatoshi /kernelbench-mega-traces KernelBench-Mega agent traces Coding agents writing full GPU megakernels across Blackwell / H100 / B200, scored as speedup over reference; contamination-audited (23 verified cells). Each .jsonl file is one agent run in Claude-Code session format, viewable with the agent trace viewer. Filename = run id; manifest.csv maps each run to model / harness / problem / GPU / score. 23 agent traces · live leaderboard: https://kernelbench.com/mega Secrets redacted. Full reasoning for… See the full description on the dataset page: https://huggingface.co/datasets/Infatoshi/kernelbench-mega-traces.tabularn<1K18 likes6.7k downloads23h agoHugging Face02thethanksforthegod /quran-asr-mega-corpustabular10K<n<100K1 likes1.2k downloads14d agoHugging Face03Mr-Philo /dolma3_dolmino_megatron_tokenize Dolma 3 / Dolmino Megatron-LM indexed dataset This repository contains immutable Megatron-LM indexed datasets (.bin and .idx) produced from pinned Dolma 3 and Dolmino releases. It intentionally contains no training checkpoints, experiment outputs, logs, or dataset caches. The indexed payloads were derived from these pinned public datasets: allenai/dolma3_mix-150B-1025@afa92bfb22366821c5e6cd427cdd036b34b713ef… See the full description on the dataset page: https://huggingface.co/datasets/Mr-Philo/dolma3_dolmino_megatron_tokenize.tabulartext-generationn<1K0 likes193 downloads24d agoHugging Face04nyu-dice-lab /lm-eval-results-Eurdem-megatron_2.1_MoE_2x7B-private Dataset Card for Evaluation run of Eurdem/megatron_2.1_MoE_2x7B Dataset automatically created during the evaluation run of model Eurdem/megatron_2.1_MoE_2x7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Eurdem-megatron_2.1_MoE_2x7B-private.tabular100K<n<1M0 likes85 downloads2y agoHugging Face05meg /requeststabularn<1K0 likes78 downloads2y agoHugging Face06open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8-details.tabular10K<n<100K0 likes58 downloads2y agoHugging Face07open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8.9-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8.9 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8.9 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8.9-details.tabular10K<n<100K0 likes58 downloads2y agoHugging Face08tyzhu /megamath-web-pro-max-splittedtabular10M<n<100M0 likes45 downloads3d agoHugging Face09open-llm-leaderboard /prithivMLmods__Megatron-Corpus-14B-Exp-detailsgated Dataset Card for Evaluation run of prithivMLmods/Megatron-Corpus-14B-Exp Dataset automatically created during the evaluation run of model prithivMLmods/Megatron-Corpus-14B-Exp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Megatron-Corpus-14B-Exp-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face10ThuraAung1601 /cd_hparam_search_whisper_th_megaspeech_v3tabularn<1K0 likes29 downloads22d agoHugging Face11open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v4-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v4 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v4-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face12open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v9.2-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v9.2 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v9.2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v9.2-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face13open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8.7-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8.7 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8.7 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8.7-details.tabular10K<n<100K0 likes23 downloads2y agoHugging Face14open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v5-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v5 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v5 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v5-details.tabular10K<n<100K0 likes19 downloads2y agoHugging Face15meg /requests_debugtabularn<1K0 likes18 downloads2y agoHugging Face16open-llm-leaderboard /prithivMLmods__Megatron-Corpus-14B-Exp.v2-detailsgated Dataset Card for Evaluation run of prithivMLmods/Megatron-Corpus-14B-Exp.v2 Dataset automatically created during the evaluation run of model prithivMLmods/Megatron-Corpus-14B-Exp.v2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Megatron-Corpus-14B-Exp.v2-details.tabular10K<n<100K0 likes15 downloads2y agoHugging Face17open-llm-leaderboard /aws-prototyping__MegaBeam-Mistral-7B-512k-detailsgated Dataset Card for Evaluation run of aws-prototyping/MegaBeam-Mistral-7B-512k Dataset automatically created during the evaluation run of model aws-prototyping/MegaBeam-Mistral-7B-512k The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/aws-prototyping__MegaBeam-Mistral-7B-512k-details.tabular10K<n<100K0 likes13 downloads2y agoHugging Face18open-llm-leaderboard /prithivMLmods__Megatron-Opus-14B-2.0-detailsgated Dataset Card for Evaluation run of prithivMLmods/Megatron-Opus-14B-2.0 Dataset automatically created during the evaluation run of model prithivMLmods/Megatron-Opus-14B-2.0 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Megatron-Opus-14B-2.0-details.tabular10K<n<100K0 likes13 downloads2y agoHugging Face19srbwin /agentic-trace-megacorpus-10tb Agentic Trace Megacorpus — ~2.06 TB (aggregated by reference) A reference-aggregation of 432 public trace datasets (agentic coding, reasoning, SWE, tool-use) totaling ~2.06 TB, assembled for MiniMax-M3 post-training. Nothing is re-hosted by value here yet — the loader streams directly from each source repo. Byte mirroring into this repo (to fill the 10 TB quota) is done Hub→Hub via mirror_to_hub.py. Composition (TB by relevance group) group TB agentic… See the full description on the dataset page: https://huggingface.co/datasets/srbwin/agentic-trace-megacorpus-10tb.tabularn<1K1 likes13 downloads3mo agoHugging Face20open-llm-leaderboard /prithivMLmods__Megatron-Opus-14B-2.1-detailsgated Dataset Card for Evaluation run of prithivMLmods/Megatron-Opus-14B-2.1 Dataset automatically created during the evaluation run of model prithivMLmods/Megatron-Opus-14B-2.1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Megatron-Opus-14B-2.1-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face21open-llm-leaderboard /ZeroXClem__Llama-3.1-8B-AthenaSky-MegaMix-detailsgated Dataset Card for Evaluation run of ZeroXClem/Llama-3.1-8B-AthenaSky-MegaMix Dataset automatically created during the evaluation run of model ZeroXClem/Llama-3.1-8B-AthenaSky-MegaMix The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ZeroXClem__Llama-3.1-8B-AthenaSky-MegaMix-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face22open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v9.1-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v9.1 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v9.1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v9.1-details.tabular10K<n<100K0 likes11 downloads2y agoHugging Face23open-llm-leaderboard /Junhoee__Qwen-Megumin-detailsgated Dataset Card for Evaluation run of Junhoee/Qwen-Megumin Dataset automatically created during the evaluation run of model Junhoee/Qwen-Megumin The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Junhoee__Qwen-Megumin-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face24open-llm-leaderboard /prithivMLmods__Megatron-Opus-14B-Exp-detailsgated Dataset Card for Evaluation run of prithivMLmods/Megatron-Opus-14B-Exp Dataset automatically created during the evaluation run of model prithivMLmods/Megatron-Opus-14B-Exp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Megatron-Opus-14B-Exp-details.tabular10K<n<100K0 likes9 downloads2y agoHugging Face25open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v3-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v3 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v3-details.tabular10K<n<100K0 likes9 downloads2y agoHugging Face26open-llm-leaderboard /CultriX__Qwen2.5-14B-MegaMerge-pt2-detailsgated Dataset Card for Evaluation run of CultriX/Qwen2.5-14B-MegaMerge-pt2 Dataset automatically created during the evaluation run of model CultriX/Qwen2.5-14B-MegaMerge-pt2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__Qwen2.5-14B-MegaMerge-pt2-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face27open-llm-leaderboard /prithivMLmods__Megatron-Opus-7B-Exp-detailsgated Dataset Card for Evaluation run of prithivMLmods/Megatron-Opus-7B-Exp Dataset automatically created during the evaluation run of model prithivMLmods/Megatron-Opus-7B-Exp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Megatron-Opus-7B-Exp-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face28open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v9-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v9 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v9 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v9-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face29meganariley /shelflife-datatabular1K<n<10K0 likes8 downloads5mo agoHugging Face30open-llm-leaderboard /amazon__MegaBeam-Mistral-7B-300k-detailsgated Dataset Card for Evaluation run of amazon/MegaBeam-Mistral-7B-300k Dataset automatically created during the evaluation run of model amazon/MegaBeam-Mistral-7B-300k The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/amazon__MegaBeam-Mistral-7B-300k-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.