CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alexshpunt /explicit-edit-benchmark Explicit Edit Benchmark 226 deterministic exact-edit tasks, run by different agents, harnesses, models and configurations. Every observation records what the harness did and whether the resulting files matched byte for byte. Source code and benchmark runner: GitHub — Explicit Edit Benchmark Open the interactive Explorer to compare agents, harnesses, models, versions, reasoning modes, correctness, recovery, time, cost and tokens. Leaderboard by model route Score v2… See the full description on the dataset page: https://huggingface.co/datasets/alexshpunt/explicit-edit-benchmark.tabulartext-generationn<1K2 likes7.8k downloads2d agoHugging Face02WillHeld /discrim-eval-explicit-subsetstabular1K<n<10K0 likes30 downloads1y agoHugging Face03connections-dev /connection_queries_natural_explicit_1__olmo3tabularn<1K0 likes14 downloads9mo agoHugging Face04connections-dev /connection_queries_natural_explicit_1__qwqtabularn<1K0 likes10 downloads9mo agoHugging Face05connections-dev /connection_queries_natural_explicit_1__gpt-4.1-mini-2025-04-14tabularn<1K0 likes6 downloads9mo agoHugging Face06open-llm-leaderboard /marcuscedricridia__etr1o-explicit-v1.1-detailsgated Dataset Card for Evaluation run of marcuscedricridia/etr1o-explicit-v1.1 Dataset automatically created during the evaluation run of model marcuscedricridia/etr1o-explicit-v1.1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/marcuscedricridia__etr1o-explicit-v1.1-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face07open-llm-leaderboard /marcuscedricridia__etr1o-explicit-v1.2-detailsgated Dataset Card for Evaluation run of marcuscedricridia/etr1o-explicit-v1.2 Dataset automatically created during the evaluation run of model marcuscedricridia/etr1o-explicit-v1.2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/marcuscedricridia__etr1o-explicit-v1.2-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face08connections-dev /connection_queries_natural_explicit_1__qwen32btabularn<1K0 likes5 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.