CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01R2E-Gym /R2E-Gym-Litetabular10K<n<100K1 likes68k downloads2y agoHugging Face02R2E-Gym /R2E-Gym-V1tabular1K<n<10K2 likes52k downloads2mo agoHugging Face03R2E-Gym /R2E-Gym-Subsettabular1K<n<10K29 likes28k downloads2mo agoHugging Face04R2E-Gym /SWE-Bench-Verifiedtextn<1K0 likes5.1k downloads2y agoHugging Face05PrimeIntellect /R2E-Gym-Subset-Verified R2E-Gym-Subset-Verified Gold-patch-validated subset of R2E-Gym/R2E-Gym-Subset (paper). The train split contains 4,522 / 4,578 rows (98.78%) verified scoreable end-to-end: apply the gold patch, run the upstream /testbed/run_tests.sh baked into the row's image, check the parsed outcomes against expected_output_json. Changes vs upstream Validation-only subset — our passes, run in fresh sandboxes per row: one full pass at concurrency 200, then a 10× retry pass over… See the full description on the dataset page: https://huggingface.co/datasets/PrimeIntellect/R2E-Gym-Subset-Verified.tabulartext-generation1K<n<10K1 likes3.2k downloads3mo agoHugging Face06R2E-Gym /SWE-Bench-Litetextn<1K0 likes1.2k downloads2y agoHugging Face07R2E-Gym /R2EGym-SFT-Trajectoriestext1K<n<10K11 likes1k downloads2y agoHugging Face08ryankamiri /R2E-Gym-Full R2E-Gym Subset Filtered for MAGRPO Filtered subset of R2E-Gym optimized for 2-agent MAGRPO training with 7B models. Dataset Statistics Total instances: 167 Format: Issue description + Oracle files in prompt Optimized for: 2-agent collaboration, 7B models Filtering Criteria (SWE-bench Lite Style) Problem statement: >40 words (up to 500 for context window) Must have non-empty oracle patch (non-test file changes) File count: Exactly 1 oracle file (single-file… See the full description on the dataset page: https://huggingface.co/datasets/ryankamiri/R2E-Gym-Full.tabulartext-generationn<1K0 likes801 downloads10mo agoHugging Face09ryankamiri /R2E-Gym-Collabtabular1K<n<10K0 likes618 downloads9mo agoHugging Face10R2E-Gym /R2EGym-Verifier-Trajectoriestext1K<n<10K3 likes369 downloads2y agoHugging Face11open-athena /rl__24GPU_base__mix_h2_language_balanced__r2egym-nl2bash-stacktext10K<n<100K0 likes299 downloads7mo agoHugging Face12rasdani /R2E-Gym-Subset-Oracletabular1K<n<10K0 likes260 downloads1y agoHugging Face13SumanthRH /R2E-Gym-Subsettabular1K<n<10K0 likes226 downloads1y agoHugging Face14R2E-Gym /R2EGym-TestingAgent-SFT-Trajectoriestext1K<n<10K4 likes214 downloads2y agoHugging Face15DCAgent /rl__24GPU_base__swe_rebench_patched_oracle__r2egym-nl2bash-stacktext10K<n<100K0 likes210 downloads7mo agoHugging Face16R2E-Gym /R2E-TestgenAgent-Patchestextn<1K1 likes201 downloads1y agoHugging Face17zhenghaoxu /R2E-Gym-Lite-Truncate-7B-Fixedtabular1K<n<10K0 likes190 downloads1y agoHugging Face18rasdani /R2E-Gym-Subset-contexttabular1K<n<10K0 likes170 downloads1y agoHugging Face19R2E-Gym /R2EGym-VerifierTrajectories-PatchOnlytext1K<n<10K0 likes156 downloads2y agoHugging Face20open-athena /rl__24GPU_base__code-contests-noblock__r2egym-nl2bash-stacktext10K<n<100K0 likes155 downloads6mo agoHugging Face21open-athena /rl__24GPU_shaped__swe_rebench_patched_oracle__r2egym-nl2bash-stacktext10K<n<100K0 likes136 downloads6mo agoHugging Face22zhenghaoxu /R2E-Gym-Lite-RFT-no-thinktext100K<n<1M1 likes133 downloads1y agoHugging Face23open-athena /a3-rl-DCAgent_r2egym-patched-full-oracletext10K<n<100K0 likes126 downloads4mo agoHugging Face24loongsage /R2E-Gym R2E-Gym: OpenCode, Codex, and Claude Code environments This repository contains three environment variants of R2E-Gym/R2E-Gym-Subset. Each variant contains the same 4,578 tasks in five Parquet shards, with the original 14-column schema and task order. Only docker_image is replaced with the corresponding public image containing the selected agent. Configuration Files Docker image prefix opencode opencode/data/*.parquet docker.io/loongsage/r2e-gym:opencode_ codex… See the full description on the dataset page: https://huggingface.co/datasets/loongsage/R2E-Gym.tabulartext-generation10K<n<100K0 likes126 downloads1d agoHugging Face25zhenghaoxu /R2E-Gym-Lite-Truncate-7Btabular1K<n<10K0 likes119 downloads1y agoHugging Face26zhenghaoxu /R2E-Gym-Lite-RFTtext10K<n<100K0 likes115 downloads1y agoHugging Face27open-athena /r2egym-patched-full-oracle-qwen3.5-122b-131k-opencode-literal-rescue-traces Agent trace dataset Decoding the literal token IDs The prompt_token_ids / completion_token_ids / logprobs columns are the verbatim tokens the serving engine emitted, stored PER AGENT STEP as a list-of-lists (one inner list per turn). To turn them back into text you MUST use the exact tokenizer the model was served with — a generic same-family tokenizer will decode word tokens to garbage. Served model / tokenizer source: Qwen/Qwen3.5-122B-A10B-FP8 from transformers… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/r2egym-patched-full-oracle-qwen3.5-122b-131k-opencode-literal-rescue-traces.text1K<n<10K0 likes114 downloads2mo agoHugging Face28synthetic-code-training /r2egym_gpt5mini_1500itext1K<n<10K0 likes109 downloads8d agoHugging Face29DCAgent /e1_gpt_long_r2egym_sandboxes_4x_glm_4.7_traces_jupitertext10K<n<100K0 likes86 downloads5mo agoHugging Face30oscarfco /R2E-Gym-Yiming-Combined-1p6pcttabular1K<n<10K0 likes83 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.