CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Lansechen /details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192 Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192 Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 16 run(s). Each run can be found as a specific split in each configuration, the split… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-50b-batch32-epoch1-8192.text1K<n<10K0 likes275 downloads2y agoHugging Face02Lansechen /details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192 Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192 Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 14 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-2k-simplified-batch32-epoch1-8192.text1K<n<10K0 likes249 downloads2y agoHugging Face03HaoranoLee /20260722_3D_test_epoch1text0 likes185 downloads2mo agoHugging Face04lixiaochuan2020 /acm-browsecompplus-teacher-logprobs-qwen3.5-9b-epoch1 lixiaochuan2020/acm-browsecompplus-teacher-logprobs-qwen3.5-9b-epoch1 Teacher (Qwen3.5-397B-A17B) top-20 forward-KL log-prob annotations for offline on-policy distillation (OPD) of Qwen3.5-9B on BrowseComp-Plus train680 (MemTool regime). Trains: OPD iter-1 Annotates the rollouts of: base model rollouts (react+memtool ×4) One .npz per (question, rep) trajectory · 1145 files. Schema (per file, numpy.load) key shape dtype meaning input_ids (L,) int32… See the full description on the dataset page: https://huggingface.co/datasets/lixiaochuan2020/acm-browsecompplus-teacher-logprobs-qwen3.5-9b-epoch1.0 likes154 downloads2mo agoHugging Face05lixiaochuan2020 /acm-browsecompplus-train-rollouts-qwen3.5-9b-epoch1 lixiaochuan2020/acm-browsecompplus-train-rollouts-qwen3.5-9b-epoch1 BrowseComp-Plus train680 pass@4 rollouts (MemTool regime) from qwen3.5-9b-opd_iter1 — one row per (question, rep) in rollouts.jsonl. Fields: question, gold, final_answer, correct_gpt5 (GPT-5 judge), num_turns, tool_call_counts, and full trajectories (raw_history + folded history + mem_operations + token_trajectory). Stats: 2720 rollouts / 680 questions / reps [1, 2, 3, 4] · pass@1 70.0% · pass@4 83.2% (GPT-5). 0 likes151 downloads2mo agoHugging Face06Lansechen /details_Lansechen__Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384 Dataset Card for Evaluation run of Lansechen/Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384 Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 9 run(s). Each run can be found as a specific split in each configuration, the split being named using… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-7B-Instruct-Distill-om220k-1k-origin-batch32-epoch1-16384.text1K<n<10K0 likes92 downloads2y agoHugging Face07epoch14 /fly-sim-runs0 likes57 downloads14d agoHugging Face08yunjae-won /mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_160k_ratiotext100K<n<1M0 likes56 downloads1y agoHugging Face09Lansechen /details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192 Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192 Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192. The dataset is composed of 2 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Instruct-Distill-om220k-1k-simplified-fem8192-batch32-epoch1-8192.text1K<n<10K0 likes52 downloads2y agoHugging Face10Stage-org /appworld-qwen35-4b-agent-rl-epoch1-tool-eval0 likes46 downloads8d agoHugging Face11yunjae-won /mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_multisample_2.5ktext10K<n<100K0 likes45 downloads1y agoHugging Face12Lansechen /details_Lansechen__Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192 Dataset Card for Evaluation run of Lansechen/Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192 Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-1.5B-Instruct-Distill-om220k-1k-simplified-batch32-epoch1-8192.text1K<n<10K0 likes43 downloads2y agoHugging Face13yunjae-won /mp_mistral7bv3_sft_dpo_beta1e-1_epoch1_40k_n16tabular10K<n<100K0 likes42 downloads1y agoHugging Face14Stage-org /appworld-qwen35-4b-from-27b-epoch1-tool-eval0 likes41 downloads7d agoHugging Face15yunjae-won /mp_mistral7bv3_sft_dpo_beta2e-2_epoch1_20k_n8tabular10K<n<100K0 likes40 downloads1y agoHugging Face16Stage-org /appworld-qwen35-4b-agent-rl-semi_hard-epoch1-tool-eval0 likes40 downloads7d agoHugging Face17yunjae-won /mpg27_mistral7bv3_sft_dpo_beta5e-2_epoch1_40k_multisample_ratiotabular10K<n<100K0 likes38 downloads1y agoHugging Face18Stage-org /appworld-qwen35-4b-agent-rl-epoch12-tool-eval0 likes38 downloads2d agoHugging Face19Stage-org /appworld-qwen35-9b-from-27b-epoch1-tool-eval0 likes37 downloads7d agoHugging Face20Stage-org /4b-solvability-200-nyshot-27b-z-epoch3-agent-rl-epoch1-tool-eval0 likes37 downloads5d agoHugging Face21Stage-org /appworld-qwen35-4b-A-200-epoch3-agent-rl-epoch1-tool-eval0 likes36 downloads7d agoHugging Face22Stage-org /appworld-4b-strat-300-LH-27b-z-e1-iter3-lowlr-epoch2-agent-rl-epoch1-tool-eval0 likes36 downloads2d agoHugging Face23yunjae-won /mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_40ktext10K<n<100K0 likes35 downloads1y agoHugging Face24Stage-org /appworld-qwen35-4b-agent-rl-hard-epoch1-tool-eval0 likes35 downloads8d agoHugging Face25Stage-org /appworld-qwen35-4b-solvability-200-27b-fixed-epoch3-agent-rl-epoch1-tool-eval0 likes34 downloads5d agoHugging Face26Stage-org /appworld-qwen35-4b-agent-rl-epoch10-tool-eval0 likes34 downloads3d agoHugging Face27yunjae-won /mp_mistral7bv3_sft_dpo_beta2e-2_epoch1_160k_ratiotext100K<n<1M0 likes32 downloads1y agoHugging Face28Stage-org /appworld-4b-strat-300-4b-z-epoch3-agent-rl-epoch1-tool-eval0 likes32 downloads3d agoHugging Face29Stage-org /appworld-4b-strat-300-LH-27b-z-e2-iter3-epoch3-agent-rl-epoch1-tool-eval0 likes32 downloads3d agoHugging Face30Stage-org /appworld-4b-strat-300-LH-27b-z-e1-iter3-lowlr-epoch2-agent-rl-epoch1-low_lr-tool-eval0 likes32 downloads2d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.