CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tanhuajie2001 /Reason-RFT-CoT-Dataset 🤗 Reason-RFT CoT Dateset The full dataset used in our project "Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning".   ⭐️ Project   │   🌎 Github   │   🔥 Models   │   📑 ArXiv   │   💬 WeChat   🤖 RoboBrain: Aim to Explore ReasonRFT Paradigm to Enhance RoboBrain's Embodied Reasoning Capabilities. ♣️ Quick Start Please refer to Dataset Preparation 🔥 Overview Visual reasoning abilities play a crucial role in understanding complex multimodal… See the full description on the dataset page: https://huggingface.co/datasets/tanhuajie2001/Reason-RFT-CoT-Dataset.imagereinforcement-learning100K<n<1M11 likes2k downloads1y agoHugging Face02IffYuan /Embodied-R1.5-RFT-Dataset Embodied-R1.5-RFT-Dataset 🌐 Project Page &nbsp;|&nbsp; 📄 arXiv &nbsp;|&nbsp; 💻 Code &nbsp;|&nbsp; 🧰 EmbodiedEvalKit &nbsp;|&nbsp; 🤗 Models & Datasets 🗓️ Update — 2026-08-20 (20260820). All 28 Stage 2 RFT JSON annotation files have been uploaded to rft_datasets_json/. The complete JSON ↔ media archive mapping is documented in the Dataset composition table below. ⚠️ Partial release. This repository currently contains only a subset of the full Stage 2 RFT data… See the full description on the dataset page: https://huggingface.co/datasets/IffYuan/Embodied-R1.5-RFT-Dataset.imageimage-text-to-text2 likes754 downloads1mo agoHugging Face03HenryZ07 /RFT-reference-trajectoryvideon<1K0 likes233 downloads22d agoHugging Face04kevinshin /wildchat-creative-writing-3k-rfttext1K<n<10K0 likes214 downloads1y agoHugging Face05huangjy-pku /3D-RFT-Reasoningimage0 likes167 downloads3mo agoHugging Face06cjfcsjt /142_sft_rft_dpo_simpo_v2 webshopv_sft300_3hist_rft_dpo_v2_simpo2.0 This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/LLaMA-Factory/models/qwen2_vl_lora_sft_rft_v2 on the vl_dpo_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft_dpo_simpo_v2.imagen<1K0 likes134 downloads2y agoHugging Face07zhenghaoxu /R2E-Gym-Lite-RFT-no-thinktext100K<n<1M1 likes128 downloads1y agoHugging Face08kelatte /qwen3-8b-dpo-agentgym-rft-datatext1K<n<10K0 likes116 downloads3mo agoHugging Face09felixZzz /4b_rft_response-2-custom_student_responsetabular100K<n<1M0 likes110 downloads1y agoHugging Face10robotflow /rftransThis repo contains the dataset used in RFTrans, generated by the Data Generator, powered by RFUniverse. The train and val folder contain the proposed synthetic dataset. You can also generate your own dataset with the tools and assets provided by us. Resources.zip contains the assets we used to generate the dataset. To note, we do not claim to own the copyright of these assets. The cleargrasp folder contains the example train set generated with the models from ClearGrasp. It's the train set we… See the full description on the dataset page: https://huggingface.co/datasets/robotflow/rftrans.1 likes101 downloads2y agoHugging Face11HKUST-DSAIL /Graph-R1-RFT-COT-30K Dataset Card: Graph-CoT-30k Dataset Details Dataset Name: Graph-CoT-30k Dataset Creator: HKUST-DSAIL Dataset Version: 1.0 Release Date: August 2025 Description Graph-CoT-30k is a large-scale, high-quality instruction tuning dataset designed to enhance the reasoning capabilities of large language models (LLMs) on complex graph-theoretic problems. It contains 30,000 question-answer (QA) pairs, each featuring ultra-long chain-of-thought (CoT) reasoning… See the full description on the dataset page: https://huggingface.co/datasets/HKUST-DSAIL/Graph-R1-RFT-COT-30K.textquestion-answering10K<n<100K1 likes95 downloads1y agoHugging Face12zhenghaoxu /R2E-Gym-Lite-RFTtext10K<n<100K0 likes88 downloads1y agoHugging Face13cjfcsjt /142_sft_rft_v2 sft_rft_v2 This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/models/qwen2_vl_lora_sft_webshopv_300 on the vl_finetune_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning_rate: 0.0001… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft_v2.imagen<1K0 likes84 downloads2y agoHugging Face14lmquan /juno-landmark-rftimage10K<n<100K0 likes84 downloads11mo agoHugging Face15cjfcsjt /142_sft_rft sft_rft This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/models/qwen2_vl_lora_sft_webshopv_300 on the vl_finetune_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning_rate: 0.0001… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft.imagen<1K0 likes83 downloads2y agoHugging Face16mm-vl /x2x_rft_22kimage10K<n<100K1 likes80 downloads1y agoHugging Face17Liang0223 /Qwen-2.5-Math-1.5B-RFT-DPO-Data0 likes80 downloads1y agoHugging Face18felixZzz /4b_rft_response-7-custom_student_response-verifiedtabular100K<n<1M0 likes79 downloads1y agoHugging Face19felixZzz /4b_rft_response-5-custom_student_response-verifiedtabular100K<n<1M0 likes79 downloads1y agoHugging Face20felixZzz /4b_rft_response-3-custom_student_response-verifiedtabular100K<n<1M0 likes77 downloads1y agoHugging Face21felixZzz /4b_rft_response-4-custom_student_response-verifiedtabular100K<n<1M0 likes72 downloads1y agoHugging Face22felixZzz /4b_rft_response-1-custom_student_response-verifiedtabular100K<n<1M0 likes70 downloads1y agoHugging Face23GraphWiz /GraphInstruct-RFT-72Ktextquestion-answering10K<n<100K10 likes66 downloads3y agoHugging Face24cjfcsjt /142_sft_rft_dpo_v2 webshopv_sft300_3hist_rft_dpo_v2 This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/LLaMA-Factory/models/qwen2_vl_lora_sft_rft_v2 on the vl_dpo_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training:… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft_dpo_v2.imagen<1K0 likes63 downloads2y agoHugging Face25JakeOh /rft-finetune-llama-3.2-1b-mathtext100K<n<1M0 likes60 downloads2y agoHugging Face26mm-vl /x2x_rft_16kimage10K<n<100K1 likes60 downloads1y agoHugging Face27bxw315-umd /dpv-rft-v3.1image10K<n<100K0 likes59 downloads11mo agoHugging Face28bxw315-umd /dpv-rft-v3.2image100K<n<1M0 likes55 downloads11mo agoHugging Face29ISdept /libero-rft-spatial-full1tabular10K<n<100K0 likes54 downloads22d agoHugging Face30felixZzz /4b_rft_response-4-custom_student_responsetabular100K<n<1M0 likes48 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.