CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tanhuajie2001 /Reason-RFT-CoT-Dataset 🤗 Reason-RFT CoT Dateset The full dataset used in our project "Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning".   ⭐️ Project   │   🌎 Github   │   🔥 Models   │   📑 ArXiv   │   💬 WeChat   🤖 RoboBrain: Aim to Explore ReasonRFT Paradigm to Enhance RoboBrain's Embodied Reasoning Capabilities. ♣️ Quick Start Please refer to Dataset Preparation 🔥 Overview Visual reasoning abilities play a crucial role in understanding complex multimodal… See the full description on the dataset page: https://huggingface.co/datasets/tanhuajie2001/Reason-RFT-CoT-Dataset.imagereinforcement-learning100K<n<1M11 likes2k downloads1y agoHugging Face02kevinshin /wildchat-creative-writing-3k-rfttext1K<n<10K0 likes214 downloads1y agoHugging Face03cjfcsjt /142_sft_rft_dpo_simpo_v2 webshopv_sft300_3hist_rft_dpo_v2_simpo2.0 This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/LLaMA-Factory/models/qwen2_vl_lora_sft_rft_v2 on the vl_dpo_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft_dpo_simpo_v2.imagen<1K0 likes134 downloads2y agoHugging Face04zhenghaoxu /R2E-Gym-Lite-RFT-no-thinktext100K<n<1M1 likes128 downloads1y agoHugging Face05kelatte /qwen3-8b-dpo-agentgym-rft-datatext1K<n<10K0 likes116 downloads3mo agoHugging Face06felixZzz /4b_rft_response-2-custom_student_responsetabular100K<n<1M0 likes110 downloads1y agoHugging Face07HKUST-DSAIL /Graph-R1-RFT-COT-30K Dataset Card: Graph-CoT-30k Dataset Details Dataset Name: Graph-CoT-30k Dataset Creator: HKUST-DSAIL Dataset Version: 1.0 Release Date: August 2025 Description Graph-CoT-30k is a large-scale, high-quality instruction tuning dataset designed to enhance the reasoning capabilities of large language models (LLMs) on complex graph-theoretic problems. It contains 30,000 question-answer (QA) pairs, each featuring ultra-long chain-of-thought (CoT) reasoning… See the full description on the dataset page: https://huggingface.co/datasets/HKUST-DSAIL/Graph-R1-RFT-COT-30K.textquestion-answering10K<n<100K1 likes95 downloads1y agoHugging Face08zhenghaoxu /R2E-Gym-Lite-RFTtext10K<n<100K0 likes88 downloads1y agoHugging Face09cjfcsjt /142_sft_rft_v2 sft_rft_v2 This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/models/qwen2_vl_lora_sft_webshopv_300 on the vl_finetune_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning_rate: 0.0001… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft_v2.imagen<1K0 likes84 downloads2y agoHugging Face10lmquan /juno-landmark-rftimage10K<n<100K0 likes84 downloads11mo agoHugging Face11cjfcsjt /142_sft_rft sft_rft This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/models/qwen2_vl_lora_sft_webshopv_300 on the vl_finetune_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning_rate: 0.0001… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft.imagen<1K0 likes83 downloads2y agoHugging Face12mm-vl /x2x_rft_22kimage10K<n<100K1 likes80 downloads1y agoHugging Face13felixZzz /4b_rft_response-7-custom_student_response-verifiedtabular100K<n<1M0 likes79 downloads1y agoHugging Face14felixZzz /4b_rft_response-5-custom_student_response-verifiedtabular100K<n<1M0 likes79 downloads1y agoHugging Face15felixZzz /4b_rft_response-3-custom_student_response-verifiedtabular100K<n<1M0 likes77 downloads1y agoHugging Face16felixZzz /4b_rft_response-4-custom_student_response-verifiedtabular100K<n<1M0 likes72 downloads1y agoHugging Face17felixZzz /4b_rft_response-1-custom_student_response-verifiedtabular100K<n<1M0 likes70 downloads1y agoHugging Face18GraphWiz /GraphInstruct-RFT-72Ktextquestion-answering10K<n<100K10 likes66 downloads3y agoHugging Face19cjfcsjt /142_sft_rft_dpo_v2 webshopv_sft300_3hist_rft_dpo_v2 This model is a fine-tuned version of /mnt/nvme0n1p1/hongxin_li/jingfan/LLaMA-Factory/models/qwen2_vl_lora_sft_rft_v2 on the vl_dpo_data dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training:… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/142_sft_rft_dpo_v2.imagen<1K0 likes63 downloads2y agoHugging Face20JakeOh /rft-finetune-llama-3.2-1b-mathtext100K<n<1M0 likes60 downloads2y agoHugging Face21mm-vl /x2x_rft_16kimage10K<n<100K1 likes60 downloads1y agoHugging Face22bxw315-umd /dpv-rft-v3.1image10K<n<100K0 likes59 downloads11mo agoHugging Face23bxw315-umd /dpv-rft-v3.2image100K<n<1M0 likes55 downloads11mo agoHugging Face24felixZzz /4b_rft_response-4-custom_student_responsetabular100K<n<1M0 likes48 downloads1y agoHugging Face25felixZzz /4b_rft_response-2-custom_student_response-verifiedtabular100K<n<1M0 likes48 downloads1y agoHugging Face26bxw315-umd /dpv-rft-v3.x-potentialsimage100K<n<1M0 likes47 downloads7mo agoHugging Face27JakeOh /rft-finetune-llama-3.2-1b-math-k10text100K<n<1M0 likes46 downloads2y agoHugging Face28kimihailv /kk-onpolicy-rfttabular10K<n<100K0 likes45 downloads16d agoHugging Face29kimihailv /kk-rft-traintabular10K<n<100K0 likes43 downloads16d agoHugging Face30felixZzz /4b_rft_response-2-student_logpstext10K<n<100K0 likes41 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.