CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Seenka /directv-zocalos-5fps Dataset Card for "directv-zocalos-5fps" More Information needed image10K<n<100K0 likes162 downloads3y agoHugging Face02cyttic /heb-directfit-1mimage100K<n<1M0 likes158 downloads3mo agoHugging Face03WaltonFuture /GeoQA-8K-direct-synthesizingThis dataset supports the unsupervised post-training of multi-modal large language models (MLLMs) as described in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO. It's designed to enable continual self-improvement without external supervision, using a self-rewarding mechanism based on majority voting over multiple sampled responses. The dataset is used to improve the reasoning ability of MLLMs, as demonstrated by significant improvements on benchmarks like MathVista… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/GeoQA-8K-direct-synthesizing.imageimage-text-to-text1K<n<10K1 likes68 downloads1y agoHugging Face04WaltonFuture /geometry3k-direct-synthesizingThis dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO to improve the reasoning abilities of multi-modal large language models (MLLMs). It contains image-text pairs, where each pair consists of an image, a problem described in text, and the corresponding answer. The dataset is designed for unsupervised post-training of MLLMs. 🐙 GitHub Repo: waltonfuture/MM-UPT 📜 Paper (arXiv): Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/geometry3k-direct-synthesizing.imageimage-text-to-text1K<n<10K1 likes50 downloads1y agoHugging Face05gankun /G-Directed-Graph Dataset Card Add more information here This dataset was produced with DataDreamer 🤖💤. The synthetic dataset card can be found here. imagen<1K0 likes47 downloads1y agoHugging Face06Seenka /banners-directv_peach Dataset Card for "banners-directv_peach" More Information needed image1K<n<10K0 likes44 downloads3y agoHugging Face07cyttic /heb-direct-renderimage100K<n<1M0 likes41 downloads3mo agoHugging Face08AliMertTemizsoy /visualpuzzles-direct-gemini25proimagen<1K0 likes38 downloads8mo agoHugging Face09AliMertTemizsoy /visualpuzzles-direct-gemini3proimagen<1K1 likes38 downloads8mo agoHugging Face10ThanhNX /Object_direction_1imagen<1K0 likes34 downloads3y agoHugging Face11WaltonFuture /MMR1-direct-synthesizingThis dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO to evaluate the performance of an unsupervised post-training framework (MM-UPT) for multi-modal LLMs. The dataset contains image-text pairs where the text represents a problem or question, and the corresponding answer. It is used to demonstrate the ability of MM-UPT to improve the reasoning capabilities of the Qwen2.5-VL-7B model without relying on any external supervised data. 🐙 GitHub Repo:… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/MMR1-direct-synthesizing.imageimage-text-to-text1K<n<10K1 likes31 downloads1y agoHugging Face12mertaylin /arc-agi-transduction100k-direct-ftimage100K<n<1M0 likes29 downloads2y agoHugging Face13dddraxxx /vstar_direct_attributes_seal_zoomimagen<1K0 likes29 downloads11mo agoHugging Face14Seenka /directv-zocalos-2fps Dataset Card for "directv-zocalos-2fps" More Information needed image1K<n<10K0 likes24 downloads3y agoHugging Face15eunjuri /direction_horizontal_tacThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "Unitree_G1_Inspire", "total_episodes": 1, "total_frames": 775, "total_tasks": 1, "total_videos": 0, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:1" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/eunjuri/direction_horizontal_tac.imageroboticsn<1K0 likes24 downloads1y agoHugging Face16Seenka /directv-zocalos-agosto-5fps Dataset Card for "directv-zocalos-agosto-5fps" More Information needed imagen<1K0 likes21 downloads3y agoHugging Face17Seenka /directv-zocalos_1.0fps_03-07-2023_05-07-2023 Dataset Card for "directv-zocalos_1.0fps_03-07-2023_05-07-2023" More Information needed imagen<1K0 likes13 downloads3y agoHugging Face18yobro4619 /direct_difficulty_final_reviewimagen<1K0 likes13 downloads7mo agoHugging Face19Seenka /directv-zocalos-agosto-5fps_vectors Dataset Card for "directv-zocalos-agosto-5fps_vectors" More Information needed imagen<1K0 likes12 downloads3y agoHugging Face20AliMertTemizsoy /puzzlevqa-direct-gemini3proimage1K<n<10K1 likes12 downloads8mo agoHugging Face21Seenka /directv-zocalos-new-test-1fps Dataset Card for "directv-zocalos-new-test-1fps" More Information needed imagen<1K0 likes11 downloads3y agoHugging Face22Seenka /directv-zocalos_1.0fps_03-08-2023_05-08-2023 Dataset Card for "directv-zocalos_1.0fps_03-08-2023_05-08-2023" More Information needed imagen<1K0 likes11 downloads3y agoHugging Face23AliMertTemizsoy /mmiq-direct-gemini25proimage1K<n<10K0 likes11 downloads8mo agoHugging Face24Seenka /directv-zocalos Dataset Card for "directv-zocalos" More Information needed imagen<1K0 likes9 downloads3y agoHugging Face25mertaylin /arc-agi-ft-direct-v1imagen<1K0 likes9 downloads2y agoHugging Face26teddyk251 /self-imagine-direct-inferenceimage10K<n<100K0 likes9 downloads1y agoHugging Face27stckmn /ocr-output-Directive017-1761352144 Document OCR using LightOnOCR-0.9B-32k-1025 This dataset contains OCR results from images in stckmn/ocr-input-Directive017-1761352140 using LightOnOCR, a fast and compact 1B OCR model. Processing Details Source Dataset: stckmn/ocr-input-Directive017-1761352140 Model: lightonai/LightOnOCR-0.9B-32k-1025 Vocabulary Size: 32k tokens Number of Samples: 10 Processing Time: 1.7 min Processing Date: 2025-10-25 00:32 UTC Configuration Image Column: image Output… See the full description on the dataset page: https://huggingface.co/datasets/stckmn/ocr-output-Directive017-1761352144.imagen<1K0 likes9 downloads11mo agoHugging Face28Seenka /directv-zocalos_1.0fps_21-08-2023_24-08-2023 Dataset Card for "directv-zocalos_1.0fps_21-08-2023_24-08-2023" More Information needed imagen<1K0 likes8 downloads3y agoHugging Face29stckmn /ocr-output-Directive017-1761354526 Document OCR using NuMarkdown-8B-Thinking This dataset contains markdown-formatted OCR results from images in stckmn/ocr-input-Directive017-1761354522 using NuMarkdown-8B-Thinking. Processing Details Source Dataset: stckmn/ocr-input-Directive017-1761354522 Model: numind/NuMarkdown-8B-Thinking Number of Samples: 21 Processing Time: 3.8 minutes Processing Date: 2025-10-25 01:17 UTC Configuration Image Column: image Output Column: markdown Dataset Split:… See the full description on the dataset page: https://huggingface.co/datasets/stckmn/ocr-output-Directive017-1761354526.imagen<1K0 likes8 downloads11mo agoHugging Face30AliMertTemizsoy /bilsem-batch-puzzlevqa-direct-gemini-2.5-proimage1K<n<10K0 likes8 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.