CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01SeonghuJeon /robocasa_cosmos24_success300_env_depthimagen<1K0 likes9.3k downloads4mo agoHugging Face02seongwon980 /Nameonly_generated Just Say the Name: Online Continual Learning with Category Names Only via Data Generation We provide the dataset used for Name-only continual learning, generated using Stable Diffusion XL, DALL.E-2, CogView2, and DeepFloyd IF models. Disclaimer This dataset is created solely for academic purposes. We minimized human intervention to ensure a fair comparison with the baseline methods discussed in our paper. Despite our efforts, the extensive size of the dataset prevented us… See the full description on the dataset page: https://huggingface.co/datasets/seongwon980/Nameonly_generated.image10K<n<100K1 likes3.5k downloads2y agoHugging Face03SeonghoonYu /Robotwinimage100K<n<1M0 likes2.1k downloads3mo agoHugging Face04seonglae /vls-100k VLS 100K 74,936 MS-COCO images paired with a long written description, a one-sentence spoken summary of that description, the spoken audio, and that audio pre-encoded with a neural codec. Images and audio are embedded in the parquet, so the viewer renders them and one call opens the set: from datasets import load_dataset ds = load_dataset("seonglae/vls-100k", split="train") ds[0]["image"] # PIL image ds[0]["audio"] # decoded waveform ds[0]["sst"] # the sentence that was… See the full description on the dataset page: https://huggingface.co/datasets/seonglae/vls-100k.audiotext-to-speech10K<n<100K0 likes258 downloads28d agoHugging Face05seonglae /vls-10k VLS 10K 9,987 MS-COCO images paired with a long written description, a one-sentence spoken summary of that description, the spoken audio, and that audio pre-encoded by two neural codecs. Images and audio are embedded in the parquet, so the viewer renders them and one call opens the set: from datasets import load_dataset ds = load_dataset("seonglae/vls-10k", split="train") ds[0]["image"] # PIL image ds[0]["audio"] # decoded waveform ds[0]["sst"] # the sentence that was… See the full description on the dataset page: https://huggingface.co/datasets/seonglae/vls-10k.audiotext-to-speech1K<n<10K0 likes173 downloads28d agoHugging Face06seosiju /BigEarthNet-S1image100K<n<1M0 likes126 downloads1y agoHugging Face07Gingiris /gingiris-seo-geo SEO & GEO Growth Playbook Rank on Google AND get cited by AI search engines — the dual-engine SEO + GEO strategy guide behind ~32K impressions/month, built from AFFiNE's 60K organic stars + 150 AI startup consultations English | 中文 📦 Install npx skills add Gingiris-1031/gingiris-seo-geo Then ask your AI agent: "Audit my SaaS site for SEO/GEO" or "Make my blog visible to ChatGPT and Perplexity" Installs the complete SEO + GEO… See the full description on the dataset page: https://huggingface.co/datasets/Gingiris/gingiris-seo-geo.imagetext-generationn<1K1 likes111 downloads2mo agoHugging Face08seooyxx /canvas CANVAS This repository contains the dataset accompanying the paper CANVAS: A Benchmark for Vision-Language Models on Tool-Based UI Design (AAAI 2026). CANVAS designed to evaluate a VLM's capability to generate a UI design with tool invocations in two tasks: ADesign Replication and BDesign Modification. A. Design Replication. Single Tasks involving restoring (implementing) a given UI image as is. B. Design Modification. Multiple Tasks involving modifying or transforming existing… See the full description on the dataset page: https://huggingface.co/datasets/seooyxx/canvas.image2 likes107 downloads9mo agoHugging Face09seokwon99 /MAVIS MAVIS: A Benchmark for Multimodal Source Attribution in Long-form Visual Question Answering 📖 Paper | 💻 Evaluation Dataset Summary MAVIS is a new dataset for open-domain, long-form visual question answering, characterized by three key features: (1) the questions incorporate input images, requiring visual understanding to correctly interpret the user’s intent; (2) the desired answers are long-form, necessitating the retrieval and synthesis of diverse information rather… See the full description on the dataset page: https://huggingface.co/datasets/seokwon99/MAVIS.imagequestion-answeringn<1K1 likes106 downloads8mo agoHugging Face10SeonghoonYu /vlabench_primitive_ft_lerobot_224 vlabench_primitive_ft_lerobot_224 VLABench/vlabench_primitive_ft_lerobot (the official pi0.5 VLABench fine-tuning set, LeRobot format) with every camera image re-encoded from 480x480 to 224x224 (bit-identical to openpi's resize_with_pad applied at load time), so that training does not pay for the 480px decode + resize and the LeRobot Arrow cache stays small. Same episodes, same schema and metadata; only the image resolution differs. Made with… See the full description on the dataset page: https://huggingface.co/datasets/SeonghoonYu/vlabench_primitive_ft_lerobot_224.image100K<n<1M0 likes102 downloads13d agoHugging Face11seokwon99 /MAVIS_documentsimage10K<n<100K0 likes98 downloads8mo agoHugging Face12SeongryongJung /chemistry-jsdhist3-summary Chemistry SDPO jsdhist3 Training History This repository contains summarized training-history artifacts for four completed Chemistry SDPO Qwen3-4B runs from /workspace/SDPO-new-clean. Images Files data/jsd_history.csv: per-step JSD scalar history and validation metrics parsed from the training log. data/run_summary.csv: one-row-per-run summary with final/best validation reward and histogram event counts.… See the full description on the dataset page: https://huggingface.co/datasets/SeongryongJung/chemistry-jsdhist3-summary.imagen<1K0 likes94 downloads2mo agoHugging Face13Hee-Seon /save-coco-text-tmpimagen<1K0 likes86 downloads8mo agoHugging Face14SeonghoonYu /Robotwin-12tasksimage100K<n<1M0 likes52 downloads3mo agoHugging Face15seomh /Drakeimage0 likes16 downloads1y agoHugging Face16seojing /pickplace_pandaThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": null, "total_episodes": 12, "total_frames": 3605, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 10, "splits": { "train": "0:12" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/seojing/pickplace_panda.imagerobotics1K<n<10K0 likes11 downloads7mo agoHugging Face17seokheeyam /funsd-bank-paragraph-test3imagen<1K0 likes8 downloads2y agoHugging Face18SEOje /japdongsani train_2024-06-03-09-50-18 This model is a fine-tuned version of yanolja/EEVE-Korean-Instruct-10.8B-v1.0 on the assay dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning_rate: 0.0003 train_batch_size: 2… See the full description on the dataset page: https://huggingface.co/datasets/SEOje/japdongsani.imagen<1K0 likes6 downloads2y agoHugging Face19seochan99 /ecb-datasets ECB Datasets: Cultural Bias Evaluation in Generative Image Models Overview This dataset contains human evaluation data for cultural bias analysis in image generation models, supporting the research paper "Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models". ECB stands for "Evaluation Cultural Bias". The dataset includes prompts, generated images, and evaluation metrics across different countries and cultural contexts. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/seochan99/ecb-datasets.imagetext-to-image1K<n<10K0 likes6 downloads11mo agoHugging Face20seoheelee /face-test-02-mmppimagen<1K0 likes5 downloads2y agoHugging Face21KimJiwoo /My-Seoul-Imagesimage1K<n<10K0 likes5 downloads10mo agoHugging Face22jmha02 /seoul_subwayKorean Subway images, filmed from Seoul Station to Seoul National University. image1K<n<10K0 likes3 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.