CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01NJU-LINK /WebCompass WebCompass A unified multimodal benchmark for evaluating LLMs' ability to generate, edit, and repair functional web pages. WebCompass spans three input modalities — text design documents, reference screenshots, and video demonstrations — and three task families — generation, editing, and repair. GitHub: NJU-LINK/WebCompass Project Page: nju-link.github.io/WebCompass Quick Start from datasets import load_dataset # Generation tasks (existing) ds_text =… See the full description on the dataset page: https://huggingface.co/datasets/NJU-LINK/WebCompass.imagetext-generationn<1K6 likes1.6k downloads4mo agoHugging Face02lingada /3DHarnessBench Probing Agentic 3D-to-Code Capabilities of Frontier Vision-Language Models Project Page · GitHub · Arxiv Abstract 3DHarnessBench evaluates the agentic capacity of frontier vision-language models (VLMs) to recover 3D geometry as executable Blender Python code from multiple forms of target evidence. Rather than restricting every system to a single fixed input, the benchmark compares four progressively richer harnesses: Single-view, Multi-view, Active… See the full description on the dataset page: https://huggingface.co/datasets/lingada/3DHarnessBench.image1K<n<10K2 likes1.6k downloads6h agoHugging Face03lingamvamshikrishnareddy /ramanv-image-editinggated ramanv-image-editing Image editing dataset for training FLUX.1-Kontext / InstructPix2Pix style models. Size 592,141 total editing pairs Sources: ultraedit Schema Each shard tar contains {uid}_src.jpg, {uid}_edit.jpg, {uid}_mask.png (where available). Metadata per record: instruction, prompt, edit_type, caption_before/after, license, sha256. Licenses MagicBrush, InstructPix2Pix, Pico-Banana, HumanEdit: CC-BY-4.0 UltraEdit, AnyEdit… See the full description on the dataset page: https://huggingface.co/datasets/lingamvamshikrishnareddy/ramanv-image-editing.image1K<n<10K7 likes536 downloads24d agoHugging Face04prashant0919 /nepali-synthetic-ocr-lines Nepali Synthetic OCR/HTR Document Line Dataset A dataset of synthetic Devanagari text line images imitating historical and official scanned document conditions, designed for OCR (Optical Character Recognition) and HTR (Handwritten Text Recognition) models such as TrOCR, CRNN, and PaddleOCR. This dataset was generated using the Mountmind PeakOCR Studio synthetic corpus generator pipeline, introducing realistic document aging artifacts like: Skew Angle Rotations (Hough line… See the full description on the dataset page: https://huggingface.co/datasets/prashant0919/nepali-synthetic-ocr-lines.imageimage-to-text1K<n<10K1 likes24 downloads4mo agoHugging Face05linxxx3 /pixmo_images_badimage100K<n<1M0 likes15 downloads3mo agoHugging Face06AlhitawiMohammed22 /lines_hu_v5image10K<n<100K0 likes9 downloads3y agoHugging Face07LinYC2024 /benchDriveimagequestion-answering1K<n<10K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.