CoolFace
20 results

concept

inclusionAI /ConceptEdit-12M ConceptEdit: Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision &nbsp;&nbsp;&nbsp; ConceptEdit-12M is a large-scale image editing dataset. Each sample is stored as a triplet: a source image, an edited image, a JSON metadata file describing the edit instruction, edit category, relative image paths, and VQA-style quality checks. The dataset is packaged as multiple .tar shards. All paths inside the tar files and JSON files are relative paths; no… See the full description on the dataset page: https://huggingface.co/datasets/inclusionAI/ConceptEdit-12M.image-to-image10M<n<100M47 likes61k downloads28d agoHugging Facegoogle-research-datasets /conceptual_captions Dataset Card for Conceptual Captions Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/google-research-datasets/conceptual_captions.imageimage-to-text1M<n<10M111 likes12k downloads2y agoHugging Faceveerlosar /rule-ling-concepts0 likes6.6k downloads13m agoHugging Facelaion /conceptual-captions-12m-webdatasetimage10K<n<100K34 likes6.2k downloads4y agoHugging Facewebshart /conceptual-captions-12m-webdataset-metadata Conceptual Captions 12M — Webshart metadata indices Per-shard webshart metadata indices for laion/conceptual-captions-12m-webdataset: 1,100 JSON files under data/, one per source tar shard, mirroring the source's shard layout. Each index records every tar member's byte offset and length (enabling ranged reads without downloading whole shards), image geometry (width/height for aspect bucketing), and — as of August 2026 — embedded captions for all 10,994,853 samples, coalesced… See the full description on the dataset page: https://huggingface.co/datasets/webshart/conceptual-captions-12m-webdataset-metadata.1 likes4.8k downloads1mo agoHugging Faceconceptnet5 /conceptnet5 Dataset Card for Conceptnet5 Dataset Summary ConceptNet is a multilingual knowledge base, representing words and phrases that people use and the common-sense relationships between them. The knowledge in ConceptNet is collected from a variety of resources, including crowd-sourced resources (such as Wiktionary and Open Mind Common Sense), games with a purpose (such as Verbosity and nadya.jp), and expert-created resources (such as WordNet and JMDict). You can browse what… See the full description on the dataset page: https://huggingface.co/datasets/conceptnet5/conceptnet5.texttext-classification10M<n<100M26 likes4.4k downloads3y agoHugging Face