CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01stzhao /AnyWord-3MDataset from AnyText: Multilingual Visual Text Generation And Editing. Dataset description from Anytext Team: Currently, there is a relative scarcity of public datasets for text generation tasks, especially those involving non-Latin script languages. To address this, we introduce a large-scale multilingual dataset called AnyWord-3M. The images in this dataset are sourced from Noah-Wukong, LAION-400M, and OCR recognition datasets such as ArT, COCO-Text, RCTW, LSVT, MLT, MTWI, ReCTS, etc. These… See the full description on the dataset page: https://huggingface.co/datasets/stzhao/AnyWord-3M.imagetext-to-image1M<n<10M17 likes7k downloads2y agoHugging Face02PKU-Alignment /align-anything Overview: Align-Anything Dataset A Comprehensive All-Modality Alignment Dataset with Fine-grained Preference Annotations and Language Feedback. 🏠 Homepage | 🤗 Align-Anything Dataset | 🤗 T2T_Instruction-tuning Dataset | 🤗 TI2T_Instruction-tuning Dataset | 👍 Our Official Code Repo Our world is inherently multimodal. Humans perceive the world through multiple senses, and Language Models should operate similarly. However, the development of Current Multi-Modality Foundation Models… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/align-anything.audioany-to-any10K<n<100K48 likes6.3k downloads1y agoHugging Face03Bin1117 /AnyEdit Celebrate! AnyEdit resolved the data alignment with the re-uploading process (but the view filter is not working:(, though it has 25 edit types). You can view the validation split for a quick look. You can also refer to anyedit-split dataset to view and download specific data for each editing type. Dataset Card for AnyEdit-Dataset Instruction-based image editing aims to modify specific image elements with natural language instructions. However, current models in this domain often… See the full description on the dataset page: https://huggingface.co/datasets/Bin1117/AnyEdit.imagetext-to-image1M<n<10M28 likes3.5k downloads2y agoHugging Face04hanamizuki-ai /anything-v3.0-glazed Dataset Card for Anything v3.0 Glazed Samples Dataset Description Dataset Summary This dataset contains image samples originally generated by Linaqruf/anything-v3.0 and subsequently processed by Glaze tool. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/anything-v3.0-glazed.imageimage-classification10K<n<100K5 likes3.2k downloads3y agoHugging Face05nvidia /describe-anything-dataset Describe Anything: Detailed Localized Image and Video Captioning NVIDIA, UC Berkeley, UCSF Long Lian, Yifan Ding, Yunhao Ge, Sifei Liu, Hanzi Mao, Boyi Li, Marco Pavone, Ming-Yu Liu, Trevor Darrell, Adam Yala, Yin Cui [Paper] | [Code] | [Project Page] | [Video] | [HuggingFace Demo] | [Model/Benchmark/Datasets] | [Citation] Dataset Card for Describe Anything Datasets Datasets used in the training of describe anything models (DAM). The datasets are in tar files. These… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/describe-anything-dataset.imageimage-to-text100K<n<1M59 likes1.8k downloads1y agoHugging Face06depth-anything /DA3-BENCH DA3-BENCH: Depth Anything 3 Evaluation Benchmark This repository contains processed benchmark datasets for evaluating Depth Anything 3 depth estimation and visual geometry models. The datasets are provided in a convenient, ready-to-use format for research and evaluation purposes. About Depth Anything 3 Depth Anything 3 (DA3) is a state-of-the-art model that predicts spatially consistent geometry from an arbitrary number of visual inputs, with or without known… See the full description on the dataset page: https://huggingface.co/datasets/depth-anything/DA3-BENCH.imagedepth-estimation10K<n<100K5 likes1.4k downloads10mo agoHugging Face07sasha /prof_images_blip__andite-anything-v4.0 Dataset Card for "prof_images_blip__andite-anything-v4.0" More Information needed image1K<n<10K0 likes1.3k downloads3y agoHugging Face08Bin1117 /anyedit-splitimagetext-to-image1M<n<10M2 likes1.2k downloads2y agoHugging Face09WensongSong /AnyInsertion AnyInsertion Wensong Song · Hong Jiang · Zongxing Yang · Ruijie Quan · Yi Yang Zhejiang University   |   Harvard University   |   Nanyang Technological University News [2025.5.9] Release new AnyInsertion v1 text- and mask-prompt dataset on HuggingFace. [2025.5.7] Release AnyInsertion v1 text prompt dataset on HuggingFace. [2025.4.24] Release AnyInsertion v1 mask prompt dataset on HuggingFace. Summary This is the dataset proposed in… See the full description on the dataset page: https://huggingface.co/datasets/WensongSong/AnyInsertion.imageimage-to-image10K<n<100K9 likes1.2k downloads1y agoHugging Face10WenhaoWang /AnyPatternThe dataset proposed in our paper "AnyPattern: Towards In-context Image Copy Detection". Please go to Github for the code about how to use this dataset. Here, we show how to download this dataset. anypattern_v31 for letter in {a..z}; do wget https://huggingface.co/datasets/WenhaoWang/AnyPattern/resolve/main/train/anypattern_v31_part_a$letter done wget https://huggingface.co/datasets/WenhaoWang/AnyPattern/resolve/main/train/anypattern_v31_part_ba cat anypattern_v31_part_a{a..z}… See the full description on the dataset page: https://huggingface.co/datasets/WenhaoWang/AnyPattern.imagefeature-extractionn<1K0 likes1.2k downloads11mo agoHugging Face11WensongSong /AnyInsertion_V1 AnyInsertion Wensong Song · Hong Jiang · Zongxing Yang · Ruijie Quan · Yi Yang Zhejiang University   |   Harvard University   |   Nanyang Technological University News [2025.5.9] Release new AnyInsertion v1 text- and mask-prompt dataset on HuggingFace. [2025.5.7] Release AnyInsertion v1 text prompt dataset on HuggingFace. [2025.4.24] Release AnyInsertion v1 mask prompt dataset on HuggingFace. Summary This is the dataset proposed in… See the full description on the dataset page: https://huggingface.co/datasets/WensongSong/AnyInsertion_V1.image100K<n<1M4 likes1k downloads1y agoHugging Face12anyi0791 /imgbedimagen<1K0 likes684 downloads37m agoHugging Face13Changyeli03 /Align-Anything-Cosiimage10K<n<100K0 likes658 downloads2y agoHugging Face14anywaylabs /synthetic-driver-monitoring-detection Synthetic DMS – Driver Monitoring System Dataset by AnywayLabs.ai Need a custom synthetic dataset for your own road safety detection use case? This dataset is an open-source sample of our synthetic data generation work at AnywayLabs. If you're working on: industrial defect detection visual inspection supervised anomaly detection hard-to-collect defect classes synthetic data for computer vision training You can request a custom synthetic dataset here, or email:… See the full description on the dataset page: https://huggingface.co/datasets/anywaylabs/synthetic-driver-monitoring-detection.imageobject-detection1K<n<10K0 likes499 downloads4mo agoHugging Face15depth-anything /DA-2K DA-2K Evaluation Benchmark Introduction DA-2K is proposed in Depth Anything V2 to evaluate the relative depth estimation capability. It encompasses eight representative scenarios of indoor, outdoor, non_real, transparent_reflective, adverse_style, aerial, underwater, and object. It consists of 1K diverse high-quality images and 2K precise pair-wise relative depth annotations. Please refer to our paper for details in constructing this benchmark. Usage Please… See the full description on the dataset page: https://huggingface.co/datasets/depth-anything/DA-2K.image1K<n<10K17 likes468 downloads2y agoHugging Face16shenzhebei /anyeditimage1M<n<10M0 likes444 downloads6mo agoHugging Face17AnyModal /flickr30kimage10K<n<100K1 likes428 downloads2y agoHugging Face18Changyeli03 /Align-Anything-L0image10K<n<100K0 likes381 downloads2y agoHugging Face19allenai /pixmo-ask-model-anything PixMo-AskModelAnything PixMo-AskModelAnything is an instruction-tuning dataset for vision-language models. It contains human-authored question-answer pairs about diverse images with long-form answers. PixMo-AskModelAnything is a part of the PixMo dataset collection and was used to train the Molmo family of models Quick links: 📃 Paper 🎥 Blog with Videos Loading data = datasets.load_dataset("allenai/pixmo-ask-model-anything", split="train") Data Format… See the full description on the dataset page: https://huggingface.co/datasets/allenai/pixmo-ask-model-anything.imagevisual-question-answering100K<n<1M8 likes378 downloads2y agoHugging Face20Changyeli03 /Align-Anything-Coccurimage10K<n<100K0 likes369 downloads2y agoHugging Face21anyaeross /MilitaryAircraftRecognitionimage10K<n<100K2 likes336 downloads1y agoHugging Face22anywaylabs /synthetic-mvtec-ad-defect-detection Synthetic MVTec AD – Defect Detection Dataset by AnywayLabs.ai Need a custom synthetic dataset for your own defect detection use case? This dataset is an open-source sample of our synthetic data generation work at AnywayLabs. If you're working on: industrial defect detection visual inspection supervised anomaly detection hard-to-collect defect classes synthetic data for computer vision training You can request a custom synthetic dataset here, or email:… See the full description on the dataset page: https://huggingface.co/datasets/anywaylabs/synthetic-mvtec-ad-defect-detection.imageobject-detectionn<1K1 likes317 downloads4mo agoHugging Face23dnth /pixmo-ask-model-anything-imagesimage100K<n<1M1 likes267 downloads2y agoHugging Face24anyangsong /Blockchain-Sensitive-Detect-Datagated Blockchain-Sensitive-Detect-Data English README 复旦大学附属儿科医院-区块链敏感信息检测项目的多模态完整测试数据集。 项目仓库:https://github.com/anyangsong/Blockchain-Sensitive-Detect 数据集:https://huggingface.co/datasets/anyangsong/Blockchain-Sensitive-Detect-Data checkpoints:https://huggingface.co/anyangsong/Blockchain-Sensitive-Detect-Checkpoints 数据以原始文件夹组织,覆盖文本、音频、图像与视频等样本。 该仓库不提供统一的 CSV、Parquet 或 JSONL 清单;类别信息主要由目录名和文件名携带。 内容警告: 数据集包含辱骂、性内容、暴力、政治相关内容、误导性医疗信息、欺诈信息。使用者应仅在具备适当访问控制、伦理审查和当地法律依据的环境中处理这些内容。… See the full description on the dataset page: https://huggingface.co/datasets/anyangsong/Blockchain-Sensitive-Detect-Data.audiotext-classificationn<1K0 likes256 downloads21d agoHugging Face25pollen-robotics /anyskin_slip_detection Slip Detection Dataset A Reachy 2's gripper has been equipped with an Anyskin tactile sensor, composed of 5 magnetic sensors. This dataset aims to reproduce the slip detection task described in the paper. Capture protocol A list of 16 objects with various hardness levels has been selected (see pictures folder). The gripper closure was manually triggered. Then the operator pulled the object without removing it from the gripper. Four forces were applied in different… See the full description on the dataset page: https://huggingface.co/datasets/pollen-robotics/anyskin_slip_detection.imagetabular-classificationn<1K0 likes253 downloads2y agoHugging Face26NickKolok /regs-anythingv3imagen<1K0 likes222 downloads3y agoHugging Face27TatiTamati /Anything2Real_demo_dataset Overview This dataset contains paired reference–target image sets, where each set includes: The generation is used tori29umai/QwenImageEdit2509_LoRA/QIE_image2rakugaki_V1_dim4_1e-3-000012.safetensors to generate abstract image A reference image, A corresponding text caption describing the from reference image to target image, A target image. It is designed for tasks such as image editing, style transfer, instruction-following image generation, or controlled image transformation.… See the full description on the dataset page: https://huggingface.co/datasets/TatiTamati/Anything2Real_demo_dataset.imagen<1K0 likes202 downloads8mo agoHugging Face28nrl-ai /anylearning-data AnyLearning datasets This repository contains reproducible sample datasets used to develop and test AnyLearning OSS. Dataset licenses are recorded in LICENSES.md. The repository's scripts and original documentation are Apache-2.0, but that license does not override the terms of any dataset. Check the dataset license before use. Licence-cleared Task Dataset Licence Image classification ZhangLabData: Chest X-Ray CC BY 4.0 Object detection Safety Helmet… See the full description on the dataset page: https://huggingface.co/datasets/nrl-ai/anylearning-data.imageimage-classification1K<n<10K0 likes201 downloads27d agoHugging Face29anywaylabs /synthetic-suspended-load-detection Synthetic Suspended Load Detection Dataset by AnywayLabs.ai Dataset Summary This dataset contains fully synthetic images for suspended load detection in factory and construction environments, generated by AnywayLabs.ai. The dataset is designed for supervised object detection of workers and suspended loads, covering multiple camera viewpoints (FOV). It targets real-world workplace safety monitoring and is intended to train models that detect dangerous proximity… See the full description on the dataset page: https://huggingface.co/datasets/anywaylabs/synthetic-suspended-load-detection.imageobject-detectionn<1K0 likes157 downloads4mo agoHugging Face30anyaeross /trashcan1image10K<n<100K0 likes155 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.