datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AnyWord-3MDataset from AnyText: Multilingual Visual Text Generation And Editing.
Dataset description from Anytext Team:
Currently, there is a relative scarcity of public datasets for text generation tasks, especially those involving non-Latin script languages. To address this, we introduce a large-scale multilingual dataset called AnyWord-3M. The images in this dataset are sourced from Noah-Wukong, LAION-400M, and OCR recognition datasets such as ArT, COCO-Text, RCTW, LSVT, MLT, MTWI, ReCTS, etc. These… See the full description on the dataset page: https://huggingface.co/datasets/stzhao/AnyWord-3M.align-anything
Overview: Align-Anything Dataset
A Comprehensive All-Modality Alignment Dataset with Fine-grained Preference Annotations and Language Feedback.
🏠 Homepage | 🤗 Align-Anything Dataset | 🤗 T2T_Instruction-tuning Dataset | 🤗 TI2T_Instruction-tuning Dataset | 👍 Our Official Code Repo
Our world is inherently multimodal. Humans perceive the world through multiple senses, and Language Models should operate similarly. However, the development of Current Multi-Modality Foundation Models… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/align-anything.AnyEdit
Celebrate! AnyEdit resolved the data alignment with the re-uploading process (but the view filter is not working:(, though it has 25 edit types). You can view the validation split for a quick look. You can also refer to anyedit-split dataset to view and download specific data for each editing type.
Dataset Card for AnyEdit-Dataset
Instruction-based image editing aims to modify specific image elements with natural language instructions. However, current models in this domain often… See the full description on the dataset page: https://huggingface.co/datasets/Bin1117/AnyEdit.anything-v3.0-glazed
Dataset Card for Anything v3.0 Glazed Samples
Dataset Description
Dataset Summary
This dataset contains image samples originally generated by Linaqruf/anything-v3.0
and subsequently processed by Glaze tool.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/anything-v3.0-glazed.describe-anything-dataset
Describe Anything: Detailed Localized Image and Video Captioning
NVIDIA, UC Berkeley, UCSF
Long Lian, Yifan Ding, Yunhao Ge, Sifei Liu, Hanzi Mao, Boyi Li, Marco Pavone, Ming-Yu Liu, Trevor Darrell, Adam Yala, Yin Cui
[Paper] | [Code] | [Project Page] | [Video] | [HuggingFace Demo] | [Model/Benchmark/Datasets] | [Citation]
Dataset Card for Describe Anything Datasets
Datasets used in the training of describe anything models (DAM).
The datasets are in tar files. These… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/describe-anything-dataset.DA3-BENCH
DA3-BENCH: Depth Anything 3 Evaluation Benchmark
This repository contains processed benchmark datasets for evaluating Depth Anything 3 depth estimation and visual geometry models. The datasets are provided in a convenient, ready-to-use format for research and evaluation purposes.
About Depth Anything 3
Depth Anything 3 (DA3) is a state-of-the-art model that predicts spatially consistent geometry from an arbitrary number of visual inputs, with or without known… See the full description on the dataset page: https://huggingface.co/datasets/depth-anything/DA3-BENCH.prof_images_blip__andite-anything-v4.0
Dataset Card for "prof_images_blip__andite-anything-v4.0"
More Information needed
anyedit-splitAnyInsertion
AnyInsertion
Wensong Song
·
Hong Jiang
·
Zongxing Yang
·
Ruijie Quan
·
Yi Yang
Zhejiang University | Harvard University | Nanyang Technological University
News
[2025.5.9] Release new AnyInsertion v1 text- and mask-prompt dataset on HuggingFace.
[2025.5.7] Release AnyInsertion v1 text prompt dataset on HuggingFace.
[2025.4.24] Release AnyInsertion v1 mask prompt dataset on HuggingFace.
Summary
This is the dataset proposed in… See the full description on the dataset page: https://huggingface.co/datasets/WensongSong/AnyInsertion.AnyPatternThe dataset proposed in our paper "AnyPattern: Towards In-context Image Copy Detection".
Please go to Github for the code about how to use this dataset.
Here, we show how to download this dataset.
anypattern_v31
for letter in {a..z}; do
wget https://huggingface.co/datasets/WenhaoWang/AnyPattern/resolve/main/train/anypattern_v31_part_a$letter
done
wget https://huggingface.co/datasets/WenhaoWang/AnyPattern/resolve/main/train/anypattern_v31_part_ba
cat anypattern_v31_part_a{a..z}… See the full description on the dataset page: https://huggingface.co/datasets/WenhaoWang/AnyPattern.AnyInsertion_V1
AnyInsertion
Wensong Song
·
Hong Jiang
·
Zongxing Yang
·
Ruijie Quan
·
Yi Yang
Zhejiang University | Harvard University | Nanyang Technological University
News
[2025.5.9] Release new AnyInsertion v1 text- and mask-prompt dataset on HuggingFace.
[2025.5.7] Release AnyInsertion v1 text prompt dataset on HuggingFace.
[2025.4.24] Release AnyInsertion v1 mask prompt dataset on HuggingFace.
Summary
This is the dataset proposed in… See the full description on the dataset page: https://huggingface.co/datasets/WensongSong/AnyInsertion_V1.imgbedAlign-Anything-Cosisynthetic-driver-monitoring-detection
Synthetic DMS – Driver Monitoring System Dataset by AnywayLabs.ai
Need a custom synthetic dataset for your own road safety detection use case?
This dataset is an open-source sample of our synthetic data generation work at AnywayLabs.
If you're working on:
industrial defect detection
visual inspection
supervised anomaly detection
hard-to-collect defect classes
synthetic data for computer vision training
You can request a custom synthetic dataset here, or email:… See the full description on the dataset page: https://huggingface.co/datasets/anywaylabs/synthetic-driver-monitoring-detection.DA-2K
DA-2K Evaluation Benchmark
Introduction
DA-2K is proposed in Depth Anything V2 to evaluate the relative depth estimation capability. It encompasses eight representative scenarios of indoor, outdoor, non_real, transparent_reflective, adverse_style, aerial, underwater, and object. It consists of 1K diverse high-quality images and 2K precise pair-wise relative depth annotations.
Please refer to our paper for details in constructing this benchmark.
Usage
Please… See the full description on the dataset page: https://huggingface.co/datasets/depth-anything/DA-2K.anyeditflickr30kAlign-Anything-L0pixmo-ask-model-anything
PixMo-AskModelAnything
PixMo-AskModelAnything is an instruction-tuning dataset for vision-language models. It contains human-authored
question-answer pairs about diverse images with long-form answers.
PixMo-AskModelAnything is a part of the PixMo dataset collection and was used to train the Molmo family of models
Quick links:
📃 Paper
🎥 Blog with Videos
Loading
data = datasets.load_dataset("allenai/pixmo-ask-model-anything", split="train")
Data Format… See the full description on the dataset page: https://huggingface.co/datasets/allenai/pixmo-ask-model-anything.Align-Anything-CoccurMilitaryAircraftRecognitionsynthetic-mvtec-ad-defect-detection
Synthetic MVTec AD – Defect Detection Dataset by AnywayLabs.ai
Need a custom synthetic dataset for your own defect detection use case?
This dataset is an open-source sample of our synthetic data generation work at AnywayLabs.
If you're working on:
industrial defect detection
visual inspection
supervised anomaly detection
hard-to-collect defect classes
synthetic data for computer vision training
You can request a custom synthetic dataset here, or email:… See the full description on the dataset page: https://huggingface.co/datasets/anywaylabs/synthetic-mvtec-ad-defect-detection.pixmo-ask-model-anything-imagesBlockchain-Sensitive-Detect-Data
Blockchain-Sensitive-Detect-Data
English README
复旦大学附属儿科医院-区块链敏感信息检测项目的多模态完整测试数据集。
项目仓库:https://github.com/anyangsong/Blockchain-Sensitive-Detect
数据集:https://huggingface.co/datasets/anyangsong/Blockchain-Sensitive-Detect-Data
checkpoints:https://huggingface.co/anyangsong/Blockchain-Sensitive-Detect-Checkpoints
数据以原始文件夹组织,覆盖文本、音频、图像与视频等样本。
该仓库不提供统一的 CSV、Parquet 或 JSONL 清单;类别信息主要由目录名和文件名携带。
内容警告: 数据集包含辱骂、性内容、暴力、政治相关内容、误导性医疗信息、欺诈信息。使用者应仅在具备适当访问控制、伦理审查和当地法律依据的环境中处理这些内容。… See the full description on the dataset page: https://huggingface.co/datasets/anyangsong/Blockchain-Sensitive-Detect-Data.anyskin_slip_detection
Slip Detection Dataset
A Reachy 2's gripper has been equipped with an Anyskin tactile sensor, composed of 5 magnetic sensors. This dataset aims to reproduce the slip detection task described in the paper.
Capture protocol
A list of 16 objects with various hardness levels has been selected (see pictures folder).
The gripper closure was manually triggered. Then the operator pulled the object without removing it from the gripper. Four forces were applied in different… See the full description on the dataset page: https://huggingface.co/datasets/pollen-robotics/anyskin_slip_detection.regs-anythingv3Anything2Real_demo_dataset
Overview
This dataset contains paired reference–target image sets, where each set includes:
The generation is used tori29umai/QwenImageEdit2509_LoRA/QIE_image2rakugaki_V1_dim4_1e-3-000012.safetensors to generate abstract image
A reference image,
A corresponding text caption describing the from reference image to target image,
A target image.
It is designed for tasks such as image editing, style transfer, instruction-following image generation, or controlled image transformation.… See the full description on the dataset page: https://huggingface.co/datasets/TatiTamati/Anything2Real_demo_dataset.anylearning-data
AnyLearning datasets
This repository contains reproducible sample datasets used to develop and test
AnyLearning OSS.
Dataset licenses are recorded in LICENSES.md. The repository's
scripts and original documentation are Apache-2.0, but that license does not
override the terms of any dataset. Check the dataset license before use.
Licence-cleared
Task
Dataset
Licence
Image classification
ZhangLabData: Chest X-Ray
CC BY 4.0
Object detection
Safety Helmet… See the full description on the dataset page: https://huggingface.co/datasets/nrl-ai/anylearning-data.synthetic-suspended-load-detection
Synthetic Suspended Load Detection Dataset by AnywayLabs.ai
Dataset Summary
This dataset contains fully synthetic images for suspended load detection in factory and construction environments, generated by AnywayLabs.ai.
The dataset is designed for supervised object detection of workers and suspended loads, covering multiple camera viewpoints (FOV). It targets real-world workplace safety monitoring and is intended to train models that detect dangerous proximity… See the full description on the dataset page: https://huggingface.co/datasets/anywaylabs/synthetic-suspended-load-detection.trashcan1
