datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
transformers-synthetic-assets
transformers-synthetic-assets
Synthetic media fixtures for Transformers tests. These assets are generated from prompts or deterministic code and are not derived from third-party source files.
transformers_image_docexample-documents
Example Documents
A small set of example documents across modalities (image, audio, video) for use in Sentence Transformers retrieval snippets and documentation. These are the kinds of files you pass to model.encode_document(...). They can safely be used as examples in your model cards if you don't want to host the example assets in your model repositories themselves.
Contents
File
Modality
doc1.jpg
image (document page)
doc2.jpg
image (document page)… See the full description on the dataset page: https://huggingface.co/datasets/sentence-transformers/example-documents.unsplash-lite
Unsplash Lite
Unsplash Lite Dataset.This dataset contains:
a subset default with about 25k images and related keywords when available. Keywords are eparated by ; and note that we kept only those where the confident score indicated by Unsplash is higher than 90%
a subset embeddings_clip-ViT-B-32 which contains precomputed embeddings of the images via the clip-ViT-B-32 model by OpenAI
a subset embeddings_metaclip-2-worldwide-s16-384 which contains precomputed embeddings of the… See the full description on the dataset page: https://huggingface.co/datasets/sentence-transformers/unsplash-lite.linear-algebra-transformersadversarial-vision-transformersCompact_VLM_filter_data
Filtration-Oriented Image-Caption Dataset
This dataset is created to train a small Vision-Language Model (VLM) that learns in-context criteria to filter noisy web-scale image-caption pairs.
We used the base Qwen2-VL-2B model to fine-tune a filtration-oriented variant, optimized to assess and filter large datasets efficiently. The goal is to build a lightweight VLM that can be deployed locally, reducing dependency on large-scale APIs and minimizing both compute costs and latency.… See the full description on the dataset page: https://huggingface.co/datasets/Dauka-transformers/Compact_VLM_filter_data.ci_outputsfashion-product-images-small
Dataset Card for "fashion-product-images-small"
More Information needed
Data was obtained from here
adversarial-vision-transformers-robustnesstransformersretouch_dataset
