datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
look-bench
LookBench: A Live and Holistic Fashion Image Retrieval Benchmark
LookBench is a large-scale, open benchmark for fashion image retrieval, designed to evaluate modern vision and vision–language models under realistic, contamination-aware settings. The benchmark emphasizes live data, domain diversity, and holistic retrieval tasks spanning both single-item and outfit-level scenarios.
This dataset accompanies the paper LookBench: A Live and Holistic Open Benchmark for Fashion Image… See the full description on the dataset page: https://huggingface.co/datasets/srpone/look-bench.hm-eval
H&M Fashion Evaluation Dataset
H&M-eval is an evaluation benchmark for fashion image-text retrieval built from the
public H&M Personalized Fashion Recommendations
catalog. It follows the same BEIR-style structure as
ZooClaw-Fashion so models
can be evaluated on both benchmarks with the same code path. Released as part of the data-agent benchmarks served via
zoodata.ai and used by agents on the
ZooClaw platform.
Released alongside ZooClaw-FashionSigLIP2
to test… See the full description on the dataset page: https://huggingface.co/datasets/srpone/hm-eval.Xiang_Handsome_Flux_SRPO_Keye_EN_Captionedzooclaw-fashion-eval
ZooClaw-Fashion Evaluation Dataset
ZooClaw-Fashion is an evaluation benchmark for fashion image-text retrieval, designed to rigorously test cross-modal retrieval models on real-world e-commerce fashion products. It features both zero-shot and in-domain query splits, enabling fine-grained analysis of model generalization. Products are sourced from the multi-brand fashion catalog provided by zoodata.ai — the data-agent stack used by agents on the ZooClaw platform.
Released… See the full description on the dataset page: https://huggingface.co/datasets/srpone/zooclaw-fashion-eval.Xiang_Handsome_Flux_SRPO_ControlNet_Frontal_lens
Example : Frontal lens results using ControlNet
Pose Ref Image
Output Image
SRPO_RL_datasets
SRPO Dataset: Reflection-Aware RL Training Data
This repository provides the multimodal reasoning dataset used in the paper:
SRPO: Enhancing Multimodal LLM Reasoning via Reflection-Aware Reinforcement Learning
We release two versions of the dataset:
39K version (modified_39Krelease.jsonl + images.zip)
Enhanced 47K+ version (47K_release_plus.jsonl + 47K_release_plus.zip)
Both follow the same unified format, containing multimodal (image–text) reasoning data with self-reflection… See the full description on the dataset page: https://huggingface.co/datasets/bruce360568/SRPO_RL_datasets.Xiang_Handsome_Flux_SRPO_Keye_ZH_CaptionedXiang_Handsome_Flux_SRPO_PicsXiang_Handsome_Flux_SRPOMale_Diff_Model_SRPO_Qwen_Image_TryOn_Pair_Repose_White_BGMale_Model_SRPO_Qwen_Image_TryOnsrpoXiang_Handsome_Flux_SRPO_Qwen_Image_AIO_RM_UpperMale_Same_Model_SRPO_Qwen_Image_TryOn_PairMale_Diff_Model_SRPO_Qwen_Image_TryOn_PairMale_Diff_Model_SRPO_Qwen_Image_TryOn_Pair_ReposeMale_Model_Flux_SRPO
