multi-image
Qwen-Image-Edit-2511-Multiple-Angles-LoRAuform3-image-text-multilingual-baseqwen-image-edit-multiple-angles-GGUFQwen-Image-Edit-2511-Multiple-Angles-LoRAdog-breeds-multiclass-image-classification-with-vitQwen-Image-Edit-2511-Multiple-Angles-LoRAQwen-Image-Edit-2511-Multiple-Angles-LoRAQwen-Image-Edit-2511-Multiple-Angles-LoRA
zen-multi-imagelaion2b_multi_korean_subset_with_image
laion2b_multi_korean_subset_with_image
img2dataset을 통해 다운로드에 성공한 Bingsu/laion2B-multi-korean-subset 이미지를 정리한 데이터셋입니다.
이미지는 9,800,137장입니다.
이미지는 짧은 쪽 길이가 256이 되도록 리사이즈 되었으며, 품질 100인 webp파일로 다운로드 되었습니다.
Usage
1. datasets
>>> from datasets import load_dataset
>>> dataset = load_dataset("Bingsu/laion2b_multi_korean_subset_with_image", streaming=True, split="train")
>>> dataset.features
{'image': Image(decode=True, id=None),
'text': Value(dtype='string', id=None)… See the full description on the dataset page: https://huggingface.co/datasets/Bingsu/laion2b_multi_korean_subset_with_image.multi_reference_image_editing
Multi-Reference Instruction-Based Image Editing Dataset
Overview
This dataset contains 20,000 high-resolution image pairs and multi-modal instructions designed for training advanced image-to-image editing models. It combines two complementary example types: 10,000 reference-grounded edits, where structural or stylistic changes are driven by up to three provided visual reference images, and 10,000 occlusion-based inpainting/outpainting edits, where the model must… See the full description on the dataset page: https://huggingface.co/datasets/molbal/multi_reference_image_editing.vgllm-spar234k-multi-image-vqa-20k-sampleMolmo2-MultiImagePoint
Molmo2 Multi-Image Pointing
This dataset contains multi-image pointing/counting metadata.
This dataset is generated by extending PixMo-Points using a semantic grouping algorithm designed to maximize coverage.
Molmo2-MultiImagePoint is a part of the Molmo2 dataset collection and was used to
provide the multi-image pointing capabilities of the Molmo2 family of models.
Quick links:
📃 Paper
🎥 Blog with Videos
Columns
image_urls: list of image URLs (original source… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-MultiImagePoint.VideoEspresso_train_multi_image
VideoEspresso
This dataset is the multi-image version.
Leaderboard
Model
Params
Frames
Overall
Narrative Analysis
Event Dynamic
Preparation Steps
Causal Analysis
Theme Analysis
Contextual Analysis
Influence Analysis
Role Analysis
Interaction Analysis
Behavior Analysis
Emotion Analysis
Cooking Process
Traffic Analysis
Situation Analysis
LLaVA-Video
72B
64
66.3%
68.4%
66.2%
74.5%
62.7%
62.3%
71.6%
62.5%
63.5%
67.7%
63.2%
60.0%
75.5%
76.7%
74.0%
LLaVA-OneVision… See the full description on the dataset page: https://huggingface.co/datasets/hshjerry0315/VideoEspresso_train_multi_image.
