CoolFace
20 results

multi-image

trl-internal-testing /zen-multi-imageimagen<1K1 likes8.9k downloads3mo agoHugging FaceBingsu /laion2b_multi_korean_subset_with_image laion2b_multi_korean_subset_with_image img2dataset을 통해 다운로드에 성공한 Bingsu/laion2B-multi-korean-subset 이미지를 정리한 데이터셋입니다. 이미지는 9,800,137장입니다. 이미지는 짧은 쪽 길이가 256이 되도록 리사이즈 되었으며, 품질 100인 webp파일로 다운로드 되었습니다. Usage 1. datasets >>> from datasets import load_dataset >>> dataset = load_dataset("Bingsu/laion2b_multi_korean_subset_with_image", streaming=True, split="train") >>> dataset.features {'image': Image(decode=True, id=None), 'text': Value(dtype='string', id=None)… See the full description on the dataset page: https://huggingface.co/datasets/Bingsu/laion2b_multi_korean_subset_with_image.imagefeature-extraction100K<n<1M6 likes2.5k downloads4y agoHugging Facemolbal /multi_reference_image_editing Multi-Reference Instruction-Based Image Editing Dataset Overview This dataset contains 20,000 high-resolution image pairs and multi-modal instructions designed for training advanced image-to-image editing models. It combines two complementary example types: 10,000 reference-grounded edits, where structural or stylistic changes are driven by up to three provided visual reference images, and 10,000 occlusion-based inpainting/outpainting edits, where the model must… See the full description on the dataset page: https://huggingface.co/datasets/molbal/multi_reference_image_editing.imageimage-to-image10K<n<100K7 likes1.5k downloads3mo agoHugging Facecvis-tmu /vgllm-spar234k-multi-image-vqa-20k-sampleimage10K<n<100K0 likes746 downloads1y agoHugging Faceallenai /Molmo2-MultiImagePoint Molmo2 Multi-Image Pointing This dataset contains multi-image pointing/counting metadata. This dataset is generated by extending PixMo-Points using a semantic grouping algorithm designed to maximize coverage. Molmo2-MultiImagePoint is a part of the Molmo2 dataset collection and was used to provide the multi-image pointing capabilities of the Molmo2 family of models. Quick links: 📃 Paper 🎥 Blog with Videos Columns image_urls: list of image URLs (original source… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-MultiImagePoint.text100K<n<1M2 likes407 downloads9mo agoHugging Facehshjerry0315 /VideoEspresso_train_multi_image VideoEspresso This dataset is the multi-image version. Leaderboard Model Params Frames Overall Narrative Analysis Event Dynamic Preparation Steps Causal Analysis Theme Analysis Contextual Analysis Influence Analysis Role Analysis Interaction Analysis Behavior Analysis Emotion Analysis Cooking Process Traffic Analysis Situation Analysis LLaVA-Video 72B 64 66.3% 68.4% 66.2% 74.5% 62.7% 62.3% 71.6% 62.5% 63.5% 67.7% 63.2% 60.0% 75.5% 76.7% 74.0% LLaVA-OneVision… See the full description on the dataset page: https://huggingface.co/datasets/hshjerry0315/VideoEspresso_train_multi_image.text100K<n<1M0 likes223 downloads1y agoHugging Face