datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VidLayer
VidLayer Dataset
📖 Introduction
We introduce VidLayer, a large-scale dataset specifically designed for layer-aware video generation. VidLayer provides aligned foreground videos, foreground masks, background videos, and raw videos, enabling supervision at both the semantic and structural levels. We believe constructing VidLayer is a foundational step toward layer-wise text-to-video modeling.
🔍 Dataset Details
We organize the VidLayer dataset… See the full description on the dataset page: https://huggingface.co/datasets/kr-cen/VidLayer.MICo-150K
MICo-150K Dataset
🌟 Catalogue
Dataset Details
Data Structure
Gallery: Decompose & Recopmose
Gallery: Human Centric Tasks
Gallery: Object Centric Tasks
Gallery: Human Object Interaction
📖 Introduction
MICo-150K is a large-scale synthetic dataset generated by Nano Banana and Nano Banana Pro, designed to advance open-source models in Multi-Image Composition (MICo).
We fine-tune a diverse set of base models—including Qwen-Image, BAGEL, OmniGen2… See the full description on the dataset page: https://huggingface.co/datasets/kr-cen/MICo-150K.VisReason
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning
VisReason is a large-scale dataset designed to advance visual Chain-of-Thought (CoT)
reasoning in multimodal large language models (MLLMs). Rather than mapping an image
directly to an answer, VisReason supervises a human-like, global-to-local reasoning
process: the model first forms a holistic hypothesis about the scene, then iteratively
zooms into salient regions (areas of interest) to collect fine-grained… See the full description on the dataset page: https://huggingface.co/datasets/kr-cen/VisReason.dictionary_krc_rusQarachay-Malqar - Russian dictionary
ru_kbd_krc_corpus_classificationJF-KRCJF-KRC-3oscar_krc_devJF-KRC-2
