CoolFace
20 results

vinci

leigangqu /VINCIE-10M Dataset Card for VINCIE-10M VINCIE: Unlocking In-context Image Editing from Video Leigang Qu, Feng Cheng, Ziyan Yang, Qi Zhao, Shanchuan Lin, Yichun Shi, Yicong Li, Wenjie Wang, Tat-Seng Chua, Lu Jiang Dataset Construction Pipeline Visual Transition Annotation. To describe visual transitions between frames, we use chain-of-thought (CoT) prompting to instruct a VLM to perform visual transition… See the full description on the dataset page: https://huggingface.co/datasets/leigangqu/VINCIE-10M.text10M<n<100M12 likes4.7k downloads1y agoHugging FaceDocTron-Hub /VinciCoder-1.6M-SFT VinciCoder: Unified Multimodal Code Generation Dataset This repository contains the datasets used for VinciCoder: Unifying Multimodal Code Generation via Coarse-to-fine Visual Reinforcement Learning, a project that introduces a unified multimodal code generation model. The framework uses a two-stage training approach, comprising a large-scale Supervised Finetuning (SFT) corpus and a Visual Reinforcement Learning (ViRL) dataset. These datasets are designed for tasks involving direct… See the full description on the dataset page: https://huggingface.co/datasets/DocTron-Hub/VinciCoder-1.6M-SFT.textimage-text-to-text1M<n<10M2 likes829 downloads10mo agoHugging FaceDocTron-Hub /VinciCoder-42k-RL VinciCoder: Unifying Multimodal Code Generation via Coarse-to-fine Visual Reinforcement Learning This repository contains the datasets used and generated in the paper VinciCoder: Unifying Multimodal Code Generation via Coarse-to-fine Visual Reinforcement Learning. The work introduces VinciCoder, a unified multimodal code generation model that addresses the limitations of single-task training paradigms. It proposes a two-stage training framework, beginning with a large-scale… See the full description on the dataset page: https://huggingface.co/datasets/DocTron-Hub/VinciCoder-42k-RL.textimage-text-to-text10K<n<100K1 likes103 downloads7mo agoHugging FaceGaviZhou /vincicoder-rl-easyr1 VinciCoder RL EasyR1 Parquet This dataset contains the reinforcement-learning data used for VinciCoder-style multimodal code generation training. The files were converted from the original VinciCoder RL parquet format into an EasyR1-compatible parquet format. The main change is the image column format. In the original files, the images column stores image data as a base64 string. In this version, images is stored as a list of image objects with raw bytes, which matches the format… See the full description on the dataset page: https://huggingface.co/datasets/GaviZhou/vincicoder-rl-easyr1.textimage-to-text10K<n<100K0 likes78 downloads4mo agoHugging FaceCyberHarem /leonardo_da_vinci_fgo Dataset of leonardo_da_vinci/レオナルド・ダ・ヴィンチ/莱昂纳多·达·芬奇 (Fate/Grand Order) This is the dataset of leonardo_da_vinci/レオナルド・ダ・ヴィンチ/莱昂纳多·达·芬奇 (Fate/Grand Order), containing 500 images and their tags. The core tags of this character are long_hair, brown_hair, blue_eyes, parted_bangs, breasts, small_breasts, which are pruned in this dataset. Images are crawled from many sites (e.g. danbooru, pixiv, zerochan ...), the auto-crawling system is powered by DeepGHS Team(huggingface organization).… See the full description on the dataset page: https://huggingface.co/datasets/CyberHarem/leonardo_da_vinci_fgo.text-to-image1K<n<10K0 likes75 downloads3y agoHugging Facevinci00 /ministral-3-benchmark-prompts Ministral 3 MLX benchmark prompts This tiny dataset contains the four fixed prompts used by the reproducible smoke benchmark for the Ministral 3 MLX 4-bit model. It is a benchmark fixture, not a training or fine-tuning dataset. Schema Each JSONL row contains: id: stable case identifier; language: prompt language; prompt: exact input sent to the model; expected_keywords: lowercase substrings used by the smoke check. The benchmark uses greedy decoding and checks… See the full description on the dataset page: https://huggingface.co/datasets/vinci00/ministral-3-benchmark-prompts.texttext-generationn<1K0 likes46 downloads4d agoHugging Face