CoolFace
20 results

languagebind

LanguageBind /Open-Sora-Plan-v1.1.0 Annotation We resized the dataset to 1080p for easier uploading. Therefore, the original annotation file might not match the video names. Please refer to this https://github.com/PKU-YuanGroup/Open-Sora-Plan/issues/312#issuecomment-2197312973 Pexels Pexels consists of multiple folders, but each folder exceeds the size limit for Huggingface uploads. Therefore, we divided each folder into 5 parts. You need to merge the 5 parts of each folder first, and then extract each… See the full description on the dataset page: https://huggingface.co/datasets/LanguageBind/Open-Sora-Plan-v1.1.0.text100K<n<1M46 likes108k downloads2y agoHugging FaceLanguageBind /UniWorld-V1 The Geneval-style dataset is sourced from BLIP3o-60k. This dataset is presented in the paper: UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation More details can be found in UniWorld-V1 Data preparation Download the data from LanguageBind/UniWorld-V1. The dataset consists of two parts: source images and annotation JSON files. Prepare a data.txt file in the following format: The first column is the root path to the image. The second… See the full description on the dataset page: https://huggingface.co/datasets/LanguageBind/UniWorld-V1.image1K<n<10K25 likes2.1k downloads1y agoHugging FaceLanguageBind /Open-Sora-Plan-v1.0.0 Open-Sora-Dataset Welcome to the Open-Sora-DataSet project! As part of the Open-Sora-Plan project, we specifically talk about the collection and processing of data sets. To build a high-quality video dataset for the open-source world, we started this project. 💪 We warmly welcome you to join us! Let's contribute to the open-source world together! Thank you for your support and contribution. If you like our project, please give us a star ⭐ on GitHub for latest update.… See the full description on the dataset page: https://huggingface.co/datasets/LanguageBind/Open-Sora-Plan-v1.0.0.text1K<n<10K66 likes924 downloads2y agoHugging FaceLanguageBind /MoE-LLaVA MoE-LLaVA: Mixture of Experts for Large Vision-Language Models If you like our project, please give us a star ⭐ on GitHub for latest update. 📰 News [2024.01.30] The paper is released. [2024.01.27] 🤗Hugging Face demo and all codes & datasets are available now! Welcome to watch 👀 this repository for the latest updates. 😮 Highlights MoE-LLaVA shows excellent performance in multi-modal learning. 🔥 High performance, but with fewer… See the full description on the dataset page: https://huggingface.co/datasets/LanguageBind/MoE-LLaVA.12 likes605 downloads2y agoHugging FaceLanguageBind /Video-LLaVA19 likes425 downloads3y agoHugging FaceLanguageBind /Video-Benchvideo1K<n<10K7 likes227 downloads3y agoHugging Face