datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DRIFT-TL-Distill-4K
DRIFT-TL-Distill-4K Dataset
This dataset contains multimodal reasoning examples with images and step-by-step thinking processes.
Paper: Directional Reasoning Injection for Fine-Tuning MLLMs
Code/Project Page: https://github.com/WikiChao/DRIFT
Dataset Structure
Each example contains:
messages: Conversation between user and assistant with image references
images: Paths to associated images
Usage
from datasets import load_dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/ChaoHuangCS/DRIFT-TL-Distill-4K.DeepEyes_train_4KFiners-4k-benchmark4k-video-annotations
4K Video Annotations — Shot Segmentation and Camera Motion
This dataset contains 12 frame-accurate shot clips segmented from five short cinematic video sequences. Every clip is paired with a detailed, manually reviewed annotation covering visible content, subject actions, shot scale, camera angle, camera movement, movement direction, stabilization, composition, lighting, color, pacing, transitions, timecodes, and technical properties.
The footage depicts a tense nighttime… See the full description on the dataset page: https://huggingface.co/datasets/LianeMarilin/4k-video-annotations.ModerationBench-4K
ModerationBench
ModerationBench is a benchmark for evaluating content moderation on real-world, multimodal social media content from Bluesky. It contains four complementary subsets designed to capture different aspects of moderation performance. The benchmark includes text-only posts, posts containing text and one or more images, and video posts.
🌐 Project Website
•
💻 Code
•
📄 Paper… See the full description on the dataset page: https://huggingface.co/datasets/ayanmaj/ModerationBench-4K.DeepEyes_train_4Khot-babes-4kGUI-C2-4Kcum-4kexotic-4kbeauty-4kfacials-4k
