datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
art-my-vibe-imagesVibeWorlding-V2-Query
VibeWorlding-V2-Query
1000 条用于 3D 场景生成评测/训练的自然语言 Query,按多维度配额(应用领域 / 场景空间 / 地形生态 /
空间布局 / 地理文化 / 气候 / 历史时期 / 视觉风格 / 复杂度)分层合成,覆盖 text_to_3d(900条)和
image_to_3d(100条,附带AI生成的参考图)两种输入模态。
数据结构
数据按比例均分成10份(split_01 ~ split_10,每份100条),每份目录结构:
split_XX/
├── query.json # 100条完整query记录
├── img/ # 该份 image_to_3d 卡片用到的参考图(仅引用到的,已去重)
└── glb_samples/ # 该份用到的glb样例(仅引用到的,已去重,来自kenney-assets-1000)
拆分方式:迭代分层抽样(iterative… See the full description on the dataset page: https://huggingface.co/datasets/yasNing/VibeWorlding-V2-Query.VIBE-Banana-Provibethinker-1.5b-atlas
VibeThinker-1.5B Brain Atlas
This is an internal-mechanics atlas for the 1.5B parameter VibeThinker model. The goal was not to benchmark end-task accuracy, but to map what the network is actually doing with its parameters: where it computes, where it stores behaviorally relevant structure, and which late-layer directions are safe to touch.
What was run
Activation census over 9,523 prompts spanning compliance, reasoning, code, math, multilingual, and refusal-style… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/vibethinker-1.5b-atlas.VIBE-Benchmark
How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing
📖 Dataset Overview
VIBE (Visual Instruction Benchmark for Image Editing) is a systematic benchmark designed to evaluate how well multimodal models follow visual instructions for image editing tasks.Unlike traditional textual instruction-guided editing benchmarks, VIBE emphasizes visual instructions, requiring models to understand spatial… See the full description on the dataset page: https://huggingface.co/datasets/VIBE-Benchmark/VIBE-Benchmark.vibes-mediaVibeEval
Vibe-Eval
A benchmark for evaluating multimodal chat models, including especially challenging examples.
[Link to paper] [Blogpost] [Github]
Dataset
Each example has the following fields:
example_id: a unique ID for the example
category: the category that this example belongs to, either difficulty-normal or difficulty-hard
prompt: the user prompt
reference: a golden reference answer for the prompt
image: an image struct (containing bytes and path keys).… See the full description on the dataset page: https://huggingface.co/datasets/RekaAI/VibeEval.VI-Bench
VI-Bench: Benchmarking Video Language Models via Video Prompt Inversion
Overview
VI-Bench is a benchmark for evaluating Video Language Models (VLMs) through Video Prompt Inversion — the task of reverse-engineering the text prompt used to generate an AI-generated video.
300 unique prompt topics across 3 difficulty levels (Easy / Medium / Hard)
~900 AI-generated videos from Hunyuan-Distill and Wan 2.1
5 evaluation dimensions: Subject, Action, Scene, Style, Camera… See the full description on the dataset page: https://huggingface.co/datasets/wulin222/VI-Bench.VIBE-Seedream4.0vibe-landing-page-arena
Vibe Landing Page Arena
A large-scale human preference dataset for evaluating AI-generated landing page design quality. 36,000 pairwise judgments from 3,492 annotators comparing landing pages generated by Claude Code, Cursor, Lovable, and Replit across 100 prompts and 4 design dimensions.
Overview
Metric
Value
Total judgments
36,000
Unique annotators
3,492
Prompts
100
Business categories
97
Design tones
82
Tools compared
4 (Claude Code, Cursor… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/vibe-landing-page-arena.vibeworlding-kenney-assets-1000
VibeWorlding 增量资产:1,000 件
给协作者的可下载交付包。 1000 件已整理的 Kenney 静态资产;包含 GLB、逐件预览、来源许可、四份合并元数据和保留失败的验证记录。不是 VibeWorlding 官方发布。
下载
下载完整资产包(955.6 MB)
解压即得到与原项目对应的目录。模型和图片与本地已验证版本逐字节相同。文件清单与 SHA-256 见 file_manifest.json,压缩包哈希见 download_manifest.json。
包含什么
1000 个唯一资产 ID、1000 个自包含 GLB、1000 张 PNG 预览。
render_in_blender/assets/item_infos.json 与 glb_corrections.json。
assets_retrieval/data/asset_cards.jsonl 与 standardized_asset_library_with_caption.csv。… See the full description on the dataset page: https://huggingface.co/datasets/anon123312/vibeworlding-kenney-assets-1000.vibeworlding-environments-10
VibeWorlding 场景准备:10 个场景
独立场景源码与验证记录,方便协作下载。 不是作者 V2 原生训练集,也不是统一通过完整物理验证的数据集。场景未与新增资产库提前绑定。
下载和打开
下载完整环境包(67.5 MB)
解压后在包根目录运行 python3 -m http.server 8766 --bind 127.0.0.1,打开 http://127.0.0.1:8766/batch-001-ten/gallery.html 看场景,或 http://127.0.0.1:8766/v2-physics-pilot/index.html 看物理测试回放。无需部署 AI 模型;HF 数据集页面供下载,HTML 预览需上述本地静态服务。
现有验证状态
10/10 通过静态准备检查;8/10 通用 Three.js JSON 重载与单视角画面对照通过。10 个场景都实际运行了 Rapier 20 秒、240 Hz 刚体仿真,5/10 满足本轮全部保守判据。保存了全部失败和逐物体轨迹。… See the full description on the dataset page: https://huggingface.co/datasets/anon123312/vibeworlding-environments-10.dataset-viber-image-generation-preference-inference-endpoints-battle-flux
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-image-generation-preference-inference-endpoints-battle-flux.VibeEvalvibe-design-arena
Vibe Design Arena
The first public pairwise human preference dataset for real-world vibe-coded web applications.
60 apps from the Vibe Coding Showcase were compared pairwise by human annotators who judged which app has better visual design based on screenshots. Every possible pair was evaluated (C(60,2) = 1,770 comparisons), with 30 human votes per pair.
Dataset Summary
Stat
Value
Apps
60
Pairwise comparisons
1,770
Human votes per comparison
30
Total… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/vibe-design-arena.VIBE-GPT-Imagevibe-testing-resultsvibe_testing_resultsvibeeval_greek
Dataset Card for Vibe-Eval Greek
The Vibe-Eval Greek dataset is a benchmark of 269 examples for evaluating multimodal chat models, including especially challenging examples. It has been manually translated into Greek from the VibeEval dataset.
Dataset Details
Each example has the following fields:
media_url: a URL where the file is hosted publicly
example_id: a unique ID for the example
category: the category that this example belongs to, either difficulty-normal or… See the full description on the dataset page: https://huggingface.co/datasets/ilsp/vibeeval_greek.VIBE-Banana-Flashvibe-testing-samplesuseful_vibesvibe-feedbacksvibe-codedQwen-Image-Edit-2509vibe-blending
TTL Paired Image Blending Dataset
This dataset contains paired images and human-readable blend attributes.
Columns
left_image: image column (left input image)
right_image: image column (right input image)
attribute_text: text describing the target blended attribute, labeled by human study
blend_difficulty: float score for blending difficulty, labeled by human study
creative_potential: float score for creative potential, labeled by human study
bucket_difficulty:… See the full description on the dataset page: https://huggingface.co/datasets/huzey/vibe-blending.FLUX2-devVIBE
Anonymous Benchmark Dataset
This repository provides an anonymized dataset release for double-blind review. It includes the benchmark data, metadata, and evaluation resources required to reproduce the main evaluation protocol.
The non-anonymous project page and repository will be linked in the camera-ready version.
VIBE-Seedream4.5OmniGen
