datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vibevoice-quran_persian-single-speakervibevoice-gptinformal_persian-single-speakerindian-law-datasetvibench
VIBench
VIBench is a benchmark for measuring vertical integration bias in
direct and agentic code generation. It contains 20 direct scenarios and
20 aligned agentic workflows covering realistic software integration
choices across cloud and API ecosystems. The bundle includes the
benchmark tasks, model metadata, system prompts, provider-proof
artifacts, and the blind detector-audit sample used for validation.
Viewer splits
The dataset viewer is intentionally simplified to… See the full description on the dataset page: https://huggingface.co/datasets/vibench-emnlp26/vibench.Vibe-Coding-InstructVibe-Coding-Instruct-V2
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]… See the full description on the dataset page: https://huggingface.co/datasets/CodeDevX/Vibe-Coding-Instruct-V2.vibench-results
VIBench Results
VIBench Results contains the retained raw generations,
detector-labeled outputs, complete runs, option-order and runtime
ablations, paper-facing summaries, figures, configs, audit files, and
static explorer indices used in the study.
The main paper evaluation covers 13 models, 15,600 direct generations,
and 2,000 agentic runs. Direct and agentic VIB are reported as
scenario-matched, share-normalized differences in affiliated-ecosystem
selection relative to strict… See the full description on the dataset page: https://huggingface.co/datasets/vibench-emnlp26/vibench-results.Gemini-3-Flash-Preview-VIBE
Gemini 3 Flash Preview VIBE
This dataset is our first attempt at an agentic coding SFT dataset.
All of the prompts for this dataset were sourced from MiniMaxAI/VIBE.
Each prompt was given to Gemini 3 Flash Preview with the follow tools and system prompt:
read_file - Read file contents from workspace
write_file - Write content to a file
edit_file - Replace text in a file
list_directory - List files and directories
search_code - Search for patterns in files
run_command - Execute… See the full description on the dataset page: https://huggingface.co/datasets/TeichAI/Gemini-3-Flash-Preview-VIBE.vibepass
VIBEPASS: Can Vibe Coders Really Pass the Vibe Check?
Authors: Srijan Bansal, Jiao Fangkai, Yilun Zhou, Austin Xu, Shafiq Joty, Semih Yavuz
TL;DR: As LLMs shift programming toward human-guided "vibe coding", agentic tools increasingly rely on models to self-diagnose and repair their own subtle faults—a capability central to autonomous software engineering yet never systematically evaluated. VIBEPASS presents the first empirical benchmark that decomposes fault-targeted reasoning into… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/vibepass.Vibe-Coding-Claude-Fable-5vibethinker-3b-jlens-traces
VibeThinker-3B J Lens Traces
The J Lens is a Jacobian lens fitted to
WeiboAI/VibeThinker-3B.
Pick one of the 18 fitted source layers and a token position, and the lens
decodes that residual-stream activation into a ranked list of vocabulary
tokens. Following one position across layers shows how the decoded ranking
changes on the way to the model's final output.
This repository stores saved results from that lens. The companion
model repository
contains the lens weights. The… See the full description on the dataset page: https://huggingface.co/datasets/JacobMolBio/vibethinker-3b-jlens-traces.vibe-coding-planning-dataset
🌌 Vibe Coding Planning Dataset
Project Overview
This dataset represents the cutting edge of Vibe-Driven Development, merging structural rigor with aesthetic intuition. Curated to facilitate high-level project planning and architectural synthesis, it serves as a foundational pillar for next-generation AI orchestration.
Dataset Specifications
Total Pairs: 5,000 unique planning instructions and responses.
Format: JSONL optimized for high-throughput… See the full description on the dataset page: https://huggingface.co/datasets/3amthoughts/vibe-coding-planning-dataset.vibeworlding-kenney-assets-1000
VibeWorlding 增量资产:1,000 件
给协作者的可下载交付包。 1000 件已整理的 Kenney 静态资产;包含 GLB、逐件预览、来源许可、四份合并元数据和保留失败的验证记录。不是 VibeWorlding 官方发布。
下载
下载完整资产包(955.6 MB)
解压即得到与原项目对应的目录。模型和图片与本地已验证版本逐字节相同。文件清单与 SHA-256 见 file_manifest.json,压缩包哈希见 download_manifest.json。
包含什么
1000 个唯一资产 ID、1000 个自包含 GLB、1000 张 PNG 预览。
render_in_blender/assets/item_infos.json 与 glb_corrections.json。
assets_retrieval/data/asset_cards.jsonl 与 standardized_asset_library_with_caption.csv。… See the full description on the dataset page: https://huggingface.co/datasets/anon123312/vibeworlding-kenney-assets-1000.vibeworlding-environments-10
VibeWorlding 场景准备:10 个场景
独立场景源码与验证记录,方便协作下载。 不是作者 V2 原生训练集,也不是统一通过完整物理验证的数据集。场景未与新增资产库提前绑定。
下载和打开
下载完整环境包(67.5 MB)
解压后在包根目录运行 python3 -m http.server 8766 --bind 127.0.0.1,打开 http://127.0.0.1:8766/batch-001-ten/gallery.html 看场景,或 http://127.0.0.1:8766/v2-physics-pilot/index.html 看物理测试回放。无需部署 AI 模型;HF 数据集页面供下载,HTML 预览需上述本地静态服务。
现有验证状态
10/10 通过静态准备检查;8/10 通用 Three.js JSON 重载与单视角画面对照通过。10 个场景都实际运行了 Rapier 20 秒、240 Hz 刚体仿真,5/10 满足本轮全部保守判据。保存了全部失败和逐物体轨迹。… See the full description on the dataset page: https://huggingface.co/datasets/anon123312/vibeworlding-environments-10.VIBE-Prompts-500000x
500,000 agentic prompts generated with GPT-OSS 120b including alot of different domains and programming languages
dataset-viber-image-generation-preference-inference-endpoints-battle-flux
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-image-generation-preference-inference-endpoints-battle-flux.cass-vibe-security-bench
CASS-Vibe Security Bench
19 teacher triples + 20 human-labeled rows + 100 real-CVE rows with reports.
from datasets import load_dataset
ds = load_dataset("bbkdevops/cass-vibe-security-bench", "truth20")
Reproduce: python eval.py (needs the repo's cass/ package for scanners;
numbers must match EVIDENCE.md scoreboard).
vcl-vibebench
VCL VibeBench
Stop trusting benchmark slides. Run it yourself.
Practical AI model prompts from Vibe Coder's Life.
This dataset mirrors the open-source GitHub suites. It is not a blended intelligence leaderboard. There is no LLM judge. 9/12 on Score means nine JavaScript helpers compiled and passed hidden unit tests — not “75% smart.”
Configs
Config
What it is
Rows
fun
Fun 1.0 — short copy-paste prompts
10
dev
Dev 1.1 — coding / debugging prompts
10… See the full description on the dataset page: https://huggingface.co/datasets/kondasviktor/vcl-vibebench.clue-vibes-tracesCodex agent traces for Clue Vibes, a Build Small Hackathon project.
Space link: https://huggingface.co/spaces/build-small-hackathon/clue-vibes
vibetrainbench-toolathlonmcq_safety
MCQ Safety
Merged safety multiple-choice dataset built from SafetyBench test-en, SALAD
Bench MCQ data, and WildGuardMix harm-category data.
Splits
Deterministic random split with seed 42:
split
rows
train
15993
valid
889
test
888
Format
Each JSONL row contains:
prompt: problem plus options formatted as A) ..., B) ...
answer: single boxed option label, e.g. \boxed{C}
source: source dataset name
metadata: JSON-encoded source and normalization… See the full description on the dataset page: https://huggingface.co/datasets/cs-552-2026-vibe-trainers/mcq_safety.vibethinker-3b-finance-sftmini-data-public-version
VibeThinker-3B Finance-Reader — SFT Training Data · PUBLIC-SAFE subset
🟢 This is vibethinker-3b-finance-sftmini-data-public-version — the redistribution-safe slice of the
full vibethinker-3b-finance-sftmini-data
dataset, containing only US-government public-domain sources (SEC EDGAR family + Federal Register).
Same schema, same pipeline, same teacher — just the legally shareable rows. (Currently private; intended to be made public.)
The supervised fine-tuning dataset… See the full description on the dataset page: https://huggingface.co/datasets/BatuhanECB/vibethinker-3b-finance-sftmini-data-public-version.Vibe-Coding-Claude-Fable-5Vibe-Coding-Claude-Fable-5dataset-viber-chat-generation-preference-inference-endpoints-battle
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-chat-generation-preference-inference-endpoints-battle.Vibe-Coding-InstructVibe-Coding-InstructVibe-Coding-InstructVibe-Coding-InstructVibe-Coding-Claude-Fable-5
