datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VenusBench-GD
VenusBench-GD: A Comprehensive Multi-Platform GUI Benchmark for Diverse Grounding Tasks
Project Page: https://ui-venus.github.io/VenusBench-GD/
Introduction
GUI grounding is a critical component in building capable GUI agents. However, existing grounding benchmarks suffer from significant limitations: they either provide insufficient data volume and narrow domain coverage, or focus excessively on a single platform and require highly specialized domain… See the full description on the dataset page: https://huggingface.co/datasets/inclusionAI/VenusBench-GD.VenusBench-CAPTCHA
VenusBench-CAPTCHA: A Real-World CAPTCHA Screenshot–Action Benchmark for GUI Agents
Evaluation Code: https://github.com/inclusionAI/UI-Venus/tree/VenusBench-CAPTCHA
Introduction
CAPTCHA solving is a practical challenge for multimodal GUI agents because it requires more than isolated visual recognition. An agent must understand the challenge instruction, identify the relevant interface region, recognize or reason about the visual target, ground the result… See the full description on the dataset page: https://huggingface.co/datasets/inclusionAI/VenusBench-CAPTCHA.VENUS-10K
Dataset Card for VENUS
Dataset Summary
Data from: Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
@article{kim2025speaking,
title={Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues},
author={Kim, Youngmin and Chung, Jiwan and Kim, Jisoo and Lee, Sunghyun and Lee, Sangkyu and Kim, Junhyeok and Yang, Cheoljong and Yu, Youngjae}… See the full description on the dataset page: https://huggingface.co/datasets/winston1214/VENUS-10K.VenusREMVenus_Case_TempVENUS-5K
Dataset Card for VENUS
Dataset Summary
Data from: Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
@article{kim2025speaking,
title={Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues},
author={Kim, Youngmin and Chung, Jiwan and Kim, Jisoo and Lee, Sunghyun and Lee, Sangkyu and Kim, Junhyeok and Yang, Cheoljong and Yu, Youngjae}… See the full description on the dataset page: https://huggingface.co/datasets/winston1214/VENUS-5K.VENUS-25K
Dataset Card for VENUS
Dataset Summary
Data from: Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
@article{kim2025speaking,
title={Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues},
author={Kim, Youngmin and Chung, Jiwan and Kim, Jisoo and Lee, Sunghyun and Lee, Sangkyu and Kim, Junhyeok and Yang, Cheoljong and Yu, Youngjae}… See the full description on the dataset page: https://huggingface.co/datasets/winston1214/VENUS-25K.VENUS-1K
Dataset Card for VENUS
Dataset Summary
Data from: Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
@article{kim2025speaking,
title={Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues},
author={Kim, Youngmin and Chung, Jiwan and Kim, Jisoo and Lee, Sunghyun and Lee, Sangkyu and Kim, Junhyeok and Yang, Cheoljong and Yu, Youngjae}… See the full description on the dataset page: https://huggingface.co/datasets/winston1214/VENUS-1K.RULER_50
RULER_50 Official-Code Qwen3 Subset
This dataset is a fixed 50-sample-per-group subset of RULER synthetic tasks.
It was generated from the official NVIDIA/RULER GitHub code, not from a
third-party pre-generated mirror.
Official generation source:
Repository: https://github.com/NVIDIA/RULER
Branch: main
Commit: 38da79d79519ef87aa46ae804f838e1eab7f86d7
Generation entrypoint: scripts/data/prepare.py
Benchmark config: scripts/synthetic.yaml
Generation settings:
tokenizer:… See the full description on the dataset page: https://huggingface.co/datasets/VenusChenyy/RULER_50.venus_tempVenusMutHub
VenusMutHub Dataset
VenusMutHub is a comprehensive collection of protein mutation data designed for benchmarking and evaluating protein language models (PLMs) on various mutation effect prediction tasks. This repository contains mutation data across multiple protein properties including enzyme activity, binding affinity, stability, and selectivity.
Dataset Overview
VenusMutHub includes:
Mutation data for hundreds of proteins across diverse functional categories… See the full description on the dataset page: https://huggingface.co/datasets/AI4Protein/VenusMutHub.Venus
Venus: A dataset for fine-grained code generation control
🎉 What is Venus? Venus is the dataset used to train Afterburner (WIP). It is an extension of the original Mercury dataset and currently includes 6 languages: Python3, C++, Javascript, Go, Rust, and Java.
🚧 What is the current progress? We are in the process of expanding the dataset to include more programming languages.
🔮 Why Venus stands out? A key contribution of Venus is that it provides runtime and memory… See the full description on the dataset page: https://huggingface.co/datasets/Elfsong/Venus.shotpath-action-diagnostic-venuslike-eval-20260709# ShotPath Action Diagnostic Venus-like Eval 20260709
This bundle contains the LLM-audited pure-operation diagnostic set for Venus-like evaluation.
Files:
action_diagnostic_pure_operation.jsonl: 908 examples after leakage audit.
images/: image files referenced by the jsonl.
scripts/eval_action_diagnostic_qwen25vl.py: Qwen2.5-VL base/LoRA evaluator.
scripts/run_action_diagnostic_venuslike_eval_server.sh: server runner for base7b, stage1 step200, stage2 step200.
Default server paths in the… See the full description on the dataset page: https://huggingface.co/datasets/purefall/shotpath-action-diagnostic-venuslike-eval-20260709.esa-venus-express-observations
ESA Venus Express Observations
Credit: NASA/JPL-Caltech
Part of a dataset collection on Hugging Face.
Dataset description
Complete observation metadata catalog from the ESA Venus Express mission, which studied Venus from 2006 to 2014.
Venus Express was a European Space Agency mission that studied the Venusian atmosphere, ionosphere, and surface environment from April 2006 until loss of contact in November 2014. It carried a suite of instruments including a… See the full description on the dataset page: https://huggingface.co/datasets/juliensimon/esa-venus-express-observations.Venus_General_TestVenus_PODvenus-cop
Venus-COP Dataset
This repository contains the Venus-COP dataset with multiple captioning methodologies for training and evaluation purposes.
Repository Structure
Folders
Each folder contains the images and captions for different dataset versions:
venus-cop/ - Base dataset (31 files)
venus-cop-b/ - "[Trigger Classifier] Prefix" method version (31 files)
venus-cop-nocap/ - No captions version for recaptioning tests (16 files)
venus-cop-v0d-recontext/ -… See the full description on the dataset page: https://huggingface.co/datasets/mushroomfleet/venus-cop.Venus_CCGVenusLibraryDocumentationVenusX_Res_Act_MF50venus_caseVenus_SFT_DataVenusX_Res_BindB_MP50VenusX_Res_Motif_MP50VenusX_Res_Act_MP50Venus_PCDVenusX_Res_Evo_MP50VenusX_Frag_Evo_MF70details_cloudyu__Venus_DPO_50
Dataset Card for Evaluation run of cloudyu/Venus_DPO_50
Dataset automatically created during the evaluation run of model cloudyu/Venus_DPO_50 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__Venus_DPO_50.VenusX_Res_Dom_MF90
