datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Uni-GUI-Desktop-1
Uni-GUI-Desktop-1
A large-scale desktop GUI agent trajectory dataset, used as part of the training data for UI-MOPD (Multi-platform On-Policy Distillation for Continual GUI Agent Learning).
Dataset Statistics
Metric
Value
Trajectories
2,685
Total Steps
~36K
Platform
Desktop (1920x1080)
Applications
10 categories
Coordinate System
Normalized to [0, 999]
Applications
App
Description
chrome
Web browsing tasks
gimp… See the full description on the dataset page: https://huggingface.co/datasets/UI-MOPD/Uni-GUI-Desktop-1.ShowUI-desktopGithub | arXiv | HF Paper | Spaces | Datasets | Quick Start
ShowUI-desktop-8K is a UI-grounding dataset focused on PC-based grounding, with screenshots and annotations originally sourced from OmniAct.
We utilize GPT-4o to augment the original annotations, enriching them with diverse attributes such as appearance, spatial relationships, and intended functionality.
You can use our rewrite strategy code to augment your own data.
If you find our work helpful, please consider citing our paper.… See the full description on the dataset page: https://huggingface.co/datasets/showlab/ShowUI-desktop.ShowUI_desktop
Desktop Dataset from ShowUI
This is a FiftyOne dataset with 7496 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Voxel51/ShowUI_desktop")
# Launch the App
session = fo.launch_app(dataset)
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/ShowUI_desktop.desktop-accessibility-screenshot-json-dumpsUni-GUI-Desktop-2
Uni-GUI-Desktop-2
A desktop GUI agent trajectory dataset collected on OSWorld environments, used as part of the training data for UI-MOPD (Multi-platform On-Policy Distillation for Continual GUI Agent Learning).
Dataset Statistics
Metric
Value
Trajectories
1,247
Total Steps
~14.8K
Platform
Desktop (1920x1080)
Applications
10 categories
Coordinate System
Normalized to [0, 1000]
Applications
App
Count
Description… See the full description on the dataset page: https://huggingface.co/datasets/UI-MOPD/Uni-GUI-Desktop-2.desktop-pii-210
Desktop PII 210
This directory is a local Hugging Face-compatible dataset package for synthetic
desktop screenshots with expected privacy and utility QA pairs.
The screenshots are generated synthetic desktop scenes. The visible sensitive
values are fictional benchmark strings, not real personal data.
Contents
images/: all 210 generated PNG screenshots.
images/metadata.jsonl: Hugging Face imagefolder metadata, one row per image.
data/train.parquet: the default Hugging… See the full description on the dataset page: https://huggingface.co/datasets/paperboy-ai/desktop-pii-210.gui-agent-desktop-sft-datadesktop-pii-500
Desktop PII 500
This directory is a local Hugging Face-compatible dataset package for synthetic
desktop screenshots with expected privacy and utility QA pairs.
It reuses the existing 210-image Desktop PII package and adds 290 accepted
supplemental screenshots generated from google/gemini-3.5-flash scenario
prompts and gpt-image-2 image generation.
The screenshots are generated synthetic desktop scenes. The visible sensitive
values are fictional benchmark strings, not real… See the full description on the dataset page: https://huggingface.co/datasets/paperboy-ai/desktop-pii-500.easyr1-103k-4MP-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktop-gta-0p0-zeroeasyr1-showui-desktop-only-4k9-omniparser-qwen-tool-call-4MP
easyr1-showui-desktop-only-4k9-omniparser-qwen-tool-call-4MP
This dataset was generated using the EasyR1 grounding dataset pipeline.
Generation Details
Generated on: 2025-08-21 12:51:54 UTC
Script: push_easyr1_to_hf.py
Data directory: /lustre/fsw/portfolios/nvr/users/aawadalla/LLaMA-Factory/data
Parameters Used
Maximum samples: 4888
Image resize (max megapixels): 4.0 MP
Prompt format: gta1_with_resolution
Output format: coordinates
Random seed: 42
Resampling… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-showui-desktop-only-4k9-omniparser-qwen-tool-call-4MP.desktop-ui-captionseasyr1-103k-4MP-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktop-0p2-zeroShowUI-desktopGithub | arXiv | HF Paper | Spaces | Datasets | Quick Start
ShowUI-desktop-8K is a UI-grounding dataset focused on PC-based grounding, with screenshots and annotations originally sourced from OmniAct.
We utilize GPT-4o to augment the original annotations, enriching them with diverse attributes such as appearance, spatial relationships, and intended functionality.
You can use our rewrite strategy code to augment your own data.
If you find our work helpful, please consider citing our paper.… See the full description on the dataset page: https://huggingface.co/datasets/ZhuOnR/ShowUI-desktop.easyr1-10k-hard-qwen7b-easy-gta1-4MP-no-gta1-filter-on-omniact-pc-e-showui-desktop
easyr1-10k-hard-qwen7b-easy-gta1-4MP-no-gta1-filter-on-omniact-pc-e-showui-desktop
This dataset was generated using the EasyR1 grounding dataset pipeline.
Generation Details
Generated on: 2025-08-26 00:22:03 UTC
Script: push_easyr1_to_hf.py
Data directory: /lustre/fsw/portfolios/nvr/users/aawadalla/LLaMA-Factory/data
Parameters Used
Maximum samples: 10000
Image resize (max megapixels): 4.0 MP
Minimum native image resolution: 0.0 MP
Prompt format:… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-10k-hard-qwen7b-easy-gta1-4MP-no-gta1-filter-on-omniact-pc-e-showui-desktop.from-desktop-task01
Video Dataset - task01
Dataset Description
This dataset contains video frames extracted from annotated video segments, along with annotations, transcriptions, and corresponding video clips.
Dataset Structure
frames/ — extracted frames (first frame from each segment)
segments/ — video clips for each annotation interval
annotations/ — original JSON annotation
transcriptions/ — transcription files (full_transcription.txt + per segment)
dataset.csv — mapping… See the full description on the dataset page: https://huggingface.co/datasets/Quazitron420/from-desktop-task01.desktop-blipdesktop-git-large-textcapsdesktop-blip-largedesktop-giteasyr1-103k-4MP-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktop-gta-0p2-zeroeasyr1-103k-4MP-stage-three-temp-1_7-RL-ui-vision-jedi-show-ui-desktop-gta-dense-rewarddesktop-git-largeremote_desktop_sceenshotsimnet1k_desktop_computereasyr1-103k-4MP-not-all-correct-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktopeasyr1-103k-4MP-stage-three-temp-1_7-RL-ui-vision-jedi-show-ui-desktop-gta-0p2-zero-dense-rewardDesktopToolsdesktop-element-detection-001
