a11y
Datasets
All datasets matching “a11y”A11y-CUA
A11y-CUA Dataset
A11y-CUA is a multimodal desktop interaction dataset for accessibility-focused computer-use agent research. It contains real task trajectories recorded on Windows across two human user groups and two computer use agents (CUAs), each operating under standard and accessibility-specific conditions. Every session captures: timestamped keyboard and mouse events, browser interaction logs, accessibility trees, screen video, and system audio.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/berkeley-hci/A11y-CUA.Reduced-A11y-CUA
Reduced A11y-CUA
A structured-data-only subset of the A11y-CUA dataset on HuggingFace.
What's different from the full dataset
The full A11y-CUA dataset includes screen recordings (screen.mp4), system audio (system_audio.wav), and microphone audio (mic_audio.wav) for every session. These files account for the large majority of the dataset's total size.
This reduced version strips all video and audio files. Every other file is identical and complete:… See the full description on the dataset page: https://huggingface.co/datasets/berkeley-hci/Reduced-A11y-CUA.A11y-CUA
A11y-CUA
Dataset Summary
A11y-CUA is a multimodal computer-use dataset of accessibility-relevant desktop task executions.
Each task folder contains:
structured desktop action logs (*.json)
accessibility tree snapshots (*_a11y_tree.json)
task metadata (metadata_*.json)
screen recording (screen.mp4)
system audio (system_audio.wav)
The dataset includes both human users and model-agent runs under multiple accessibility conditions.
Dataset Structure
Top-level… See the full description on the dataset page: https://huggingface.co/datasets/ananyagm/A11y-CUA.A11YBench
A11YBench
A Benchmark for Web Accessibility Repair
😃Dataset Summary
A11YBench consists of 60 real-world web projects, encompassing 147 web pages and 8,886 accessibility violations detected by the IBM Accessibility Checker using Check Rule 2025.09.03.
The projects vary substantially in size, from 123 to 43,198 source files and from 3,610 to 1,555,532 lines of code, covering both lightweight documentation sites and large production-grade applications.
This scale ensures… See the full description on the dataset page: https://huggingface.co/datasets/LLM4APR/A11YBench.A11y-CUA
A11y-CUA Dataset
A11y-CUA is a multimodal desktop interaction dataset for accessibility-focused computer-use agent research. It contains real task trajectories recorded on Windows across two human user groups and two computer use agents (CUAs), each operating under standard and accessibility-specific conditions. Every session captures: timestamped keyboard and mouse events, browser interaction logs, accessibility trees, screen video, and system audio.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/hamid735/A11y-CUA.Reduced-A11y-CUA
Reduced A11y-CUA
A structured-data-only subset of the A11y-CUA dataset on HuggingFace.
What's different from the full dataset
The full A11y-CUA dataset includes screen recordings (screen.mp4), system audio (system_audio.wav), and microphone audio (mic_audio.wav) for every session. These files account for the large majority of the dataset's total size.
This reduced version strips all video and audio files. Every other file is identical and complete:… See the full description on the dataset page: https://huggingface.co/datasets/Uriel12Agl/Reduced-A11y-CUA.
