datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ROMA_proactive
ROMA Proactive Streaming Dataset
Figure: Overview of ROMA's Streaming Dataset. This repository contains the Proactive subset (Green and Purple sections).
Dataset Summary
This repository contains the Proactive Interaction subset of the dataset introduced in the paper ROMA: Real-time Omni-Multimodal Assistant with Interactive Streaming Understanding.
This dataset is designed to train multimodal models for streaming video understanding, specifically focusing on tasks… See the full description on the dataset page: https://huggingface.co/datasets/EurekaTian/ROMA_proactive.trainproactive-execution-context-review-cases
Proactive Execution Context Review Cases
An original, synthetic teaching dataset for reviewing whether an AI system should prepare a next step, refresh its Context, or return a decision to a person.
What this is
Each record describes a fictional work scenario with an available Session summary, a candidate next step, and an expected review boundary. The material is intentionally small and illustrative; it is not a benchmark, model evaluation, product telemetry… See the full description on the dataset page: https://huggingface.co/datasets/ChengyiX/proactive-execution-context-review-cases.Qwen3.5_Proactive_SFTDeepSeek-R1-Distill-Data-5kReasoning-While-Asking-SFT-DatasetProactiveBenchThis is the benchmark associated with submission 19025 at ICLR 2026.
The benchmark is intended to be used with the proposed submission environments (see the source code).
The .jsonl files do not contain proper image paths but rather image path templates, as each .jsonl entry is a sample, and each sample corresponds to a different environment with its own images.
See the submitted code README for information about dataset downloading and preprocessing, and to re-run the evaluations.
