datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
apex-r1-real-world-documents
Apex-R1 Real-World Benchmark Documents
This dataset stores real-world document/data assets collected for Apex-R1 synthetic long-horizon agentic RL workspace generation.
The files are intended as seed workspace materials, not as benchmark task labels. They can be injected into APEX-style filesystem/ or .apps_data/ environments to create more realistic and diverse professional-domain tasks.
Contents
benchmark_documents/
EnterpriseBench/ # CRM invoices… See the full description on the dataset page: https://huggingface.co/datasets/mtybilly/apex-r1-real-world-documents.autoinference-realtime-mix-v1
Autoinference Real-Time Generation Mix v1
This is a prompt set for the real_time_generation serving benchmark. That profile
stands in for medium-context, single-shot interactive traffic: roughly 3000 input
tokens, 100 output tokens, one request at a time with no shared context between
requests. The usual way to run it feeds the server random token IDs of a fixed
length. This dataset keeps the same input and output shape but uses real prompts.
The reason real text matters: random… See the full description on the dataset page: https://huggingface.co/datasets/modal-labs/autoinference-realtime-mix-v1.
