datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Open-Jev
Open-Jev: typed decision datasets
Open-Jev turns a state and a question into a typed decision: a yes/no probability, a distribution over choices, independent label probabilities, or a discrete numeric/ordinal decision. This repository publishes twelve separate, frozen data configs from the Open-Jev project, together with original manifests, exact raw records, source code and reconstruction instructions.
These are controlled, mostly synthetic tasks and reference labels. They are… See the full description on the dataset page: https://huggingface.co/datasets/ZefanCai/Open-Jev.Open-Jev-v1.1
Open-Jev v1.1 data
This release publishes 326,619 typed decision records, including 147,139 training records, from the frozen community-hard-mix-v2-final mixture used for the new Open-Jev 27B training stage. It is a redistributable projection, not the exact complete training or evaluation dataset: 2,053 Wikispeedia records are omitted because a separate redistribution license for the archived graph/path data has not been verified.
Project website · Code · Previous data release ·… See the full description on the dataset page: https://huggingface.co/datasets/ZefanCai/Open-Jev-v1.1.APUS-OpenJev-Eval-Frozen80
APUS-OpenJev Eval Frozen80
Private, fixed development evaluation panel: 80 decisions from five task families, 79 parent groups. This is the same panel previously used for official Jev, xDAN-openJet low/high, Laya, and checkpoint comparisons. This publication packages existing data without resampling, changing labels, candidate order, or input content.
This is not an independent test set. It has repeatedly informed debugging and checkpoint selection. Repeated runs measure… See the full description on the dataset page: https://huggingface.co/datasets/apus-ailab/APUS-OpenJev-Eval-Frozen80.OpenJev-Vision-Research-v0.1
OpenJev Vision Research v0.1
12,832 image records, with public provenance, original synthetic scenes,
and programmatically derived decision questions.
This is an experimental research dataset for visual posterior learning and
compositional decisions, released with OpenJev.
It is not a reproduction of TypeSafe's proprietary Jev model or training method.
Three separate configurations
Config
Images
What the labels mean
License
synthetic
8,192
Exact… See the full description on the dataset page: https://huggingface.co/datasets/IamBusy/OpenJev-Vision-Research-v0.1.
