datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
INSTRUCT_JEV
INSTRUCT_JEV
INSTRUCT_JEV is an instruction corpus built from the TypeSafe AI documentation
for Jev, the first System One model. It is structured around the three TypeSafe
question primitives - Choice, Noul and Score - and mirrors the raw corpus
captured in deckerGUI-jev_corpus_RAW.
Credits
INSTRUCT_JEV is a DeckerGUI project and exists because of the work below.
Who
Contribution
Link
TypeSafe AI
Jev - the first System One model - and the Choice / Noul… See the full description on the dataset page: https://huggingface.co/datasets/ctaxnagomi/INSTRUCT_JEV.jev-bench
jev-bench
A small multiple-choice set for measuring a model that returns the probability of each option
instead of writing an answer — the Jev / TypeSafe System One style of API, where a request carries
a state and a question with named options and the response carries a distribution over them.
Ordinary multiple-choice benchmarks score the text a model generates. That says nothing about
whether the probability attached to the answer means anything, which is the whole point of… See the full description on the dataset page: https://huggingface.co/datasets/kishida/jev-bench.
