continuation
3D-dungeon-crawler-stratified-continuations-v1
Stratified conditional-continuation benchmark (v1)
Complete and frozen.
A frozen evaluation cohort of 4,000 matched pairs (8,000 episodes)
for estimating and comparing
P(M,Y,R | do(X=1), A=1) P(R=1 | do(X=1), A=1)
from a common set of pre-X contexts shared by every model arm and by the
simulator reference. No arm may select a different prefix set.
manifest.jsonl is immutable and audit.json records the result of every
audit required by issue #53, including the exact-quota… See the full description on the dataset page: https://huggingface.co/datasets/osazuwa/3D-dungeon-crawler-stratified-continuations-v1.dynamic_proteins_promise_sf_cluster_continuation
Dynamic Protein Benchmarking MVP
This repository contains a first-pass Python package for benchmarking whether protein structure-generation models recover multiple experimentally observed conformations of the same protein.
The MVP targets unconditioned multistate recovery for BioEmu and Boltz across full-MSA, shallow-MSA, and no-MSA style conditions. The initial implementation emphasizes reproducible data structures, evaluation metrics, deterministic outputs, and a notebook UI… See the full description on the dataset page: https://huggingface.co/datasets/archiitecture/dynamic_proteins_promise_sf_cluster_continuation.ai-research-berkeley-agentic-verification-harness-opt-continuation-2026-09-26
Agentic Verification Meta-Verifier Traces
This public, manually gated Dataset repository stores immutable phase snapshots
from Meta-Verifier experiments.
Each run is organized as:
experiments/<theme>/<method>/<run>/phases/
train/ # Solver, delegated-verifier, Proposer/Reflector, harness population
val/ # Full validation traces, metrics, and selected frozen harness
heldout/ # Claimed held-out stage, all scheduled cells, and final scores
Access requests are… See the full description on the dataset page: https://huggingface.co/datasets/MinjaeLee-FuriosaAI-Ext/ai-research-berkeley-agentic-verification-harness-opt-continuation-2026-09-26.qasr-arabic-speech-continuationsdsg-state-continuation
DSG State-Continuation
Training data for graph-conditioned long-form fiction generation: given the
narrative state a reader would hold after chapters 1..t-1, and a one-line brief
for chapter t, write chapter t.
Built from 215 public-domain novels (Project Gutenberg, English fiction),
segmented into chapters. 6,876 examples.
Why the state is built this way
The state is not a summary and not a retrieval index. It is a revision-aware
assertion store built causally —… See the full description on the dataset page: https://huggingface.co/datasets/GOVINDFROM/dsg-state-continuation.qwen_continuation_dataset
Qwen Continuation Dataset
Generated with qwen_continuation_dataset.
Statistics
Shards
42
Examples
235
Shard size
1
Updated
2026-07-05 14:47 UTC
Usage
from datasets import load_dataset
ds = load_dataset("Zhuzhik/qwen_continuation_dataset")
ds = load_dataset("Zhuzhik/qwen_continuation_dataset", streaming=True)
Fields
Field
Description
source_id
source document ID
source_name
source dataset (fineweb… See the full description on the dataset page: https://huggingface.co/datasets/Zhuzhik/qwen_continuation_dataset.
