crawler
Datasets
All datasets matching “crawler”3D-dungeon-crawler-video-v2
3D Dungeon Crawler Video v2
32,000 deterministic 28-second observational Unity episodes.
The canonical split contains 16,000 pretrain, 14,000 training,
1,000 test, and 1,000 evaluation episodes.
Unity renders at 512x288 for supersampling. Videos are stored at
256x144, 30 fps, H.264. Training samples every third frame,
yielding 280 frames and an 18x32 visual-token grid per episode.
manifest.jsonl is authoritative for asset paths. Each record points to one MP4 and one
NPZ… See the full description on the dataset page: https://huggingface.co/datasets/osazuwa/3D-dungeon-crawler-video-v2.3D-dungeon-crawler-stratified-continuations-v1
Stratified conditional-continuation benchmark (v1)
Complete and frozen.
A frozen evaluation cohort of 4,000 matched pairs (8,000 episodes)
for estimating and comparing
P(M,Y,R | do(X=1), A=1) P(R=1 | do(X=1), A=1)
from a common set of pre-X contexts shared by every model arm and by the
simulator reference. No arm may select a different prefix set.
manifest.jsonl is immutable and audit.json records the result of every
audit required by issue #53, including the exact-quota… See the full description on the dataset page: https://huggingface.co/datasets/osazuwa/3D-dungeon-crawler-stratified-continuations-v1.v2-crawler3D-dungeon-crawler-video-v2-leaky-xor-supplementai-crawler-index
AI Crawler Index
150 web crawlers and AI user agents from 74 operators — what each one is for,
what blocking it costs you, and the IP ranges its operator publishes.
Plus a compiled user-agent regex and the union of 1997 IPv4 and 1062 IPv6
prefixes from 15 operator-published range files.
Home: https://www.pathwren.workers.dev/c/huggingface-datasets/ · CC0 · no signup, no key.
What this is, plainly
This is an independent, non-commercial automated project. It is run… See the full description on the dataset page: https://huggingface.co/datasets/pathwren/ai-crawler-index.3D-dungeon-crawler-video
3D Dungeon Crawler Video
Deterministic 23-second Unity episodes for observational world-model training. The 10,000 episodes are split into 8,000 train, 1,000 validation, and 1,000 test episodes.
Unity renders at 512x288 for supersampling; each clips/*.mp4 is area-downscaled and stored at 256x144 and 30 fps. Training samples every third frame, yielding 230 frames and an 18x32 tokenizer grid. Matching arrays/*.npz files contain dag_ticks, action_tokens, and final_state.
All… See the full description on the dataset page: https://huggingface.co/datasets/osazuwa/3D-dungeon-crawler-video.
