v-jepa
Datasets
All datasets matching “v-jepa”behavior-1k-2025-challenge-vjepa2-vitg-demo-embeddings
V-JEPA 2 ViT-G Embeddings — BEHAVIOR-1K 2025 Challenge Demos (62h)
Precomputed video embeddings for a 62-hour subsample of the
BEHAVIOR-1K 2025 challenge demonstrations,
extracted with the V-JEPA 2 ViT-g encoder.
The goal is to make downstream experimentation faster and more reproducible by eliminating
repeated video decoding and encoder forward passes — lowering the barrier for teams
without access to large GPU clusters.
Field
Value
Source dataset… See the full description on the dataset page: https://huggingface.co/datasets/quastAI/behavior-1k-2025-challenge-vjepa2-vitg-demo-embeddings.ego10k-vjepa-latents
Ego10k V-JEPA Latents Dataset
This dataset contains compressed, highly-informative Video Joint Embedding Predictive Architecture (V-JEPA) latents extracted from Ego-centric industrial manufacturing videos.
Dataset Structure
The dataset is partitioned into roughly 1GB .parquet chunks using PyArrow.
Data Source and Preprocessing
The latent embeddings in this dataset were systematically extracted from the Ego10k Master Dataset provided by build.ai. The… See the full description on the dataset page: https://huggingface.co/datasets/rookierufus/ego10k-vjepa-latents.bridgev2_vjepa21_latent_shardedVJEPA-LATENTS-L2NORMsokoban-10k-vjepa2-tokenizedlewm-vjepa21-state-moe-shared-residual-wo-tokencl-analysis
LEWM V-JEPA2.1 State-MoE Shared-Residual Analysis
This dataset contains PNG visualizations and adjacent-epoch absolute-delta
summaries for the experiment
lewm_reasoning_vjepa21_vitL_tokens_100tasks_stateMoE_sharedResidual_pair_topkExcess_flat_headaware_woTokenCL_modify.
Contents
metadata.csv: parsed stage, epoch, layer, condition, scope, and image paths.
gallery_manifest.json: manifest consumed by the companion Static HTML Space.
summary.json: aggregate image… See the full description on the dataset page: https://huggingface.co/datasets/zoeloopy/lewm-vjepa21-state-moe-shared-residual-wo-tokencl-analysis.
