world-models
WorldModel-Stabletoolbench-Llama3.1-8B-i1-GGUFWorldModel-Stabletoolbench-Qwen2.5-7B-i1-GGUFWorldModel-Stabletoolbench-Llama3.1-8B-GGUFWorldModel-Stabletoolbench-Qwen2.5-7B-GGUFWorldModel-Stabletoolbench-Llama3.1-8BWorldModel-Sciworld-Qwen2.5-7BWorldModel-Stabletoolbench-Qwen2.5-7BWorldModel-Sciworld-Llama3.1-8B
Hybrid_Neural_World_Models
Hybrid Neural World Models
Training, validation, test, and out-of-distribution (OOD) trajectories for the
three physical systems used in Hybrid Neural World Models (Pranav Lakshmanan,
Paras Chopra). The accompanying code, checkpoints, and paper define a single
neural surrogate that predicts states at any horizon plus a step-doubling
trust signal that flags when its forecasts can be trusted.
This repo contains the raw trajectory data only. Models / training code live
separately.… See the full description on the dataset page: https://huggingface.co/datasets/PraLak/Hybrid_Neural_World_Models.repro-causal-jepa-learning-world-models-through-object-level-latent-masking-traces
Agent traces
Agent sessions published from a Trackio Logbook.
sparse_world_modelsso101-task2-720p-whole-arm-v3-cleanedThis dataset was created using LeRobot.
Dataset Description
Recovered and cleaned SO-101 task 2 dataset. Bad final source episodes 95 and 96 were removed. V3 data, episode metadata, and video shard indices are compact and contiguous.
License: apache-2.0
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 95,
"total_frames": 73988,
"total_tasks": 1,
"chunks_size": 1000… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task2-720p-whole-arm-v3-cleaned.world-models-eval
DreamGrasp: Processed LIBERO Manipulation Demonstrations
Does a robot policy's evaluation still mean something if it never touched a real simulator, only a world model's imagination of one?
This dataset is the shared training data behind that question, a single, ready-to-train release built from LIBERO's manipulation demonstrations (libero_spatial, libero_object, libero_goal). It provides:
Fixed, versioned train / validation / test / held-out splits, so every result trained on… See the full description on the dataset page: https://huggingface.co/datasets/ZaidGhazal/world-models-eval.browser-world-models-transitions
Browser World Models — Transitions
(before screenshot, action, after screenshot) transitions from real websites, for training and
evaluating a world model ("simulator") that predicts the consequence of a web action. Part of
the browser-world-models project.
How it was collected
An LLM policy (gpt-5.4-mini) drove vercel-labs/agent-browser
over the WebVoyager task set (642 tasks / 15 sites); the
task questions are the goals. For every action we saved a screenshot… See the full description on the dataset page: https://huggingface.co/datasets/sudac/browser-world-models-transitions.
