CoolFace
Datasetpublic

leopoldmaillard/sceneteract-grpo

SceneTeract GRPO Training Set Action-level feasibility samples for post-training a VLM against a geometric verifier. Each row is one atomic interaction — an image, a prompt, and a label that was measured rather than annotated — ready to drop into TRL's GRPOTrainer. 8,073 samples over 1,132 3D-FRONT living rooms and dining rooms and three agent profiles. from datasets import load_dataset ds = load_dataset("leopoldmaillard/sceneteract-grpo") ds["train"] # 6,473 samples / 905… See the full description on the dataset page: https://huggingface.co/datasets/leopoldmaillard/sceneteract-grpo.

sourceHugging Facecc-by-nc-4.0updated 9d agoView on Hugging Face
0likes96downloads
4 commits on main
2e4e4169d ago

Update README.md

leopoldmaillard
838563a12d ago

Upload README.md with huggingface_hub

leopoldmaillard
7b501fe12d ago

Add files using upload-large-folder tool

leopoldmaillard
517a80e12d ago

initial commit

leopoldmaillard