LIF1014/ptdbench-verl-coding-task-evaluator-dataset
PTDBench dataset snapshot: task_evaluator This repository stores the immutable runtime dataset snapshot for one materialized PTDBench task. It intentionally excludes model weights and training checkpoints. PTDBench family: verl_coding Source evaluation metric: val-core/taco/acc/mean@1 Provenance: Processed from local TACO EASY (drop picture_num != 0); 8368 train / 184 test rows; task-specific bytes are pinned. License: Apache-2.0 The artifact manifest records every hydrated… See the full description on the dataset page: https://huggingface.co/datasets/LIF1014/ptdbench-verl-coding-task-evaluator-dataset.
PTDBench dataset snapshot: task_evaluator
This repository stores the immutable runtime dataset snapshot for one materialized PTDBench task. It intentionally excludes model weights and training checkpoints.
- PTDBench family:
verl_coding - Source evaluation metric:
val-core/taco/acc/mean@1 - Provenance: Processed from local TACO EASY (drop
picture_num != 0); 8368 train / 184 test rows; task-specific bytes are pinned. - License:
Apache-2.0
The artifact manifest records every hydrated runtime path, byte size, and SHA-256. The task Dockerfile pins the repository commit after publication.
