CoolFace
Datasetpublic

archit11/verl-code-corpus-track-a-file-split

archit11/verl-code-corpus-track-a-file-split Repository-specific code corpus extracted from the verl project and split by file for training/evaluation. What is in this dataset Source corpus: data/code_corpus_verl Total files: 214 Train files: 172 Validation files: 21 Test files: 21 File type filter: .py Split mode: file (file-level holdout) Each row has: file_name: flattened source file name text: full file contents Training context This dataset… See the full description on the dataset page: https://huggingface.co/datasets/archit11/verl-code-corpus-track-a-file-split.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes47downloads
8 commits on main
b08f4fd7mo ago

Upload test.jsonl with huggingface_hub

archit11
ec035167mo ago

Upload validation.jsonl with huggingface_hub

archit11
70695657mo ago

Upload train.jsonl with huggingface_hub

archit11
089746c7mo ago

Upload training_metrics.json with huggingface_hub

archit11
7eedadc7mo ago

Upload split_manifest.json with huggingface_hub

archit11
1b659667mo ago

Upload README.md with huggingface_hub

archit11
f268abd7mo ago

Upload split dataset

archit11
47050ae7mo ago

initial commit

archit11