CoolFace
Datasetpublicgated

handwoven8588/the-stack-v2-train-xsmol-content

The Stack v2 — 12-Language Resolved Content This dataset provides resolved file content for twelve programming languages, derived from the repository/file identifiers published in bigcode/the-stack-v2-train-full-ids. The upstream dataset ships identifiers only — each file is a pointer into the Software Heritage archive. Here, those identifiers have been resolved to their actual source text so the content is directly usable, with the upstream metadata carried through unchanged.… See the full description on the dataset page: https://huggingface.co/datasets/handwoven8588/the-stack-v2-train-xsmol-content.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
1likes263downloads

No commit history came back for main. The revision may not exist, or the source declined the request.