CoolFace
Datasetpublic

Ujjwal-Tyagi/JavaScript-Code-Large

JavaScript-Code-Large JavaScript-Code-Large is a large-scale corpus of JavaScript source code comprising around 5 million JavaScript files. The dataset is designed to support research in large language model (LLM) pretraining, code intelligence, software engineering automation, and program analysis for the JavaScript ecosystem. By providing a high-volume, language-specific corpus, JavaScript-Code-Large enables systematic experimentation in JavaScript-focused model training, domain adaptation… See the full description on the dataset page: https://huggingface.co/datasets/Ujjwal-Tyagi/JavaScript-Code-Large.

sourceHugging Facemitupdated 6mo agoView on Hugging Face
1likes308downloads
java_script_only_0002.jsonl4 linesDownload Raw Back to root
1version https://git-lfs.github.com/spec/v12oid sha256:7a5d349665ffc2eb7979e284a8aa9da88dd12ddb52ecc3ab6834e5d19e30ec2a3size 17175477374