CoolFace
Datasetpublic

shailja/Verilog_GitHub

VeriGen Dataset Summary The dataset comprises Verilog modules as entries. The entries were retrieved from the GitHub dataset on BigQuery. For training [models (https://huggingface.co/shailja/fine-tuned-codegen-2B-Verilog)], we filtered entries with no of characters exceeding 20000 and duplicates (exact duplicates ignoring whitespaces). Paper: Benchmarking Large Language Models for Automated Verilog RTL Code Generation Point of Contact: contact@shailja… See the full description on the dataset page: https://huggingface.co/datasets/shailja/Verilog_GitHub.

sourceHugging Facemitupdated 3y agoView on Hugging Face
32likes403downloads

shailja/Verilog_GitHub · main · files are served by the source, never re-hosted here