CoolFace
Datasetpublic

shailja/Verilog_GitHub

VeriGen Dataset Summary The dataset comprises Verilog modules as entries. The entries were retrieved from the GitHub dataset on BigQuery. For training [models (https://huggingface.co/shailja/fine-tuned-codegen-2B-Verilog)], we filtered entries with no of characters exceeding 20000 and duplicates (exact duplicates ignoring whitespaces). Paper: Benchmarking Large Language Models for Automated Verilog RTL Code Generation Point of Contact: contact@shailja… See the full description on the dataset page: https://huggingface.co/datasets/shailja/Verilog_GitHub.

sourceHugging Facemitupdated 3y agoView on Hugging Face
32likes401downloads
3 commits on main
15420103y ago

Update README.md

shailja
25fc6264y ago

Upload Verilog_bigquery_GitHub.csv

shailja
bbc77684y ago

initial commit

shailja