CoolFace
18 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01shailja /Verilog_GitHub VeriGen Dataset Summary The dataset comprises Verilog modules as entries. The entries were retrieved from the GitHub dataset on BigQuery. For training [models (https://huggingface.co/shailja/fine-tuned-codegen-2B-Verilog)], we filtered entries with no of characters exceeding 20000 and duplicates (exact duplicates ignoring whitespaces). Paper: Benchmarking Large Language Models for Automated Verilog RTL Code Generation Point of Contact: contact@shailja Languages:… See the full description on the dataset page: https://huggingface.co/datasets/shailja/Verilog_GitHub.text100K<n<1M32 likes407 downloads3y agoHugging Face02JayZhang1 /Verilogdata4pretrainCODET5text100K<n<1M12 likes293 downloads3y agoHugging Face03bnadimi /PyraNet-Verilog PyraNet: A Multi-Layered Hierarchical Dataset for Verilog Authors: Bardia Nadimi, Ghali Omar Boutaib, Hao Zheng Paper link: https://arxiv.org/abs/2412.06947. This dataset is built on top of the VeriBest dataset, which won First Place in the LLM4HWDesign contest at the ICCAD 2024 conference. Dataset Summary This dataset, introduced in our paper PyraNet: A Large Scale Hierarchical Verilog Dataset, addresses the limitations of existing Verilog datasets… See the full description on the dataset page: https://huggingface.co/datasets/bnadimi/PyraNet-Verilog.texttext-generation100K<n<1M32 likes134 downloads1y agoHugging Face04emilgoh /verilog-dataset-v3text10K<n<100K5 likes72 downloads3y agoHugging Face05emilgoh /verilog-dataset-v2text10K<n<100K5 likes44 downloads3y agoHugging Face06LYC2004 /mg-verilog-testtextn<1K0 likes40 downloads2y agoHugging Face07wangxinze /Verilog_datatext100K<n<1M6 likes39 downloads3y agoHugging Face08emilgoh /verilog-dataset-smalltext10K<n<100K2 likes34 downloads3y agoHugging Face09develoco /PyraNet-Verilog PyraNet: A Multi-Layered Hierarchical Dataset for Verilog Authors: Bardia Nadimi, Ghali Omar Boutaib, Hao Zheng Paper link: https://arxiv.org/abs/2412.06947. This dataset is built on top of the VeriBest dataset, which won First Place in the LLM4HWDesign contest at the ICCAD 2024 conference. Dataset Summary This dataset, introduced in our paper PyraNet: A Large Scale Hierarchical Verilog Dataset, addresses the limitations of existing Verilog… See the full description on the dataset page: https://huggingface.co/datasets/develoco/PyraNet-Verilog.texttext-generation100K<n<1M0 likes31 downloads3mo agoHugging Face10Confidentssc /Verilog_GitHub VeriGen Dataset Summary The dataset comprises Verilog modules as entries. The entries were retrieved from the GitHub dataset on BigQuery. For training [models (https://huggingface.co/shailja/fine-tuned-codegen-2B-Verilog)], we filtered entries with no of characters exceeding 20000 and duplicates (exact duplicates ignoring whitespaces). Paper: Benchmarking Large Language Models for Automated Verilog RTL Code Generation Point of Contact: contact@shailja Languages:… See the full description on the dataset page: https://huggingface.co/datasets/Confidentssc/Verilog_GitHub.text100K<n<1M0 likes25 downloads8mo agoHugging Face11mrbeniwal /verilog001text10K<n<100K0 likes19 downloads10mo agoHugging Face1297kjmin /VeriLogos_Augmented_DatasetThis repository includes the augmented Verilog code dataset utilized in the study titled "Improving LLM-based Verilog Code Generation with Data Augmentation and RL (DATE25)". Please refer to the GitHub link 'https://github.com/97kjmin/VeriLogos' for more details. text1K<n<10K0 likes18 downloads2y agoHugging Face13develoco /Verilog_GitHub VeriGen Dataset Summary The dataset comprises Verilog modules as entries. The entries were retrieved from the GitHub dataset on BigQuery. For training [models (https://huggingface.co/shailja/fine-tuned-codegen-2B-Verilog)], we filtered entries with no of characters exceeding 20000 and duplicates (exact duplicates ignoring whitespaces). Paper: Benchmarking Large Language Models for Automated Verilog RTL Code Generation Point of Contact: contact@shailja… See the full description on the dataset page: https://huggingface.co/datasets/develoco/Verilog_GitHub.text100K<n<1M0 likes16 downloads3mo agoHugging Face14TJ-24 /VerilogDatasettext10K<n<100K0 likes15 downloads7mo agoHugging Face15Vishvjit2001 /Verilog_testtextn<1K0 likes9 downloads1y agoHugging Face16zasdwad /Verilog_GitHub VeriGen Dataset Summary The dataset comprises Verilog modules as entries. The entries were retrieved from the GitHub dataset on BigQuery. For training [models (https://huggingface.co/shailja/fine-tuned-codegen-2B-Verilog)], we filtered entries with no of characters exceeding 20000 and duplicates (exact duplicates ignoring whitespaces). Paper: Benchmarking Large Language Models for Automated Verilog RTL Code Generation Point of Contact: contact@shailja Languages:… See the full description on the dataset page: https://huggingface.co/datasets/zasdwad/Verilog_GitHub.text100K<n<1M0 likes8 downloads6mo agoHugging Face17AdeptOfStroggus /verilog-dataset-v2Модифицированный датасет от emilgoh/verilog-dataset-v2 для обучения нейросетей локально. Основная задача - оптимизировать датасет под возможность в том числе локальной дотренировки моделей (уровня 8b) text10K<n<100K0 likes8 downloads4mo agoHugging Face18dcmberdani /childai_verilog_bigquery_dataset_processedgatedtabular100K<n<1M3 likes1 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.