fbnhnsl/Preprocessed_Solidity_Dataset_V1
This dataset consists of 4,134 unique Solidity files. The files were gathered from three sources: Etherscan, Github and DISL dataset. Six preprocessing steps were applied: Step 1 "Cleaning": Unnecessary parts such as comments or blank lines were removed from each file. Step 2 "Formatting": Each file was converted with Prettier (and the corresponding Solidity-plugin) so that the final model only generates code in a correct format. Step 3 "Slither Analysis": Each file has been checked for… See the full description on the dataset page: https://huggingface.co/datasets/fbnhnsl/Preprocessed_Solidity_Dataset_V1.
07
1version https://git-lfs.github.com/spec/v12oid sha256:d52f0f72c5644c601da2229f4b8831da01781315a1befd01d61a860dd25e03fc3size 178995394 