GPUMODE/pytorch_scrape_inductor_data
This dataset is composed of scraping code off of github containing pytorch code, and then running torch compile on it in order to have pairs of pytorch and triton code. To spot check the data you can run from datasets import load_dataset # Load the dataset ds = load_dataset("GPUMODE/pytorch_scrape_inductor_data") # Get the first row row = ds["train"][0] print("\n=== UUID ===") print(row['uuid']) print("\n=== PYTHON CODE ===") print(row['python_code']) print("\n=== TRITON CODE ===")… See the full description on the dataset page: https://huggingface.co/datasets/GPUMODE/pytorch_scrape_inductor_data.
This repository belongs to GPUMODE on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
