GPUMODE/pytorch_scrape_inductor_data
This dataset is composed of scraping code off of github containing pytorch code, and then running torch compile on it in order to have pairs of pytorch and triton code. To spot check the data you can run from datasets import load_dataset # Load the dataset ds = load_dataset("GPUMODE/pytorch_scrape_inductor_data") # Get the first row row = ds["train"][0] print("\n=== UUID ===") print(row['uuid']) print("\n=== PYTHON CODE ===") print(row['python_code']) print("\n=== TRITON CODE ===")… See the full description on the dataset page: https://huggingface.co/datasets/GPUMODE/pytorch_scrape_inductor_data.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face