LearningOpt/pie-hq-selfplay-7b
016
7b CodeLlama model finetuned on HQ + self-play augmented data to optimize C++ programs. For more details on how the data was collected and training was done, refer to this webpage.
7b CodeLlama model finetuned on HQ + self-play augmented data to optimize C++ programs. For more details on how the data was collected and training was done, refer to this webpage.