LearningOpt/pie-hq-selfplay-13b
017
13b CodeLlama model finetuned on HQ + self-play augmented data to optimize C++ programs. For more details on how the data was collected and training was done, refer to this webpage.
13b CodeLlama model finetuned on HQ + self-play augmented data to optimize C++ programs. For more details on how the data was collected and training was done, refer to this webpage.