sammybow/luth-sft
Dataset Details This dataset includes all the data used to fine-tune Luth-0.6B-Instruct and Luth-1.7B-Instruct, enhancing their French capabilities on tasks such as instruction following, mathematics, and general knowledge. The models also improved in English thanks to knowledge transfer between the two languages. It contains ~338M tokens in French. Our data scripts are available on GitHub. Dataset Sources Scholar By Kurakura AI: Dataset… See the full description on the dataset page: https://huggingface.co/datasets/sammybow/luth-sft.
045
Duplicate from kurakurai/luth-sft
