tascib/turkish-instruction
Turkish Instruction Dataset Dataset Description This dataset is a large-scale Turkish instruction-tuning corpus created by combining multiple publicly available datasets and applying cleaning and deduplication steps. It is designed for training and evaluating large language models (LLMs) in Turkish. The dataset was prepared as part of a capstone project by students from Sabancı University. Data Sources The dataset is constructed from the following… See the full description on the dataset page: https://huggingface.co/datasets/tascib/turkish-instruction.
2219
No commit history came back for main. The revision may not exist, or the source declined the request.
