datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nutuk_soru_cevap
🇹🇷 Nutuk Question-Answer Fine-Tuning Dataset
Dataset Description
This dataset provides high-quality question-answer pairs derived from Nutuk (The Great Speech), the foundational historical text delivered by Mustafa Kemal Atatürk.
It is meticulously designed for fine-tuning Large Language Models (LLMs) on Turkish language tasks, specifically focusing on instruction-following, reading comprehension, and historical question-answering. To ensure high cognitive diversity… See the full description on the dataset page: https://huggingface.co/datasets/kmkarakaya/nutuk_soru_cevap.nutuk-soru-ve-cevaplar-veriseti
Dataset Card for Nutuk Q&A Dataset
Dataset Details
Dataset Description
This dataset contains 7,634 question-answer pairs in Turkish, extracted and curated from Mustafa Kemal Atatürk's famous speech "Nutuk" (The Great Speech). The dataset is designed for fine-tuning Turkish language models to understand and respond to questions about Turkish history, the War of Independence, and Atatürk's thoughts and philosophies.
The dataset presents conversations in a format… See the full description on the dataset page: https://huggingface.co/datasets/erenfazlioglu/nutuk-soru-ve-cevaplar-veriseti.
