iraklixyz/georgian-sft-conversations
Natively Written Georgian SFT Conversations A high-quality, general-purpose Supervised Fine-Tuning (SFT) dataset containing 56,676 rows of natively written multi-turn Georgian conversations. The dataset is specifically designed and formatted to train models for conversational chat, instruction following, and agent-like behaviors in the Georgian language. [!NOTE] As of June 2026, this is the largest cleaned, high-quality, natively written SFT conversation dataset available in… See the full description on the dataset page: https://huggingface.co/datasets/iraklixyz/georgian-sft-conversations.
025
