CoolFace
Datasetpublic

iraklixyz/georgian-sft-conversations

Natively Written Georgian SFT Conversations A high-quality, general-purpose Supervised Fine-Tuning (SFT) dataset containing 56,676 rows of natively written multi-turn Georgian conversations. The dataset is specifically designed and formatted to train models for conversational chat, instruction following, and agent-like behaviors in the Georgian language. [!NOTE] As of June 2026, this is the largest cleaned, high-quality, natively written SFT conversation dataset available in… See the full description on the dataset page: https://huggingface.co/datasets/iraklixyz/georgian-sft-conversations.

sourceHugging Facecc-by-sa-4.0updated 4mo agoView on Hugging Face
0likes25downloads
filegeorgian_sft_conversations.parquet66.5 MBdownload

iraklixyz/georgian-sft-conversations · main · files are served by the source, never re-hosted here