CoolFace
Datasetpublic

CryptoYogi/vazhi-tamil-sft-v7_0

VAZHI Tamil SFT v7.0 — Optimized for Gemma 3 1B-it SFT dataset for fine-tuning Gemma 3 1B-it as VAZHI, a Tamil AI assistant grounded in Tamil spiritual culture. Dataset Details Total: 4,172 samples (3,754 train + 418 eval) Format: Raw instruction/output JSON (Gemma chat template applied at training time) Avg answer length: 47 words (mobile-appropriate) Target model: google/gemma-3-1b-it Token Distribution Bucket %Tokens Count Purpose… See the full description on the dataset page: https://huggingface.co/datasets/CryptoYogi/vazhi-tamil-sft-v7_0.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes18downloads

CryptoYogi/vazhi-tamil-sft-v7_0 · main · files are served by the source, never re-hosted here