Yusiko/little_dataset_Azerbaijani
🇦🇿 Azerbaijani Instruction Dataset 📚 About this dataset This dataset contains 400 Azerbaijani-language instruction–response pairs, created for fine-tuning conversational and educational AI models. It follows the Alpaca-style format with "input" and "output" fields and focuses on clarity, accuracy, and linguistic richness. 🧩 Structure input → The user’s instruction or question (in Azerbaijani) output → The correct and natural Azerbaijani response Each topic includes 100 high-quality… See the full description on the dataset page: https://huggingface.co/datasets/Yusiko/little_dataset_Azerbaijani.
🇦🇿 Azerbaijani Instruction Dataset
📚 About this dataset This dataset contains 400 Azerbaijani-language instruction–response pairs, created for fine-tuning conversational and educational AI models. It follows the Alpaca-style format with "input" and "output" fields and focuses on clarity, accuracy, and linguistic richness.
🧩 Structure input → The user’s instruction or question (in Azerbaijani) output → The correct and natural Azerbaijani response
Each topic includes 100 high-quality examples.
🧠 Topics Included
- 🏛 Tarix – Historical events, ancient states, cultural evolution
- ➗ Riyazi Analiz – Mathematical limits, derivatives, integrals, and theory
- 🤖 Mexatronika və Robototexnika – Sensors, actuators, PID control, kinematics
- 💰 İqtisadiyyatın Əsasları – Supply & demand, inflation, market structures
⚙️ Format Examples {"input": "Atropatena dövləti nə vaxt yaranmışdır?", "output": "Atropatena M.Ö. 4-cü əsrdə yaranmış qədim Azərbaycan dövlətidir."}
{"input": "f(x)=x² funksiyasının törəməsini tapın.", "output": "f'(x)=2x."}
🧾 License MIT 🆓 Open for educational and research use.
💡 Notes
- 100% Azerbaijani language 🇦🇿
- Clean, grammar-checked, and fine-tuning ready
- Ideal for instruction-tuned little LLMs and dialogue-based models
⭐️ “Teaching AI to speak Azerbaijani — one dataset at a time.”
