CoolFace
Modelpublic

TeichAI/Qwen3-4B-Thinking-2507-MiniMax-M2.1-Distill-GGUF

sourceHugging Faceupdated 9mo agoView on Hugging Face
10likes876downloads
Model Card

Qwen3 4B Thinking 2507 - MiniMax M2.1 Distill

This model was trained on a reasoning dataset of MiniMax M2.1.

  • โ€”๐Ÿงฌ Datasets:
  • โ€”TeichAI/MiniMax-M2.1-8800x
  • โ€”๐Ÿ— Base Model:
  • โ€”unsloth/Qwen3-4B-Thinking-2507
  • โ€”⚡ Use cases:
  • โ€”Coding
  • โ€”Science
  • โ€”Deep Research
  • โ€”∑ Stats (Dataset)
  • โ€”Costs: $ 42.94 (USD)
  • โ€”Total tokens (input + output): 39.2 M ---

This qwen3 model was trained 2x faster with Unsloth and Huggingface's TRL library.

An Ollama Modelfile is included for easy deployment.