CoolFace
Modelpublic

dispatchAI/Llama-3.2-3B-Instruct-mobile

sourceHugging Facellama3.2updated 3mo agoView on Hugging Face
0likes46downloads
Model Card

Llama 3.2 3B Instruct - Mobile (GGUF)

The sweet spot between size and capability. When 1B isn't enough but you still need mobile compatibility.

PropertyValue
Parameters3.2 billion
Size~2.1 GB
Speed~16 tok/s (S20 FE CPU)
Quality Retention~96%

Best For

  • Complex reasoning on mobile (better than 1B)
  • Long-form content generation
  • Multi-turn conversations with context
  • Advanced RAG pipelines
  • Research assistant applications