CoolFace
Modelpublic

amritansh/replit_3B

sourceHugging Faceotherupdated 3y agoView on Hugging Face
0likes18downloads
Model Card

This is a ggml quantized version of Replit-v2-CodeInstruct-3B. Quantized to 4bit -> q4_1. To run inference you can use ggml directly or ctransformers.

  • Memory usage of model: 2GB~
  • Repo to run the model using ctransformers: https://github.com/abacaj/replit-3B-inference