CoolFace
Modelpublic

marianna13/llava-phi-2-3b-GGUF

sourceHugging Facemitupdated 3y agoView on Hugging Face
4likes79downloads
Model Card

Model Card for LLaVa-Phi-2-3B-GGUF

<!-- Provide a quick summary of what the model is/does. -->

Model Details

Model Description

<!-- Provide a longer summary of what this model is. -->

Quantized version of llava-phi-2-3b. Quantization was done using llama.cpp

  • —Developed by: LAION, SkunkworksAI & Ontocord
  • —Model type: LLaVA is an open-source chatbot trained by fine-tuning Phi-2 on GPT-generated multimodal instruction-following data. It is an auto-regressive language model, based on the transformer architecture
  • —Finetuned from model: Phi-2
  • —License: MIT

Model Sources

<!-- Provide the basic links for the model. -->

Usage

make & ./llava-cli -m ../ggml-model-f16.gguf --mmproj ../mmproj-model-f16.gguf --image /path/to/image.jpg

Evaluation

<!-- This section describes the evaluation protocols and provides the results. -->

Benchmarks

ModelParametersSQAGQATextVQAPOPE
LLaVA-1.57.3B68.062.058.385.3
MC-LLaVA-3B3B-49.638.59-
LLaVA-Phi3B68.4-48.685.0
moondream11.6B-56.339.8-
llava-phi-2-3b3B69.051.247.086.0

Image Captioning (MS COCO)

ModelBLEU_1BLEU_2BLEU_3BLEU_4METEORROUGE_LCIDErSPICE
llava-1.5-7b75.859.84533.329.457.7108.823.5
llava-phi-2-3b67.750.535.724.227.052.485.020.7