sinjab/jina-reranker-v2-base-multilingual-F16-GGUF
029
jina-reranker-v2-base-multilingual-F16-GGUF
This model was converted to GGUF format from jinaai/jina-reranker-v2-base-multilingual using llama.cpp via the ggml.ai's GGUF-my-repo space.
Refer to the original model card for more details on the model.
Model Information
- Base Model: jinaai/jina-reranker-v2-base-multilingual
- Quantization: F16
- Format: GGUF (GPT-Generated Unified Format)
- Converted with: llama.cpp
Quantization Details
This is a F16 quantization of the original model:
- F16: Full 16-bit floating point - highest quality, largest size
- Q8_0: 8-bit quantization - high quality, good balance
- Q4_K_M: 4-bit quantization with medium quality - smaller size, faster inference
Usage
This model can be used with llama.cpp and other GGUF-compatible inference engines.
# Example using llama.cpp
./llama-rerank -m jina-reranker-v2-base-multilingual-F16.ggufModel Files
Citation
If you use this model, please cite the original model:
# See original model card for citation informationLicense
This model inherits the license from the original model. Please refer to the original model card for license details.
Acknowledgements
- Original model by the authors of jinaai/jina-reranker-v2-base-multilingual
- GGUF conversion via llama.cpp by ggml.ai
- Converted and uploaded by sinjab
