CoolFace
Modelpublic

memoryco-ai/llama-nemotron-rerank-1b-v2-GGUF

sourceHugging Faceupdated 6mo agoView on Hugging Face
1likes29downloads
3 commits on main
af4bc6d6mo ago

fix: replace broken GGUFs with working rerank versions - cls.output.weight tensor fix, classifier labels, add_eos_token

bsneed
c5fa5f96mo ago

Add Q8_0 and Q4_K_M GGUF quantizations of llama-nemotron-rerank-1b-v2

bsneed
cb6b4c76mo ago

initial commit

memoryco