CoolFace
Modelpublic

Aqua00/Nemotron-3-Embed-8B-GGUF

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes1.6kdownloads
Model Card

Nemotron-3-Embed-8B-GGUF

GGUF quantizations of nvidia/Nemotron-3-Embed-8B-BF16.

Quantized with llama.cpp using a public Wikitext-2 importance matrix. Queries require the query: prefix and documents require passage: . Embeddings use mean pooling and L2 normalization.

This model is not affiliated with or endorsed by nvidia, mistral, or any other company. These are simply quantized gguf files of the original model.