models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
bge-m3-onnx-int8Qwen3-Embedding-0.6B-int8-ovmultilingual-e5-small-int8Octen-Embedding-8B-INT8Octen-Embedding-4B-INT8CodeRankEmbed-onnx-int8SEX_ROLEPLAY-3.2-1B-ao-int8da8wpplx-embed-v1-0.6b-onnx-int8-standardLlama-3-8B-LLM2Vec-ARDY-INT8BAAI-bge-m3-int8Qwen3-Embedding-0.6B-ONNX-INT8all-MiniLM-L6-v2-ct2-int8Octen-Embedding-0.6B-ONNX-INT8-FULLINT8-nemotron-3.5-asr-streaming-0.6bembeddinggemma-int8bloom-deepspeed-inference-int8bge-small-en-v1.5-int8Qwen3-Embedding-0.6B-INT8bge-m3-quantized-int8bge-m3-ct2-int8bge-m3-quantized-int8bge-small-en-v1.5-rag-int8-staticSEX_ROLEPLAY-3.2-1B-ao-int8wo-gs128twitter-int8Qwen3-VL-Embedding-8B-bnb-int8bge-int8bge-m3-openvino-int8embeddinggemma-300m-code-8L-distill-int8multilingual-e5-small-int8-dynamicQwen3-Embedding-0.6B-GPTQ-Int8
