CoolFace
Modelpublic

prithivMLmods/Qwen3-VL-8B-Thinking-abliterated-v1-GGUF

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
1likes790downloads
Model Card

Qwen3-VL-8B-Thinking-abliterated-v1-GGUF

The Qwen3-VL-8B-Thinking-abliterated-v1 from prithivMLmods is an 8B-parameter vision-language model, an abliterated (v1.0) variant of Alibaba's Qwen3-VL-8B-Thinking optimized for uncensored reasoning and captioning across complex, sensitive, nuanced, artistic, technical, or abstract visual/multimodal content while supporting diverse aspect ratios, resolutions, videos, and layouts with Interleaved-MRoPE, 32-language OCR, and 256K+ context length. It bypasses standard content filters to generate factual, descriptive, reasoning-rich outputs with variational detail control—from high-level summaries to intricate chain-of-thought analyses—leveraging the base model's superior visual agent capabilities, spatial perception, long-context video understanding, and STEM reasoning, primarily in English with multilingual prompt adaptability.[attached_file:1 equivalent] Ideal for research in content moderation/red-teaming, creative storytelling, and visual datasets excluded from mainstream models, it uses Transformers/Qwen3VLForConditionalGeneration for GPU inference (16-24GB VRAM) but may produce explicit content unsuitable for moderated production.

Qwen3-VL-8B-Thinking-abliterated-v1

File NameQuant TypeFile SizeFile Link
Qwen3-VL-8B-Thinking-abliterated-v1.IQ4_XS.ggufIQ4_XS4.59 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q2_K.ggufQ2_K3.28 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q3KL.ggufQ3KL4.43 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q3KM.ggufQ3KM4.12 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q3KS.ggufQ3KS3.77 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q4KM.ggufQ4KM5.03 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q4KS.ggufQ4KS4.8 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q5KM.ggufQ5KM5.85 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q5KS.ggufQ5KS5.72 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q6_K.ggufQ6_K6.73 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.Q8_0.ggufQ8_08.71 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.f16.ggufF1616.4 GBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.mmproj-Q8_0.ggufmmproj-Q8_0752 MBDownload
Qwen3-VL-8B-Thinking-abliterated-v1.mmproj-f16.ggufmmproj-f161.16 GBDownload

Quants Usage

(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)

Here is a handy graph by ikawrakow comparing some lower-quality quant types (lower is better):

image.png