Huffing Face

mradermacher/Qwen3-VL-8B-Instruct-abliterated-GGUF

apache-2.0 16/16 files local @ 292ee2bd5d0f

HF_ENDPOINT=https://huffingface.co \
  python -c "from huggingface_hub import snapshot_download; snapshot_download('mradermacher/Qwen3-VL-8B-Instruct-abliterated-GGUF')"

Model card

About

static quants of https://huggingface.co/prithivMLmods/Qwen3-VL-8B-Instruct-abliterated-v1

For a convenient overview and download list, visit our model page for this model.

weighted/imatrix quants are available at https://huggingface.co/mradermacher/Qwen3-VL-8B-Instruct-abliterated-i1-GGUF

Usage

If you are unsure how to use GGUF files, refer to one of TheBloke's READMEs for more details, including on how to concatenate multi-part files.

Provided Quants

(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)

LinkTypeSize/GBNotes
GGUFmmproj-Q8_00.9multi-modal supplement
GGUFmmproj-f161.3multi-modal supplement
GGUFQ2_K3.4
GGUFQ3_K_S3.9
GGUFQ3_K_M4.2lower quality
GGUFQ3_K_L4.5
GGUFIQ4_XS4.7
GGUFQ4_K_S4.9fast, recommended
GGUFQ4_K_M5.1fast, recommended
GGUFQ5_K_S5.8
GGUFQ5_K_M6.0
GGUFQ6_K6.8very good quality
GGUFQ8_08.8fast, best quality
GGUFf1616.516 bpw, overkill

Here is a handy graph by ikawrakow comparing some lower-quality quant types (lower is better):

image.png

And here are Artefact2's thoughts on the matter: https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9

FAQ / Model Request

See https://huggingface.co/mradermacher/model_requests for some answers to questions you might have and/or if you want some other model quantized.

Thanks

I thank my company, nethype GmbH, for letting me use its servers and providing upgrades to my workstation to enable this work in my free time.

Files

FileSizeStatus
.gitattributes3 kBhuffed
Qwen3-VL-8B-Instruct-abliterated.IQ4_XS.gguf4.6 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q2_K.gguf3.3 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q3_K_L.gguf4.4 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q3_K_M.gguf4.1 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q3_K_S.gguf3.8 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q4_K_M.gguf5.0 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q4_K_S.gguf4.8 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q5_K_M.gguf5.9 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q5_K_S.gguf5.7 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q6_K.gguf6.7 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.Q8_0.gguf8.7 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.f16.gguf16.4 GBhuffed
Qwen3-VL-8B-Instruct-abliterated.mmproj-Q8_0.gguf752.3 MBhuffed
Qwen3-VL-8B-Instruct-abliterated.mmproj-f16.gguf1.2 GBhuffed
README.md4 kBhuffed
mradermacher/Qwen3-VL-8B-Instruct-abliterated-GGUF โ€” fast open-weights mirror ยท Huffing Face