Qwen3 VL 30B thinking gguf

Image to text, and text to text.

quantized models Comparison

Type Bits Quality Description
IQ1 1-bit very Low Minimal footprint; worse than Q2/IQ2
Q2/IQ2 2-bit ๐ŸŸฅ Low Minimal footprint; only for tests
Q3/IQ3 3-bit ๐ŸŸง Lowโ€“Med โ€œMediumโ€ variant
Q4/IQ4 4-bit ๐ŸŸฉ Medโ€“High โ€œMediumโ€ โ€” 4-bit
**Q5 ** 5-bit ๐ŸŸฉ๐ŸŸฉ High Excellent general-purpose quant
**Q6_K ** 6-bit ๐ŸŸฉ๐ŸŸฉ๐ŸŸฉ Very High Almost FP16 quality, larger size
**Q8 ** 8-bit ๐ŸŸฉ๐ŸŸฉ๐ŸŸฉ๐ŸŸฉ Near-lossless baseline
Downloads last month
109
GGUF
Model size
31B params
Architecture
qwen3vlmoe
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for John1604/Qwen3-VL-30B-A3B-Thinking-gguf

Quantized
(31)
this model