Qwen/Qwen3.5-4B โ€” Q3_K_S Mixed-Precision GGUF

Mixed-precision GGUF quantization of Qwen/Qwen3.5-4B, generated with gguf-mixed-quant.

How to quantize

pip install gguf-mixed-quant

gguf-mixed-quant \
  --model Qwen/Qwen3.5-4B \
  --preset Q3_K_S \
  --device cuda \
  --dtype float16 \
  --output Qwen3.5-4B-Q3_K_S-mixed.gguf
Downloads last month
5
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

3-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for anazir/Qwen3.5-4B-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(409)
this model