Llama.cpp Quantization of Qwen3.5-9B-Base-ZitGen-V2 by lolzinventor

Using llama.cpp release b8914 for quantization.

Original model: lolzinventor/Qwen3.5-9B-Base-ZitGen-V2

Download a file (not the whole branch) from below:

Filename Quant type File Size Description
lolzinventor_Qwen3.5-9B-Base-ZitGen-V2-GGUF-Q4_K_M.gguf Q4_K_M 5,3G Good quality
lolzinventor_Qwen3.5-9B-Base-ZitGen-V2-GGUF-Q8_0.gguf Q8_0 8,9G High quality

Download mmproj

Download mmproj from this repository:

mmproj-lolzinventor_Qwen3.5-9B-Base-ZitGen-V2-GGUF-bf16.gguf

Or use one from Qwen3.5-9B model if you already have that downloaded.

Downloads last month
55
GGUF
Model size
9B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for thirteenbit/lolzinventor_Qwen3.5-9B-Base-ZitGen-V2-GGUF

Quantized
(2)
this model