Brianpuze/Qwen2.5-0.5B-Q4_K_M-GGUF

This repo contains GGUF quantized versions of Qwen/Qwen2.5-0.5B using llama.cpp.

Quantized Versions:

Run with llama.cpp

llama-cli --hf-repo Brianpuze/Qwen2.5-0.5B-Q4_K_M-GGUF --hf-file qwen2.5-0.5b-q4_k_m.gguf -p

Downloads last month
16
GGUF
Model size
0.5B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Brianpuze/Qwen2.5-0.5B-Q4_K_M-GGUF

Quantized
(116)
this model