gemma-4-E2B-it-qat-q4_0-assistant-GGUF

Quantized and to be used with ik_llama.cpp

README to be written ...

Downloads last month
324
GGUF
Model size
78M params
Architecture
gemma4_mtp
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for msievers/gemma-4-E2B-it-qat-q4_0-assistant-GGUF