How to use from
Docker Model Runner
docker model run hf.co/msievers/gemma-4-E2B-it-qat-q4_0-assistant-GGUF:Q4_0
Quick Links

gemma-4-E2B-it-qat-q4_0-assistant-GGUF

Quantized and to be used with ik_llama.cpp

README to be written ...

Downloads last month
307
GGUF
Model size
78M params
Architecture
gemma4_mtp
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for msievers/gemma-4-E2B-it-qat-q4_0-assistant-GGUF