How to use from the
Use from the
Transformers library
# Load model directly
from transformers import AutoModel
model = AutoModel.from_pretrained("msievers/gemma-4-E2B-it-qat-q4_0-assistant-GGUF", device_map="auto")
Quick Links

gemma-4-E2B-it-qat-q4_0-assistant-GGUF

Quantized and to be used with ik_llama.cpp

README to be written ...

Downloads last month
324
GGUF
Model size
78M params
Architecture
gemma4_mtp
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for msievers/gemma-4-E2B-it-qat-q4_0-assistant-GGUF