Instructions to use apocryphx/gemma-4-26b-a4b-it-qat-q4-apml with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use apocryphx/gemma-4-26b-a4b-it-qat-q4-apml with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir gemma-4-26b-a4b-it-qat-q4-apml apocryphx/gemma-4-26b-a4b-it-qat-q4-apml
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
gemma-4-26b-a4b-it-qat-q4.apml — Apertura model bundle
Gemma 4 26B A4B (mixture-of-experts, ~4B active) instruction-tuned, Google's QAT-Q4 weights converted to the
Apertura .apml bundle format: MLX-affine 4-bit quantization
(group size 64, 8-bit embeddings), tokenizer.json +
chat_template.jinja included, runtime mlx, architecture gemma4.
Consumed by Apertura — an Objective-C++ transformer engine + macOS chat app on MLX.
Download
hf download apocryphx/gemma-4-26b-a4b-it-qat-q4-apml --local-dir gemma-4-26b-a4b-it-qat-q4.apml
License
Gemma is provided under and subject to the
Gemma Terms of Use. By downloading this
model you agree to those terms, including the
Gemma Prohibited Use Policy.
This repository redistributes a quantized Model Derivative of
google/gemma-4-26B-A4B-it-qat-q4_0-unquantized; the same terms and use
restrictions apply to it and to any further derivatives.
- Downloads last month
- 46
Quantized
Model tree for apocryphx/gemma-4-26b-a4b-it-qat-q4-apml
Base model
google/gemma-4-26B-A4B