gemma-4-26b-a4b-it-qat-q4.apml — Apertura model bundle

Gemma 4 26B A4B (mixture-of-experts, ~4B active) instruction-tuned, Google's QAT-Q4 weights converted to the Apertura .apml bundle format: MLX-affine 4-bit quantization (group size 64, 8-bit embeddings), tokenizer.json + chat_template.jinja included, runtime mlx, architecture gemma4.

Consumed by Apertura — an Objective-C++ transformer engine + macOS chat app on MLX.

Download

hf download apocryphx/gemma-4-26b-a4b-it-qat-q4-apml --local-dir gemma-4-26b-a4b-it-qat-q4.apml

License

Gemma is provided under and subject to the Gemma Terms of Use. By downloading this model you agree to those terms, including the Gemma Prohibited Use Policy. This repository redistributes a quantized Model Derivative of google/gemma-4-26B-A4B-it-qat-q4_0-unquantized; the same terms and use restrictions apply to it and to any further derivatives.

Downloads last month
46
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for apocryphx/gemma-4-26b-a4b-it-qat-q4-apml