Instructions to use apocryphx/gemma-4-E2B-it-q4-apml with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use apocryphx/gemma-4-E2B-it-q4-apml with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir gemma-4-E2B-it-q4-apml apocryphx/gemma-4-E2B-it-q4-apml
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
File size: 1,236 Bytes
3caf3bc | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 | ---
license: gemma
base_model: google/gemma-4-E2B-it
tags:
- apertura
- apml
- mlx
- quantized
- gemma4
---
# gemma-4-E2B-it-q4.apml — Apertura model bundle
Gemma 4 E2B (elastic 2B) instruction-tuned, quantized to the **Apertura
`.apml` bundle format**: MLX-affine 4-bit (group size 64, 8-bit embeddings),
`tokenizer.json` + `chat_template.jinja` included, runtime `mlx`,
architecture `gemma4`. Post-training quantization of the bf16 release (this
family has no QAT variant).
Consumed by [Apertura](https://github.com/apocryphx/Apertura) — an
Objective-C++ transformer engine + macOS chat app on MLX. Export gate:
`--verify-bundle` (bundle reload == in-memory quantization, argmax-identical).
## Download
```sh
hf download apocryphx/gemma-4-E2B-it-q4-apml --local-dir gemma-4-E2B-it-q4.apml
```
## License
Gemma is provided under and subject to the
[Gemma Terms of Use](https://ai.google.dev/gemma/terms). By downloading this
model you agree to those terms, including the
[Gemma Prohibited Use Policy](https://ai.google.dev/gemma/prohibited_use_policy).
This repository redistributes a quantized **Model Derivative** of
`google/gemma-4-E2B-it`; the same terms and use restrictions apply to it and
to any further derivatives.
|