File size: 1,236 Bytes
3caf3bc
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
---
license: gemma
base_model: google/gemma-4-E2B-it
tags:
- apertura
- apml
- mlx
- quantized
- gemma4
---

# gemma-4-E2B-it-q4.apml — Apertura model bundle

Gemma 4 E2B (elastic 2B) instruction-tuned, quantized to the **Apertura
`.apml` bundle format**: MLX-affine 4-bit (group size 64, 8-bit embeddings),
`tokenizer.json` + `chat_template.jinja` included, runtime `mlx`,
architecture `gemma4`. Post-training quantization of the bf16 release (this
family has no QAT variant).

Consumed by [Apertura](https://github.com/apocryphx/Apertura) — an
Objective-C++ transformer engine + macOS chat app on MLX. Export gate:
`--verify-bundle` (bundle reload == in-memory quantization, argmax-identical).

## Download

```sh
hf download apocryphx/gemma-4-E2B-it-q4-apml --local-dir gemma-4-E2B-it-q4.apml
```

## License

Gemma is provided under and subject to the
[Gemma Terms of Use](https://ai.google.dev/gemma/terms). By downloading this
model you agree to those terms, including the
[Gemma Prohibited Use Policy](https://ai.google.dev/gemma/prohibited_use_policy).
This repository redistributes a quantized **Model Derivative** of
`google/gemma-4-E2B-it`; the same terms and use restrictions apply to it and
to any further derivatives.