apocryphx commited on
Commit
3caf3bc
·
verified ·
1 Parent(s): d53a20e

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +38 -0
README.md ADDED
@@ -0,0 +1,38 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: gemma
3
+ base_model: google/gemma-4-E2B-it
4
+ tags:
5
+ - apertura
6
+ - apml
7
+ - mlx
8
+ - quantized
9
+ - gemma4
10
+ ---
11
+
12
+ # gemma-4-E2B-it-q4.apml — Apertura model bundle
13
+
14
+ Gemma 4 E2B (elastic 2B) instruction-tuned, quantized to the **Apertura
15
+ `.apml` bundle format**: MLX-affine 4-bit (group size 64, 8-bit embeddings),
16
+ `tokenizer.json` + `chat_template.jinja` included, runtime `mlx`,
17
+ architecture `gemma4`. Post-training quantization of the bf16 release (this
18
+ family has no QAT variant).
19
+
20
+ Consumed by [Apertura](https://github.com/apocryphx/Apertura) — an
21
+ Objective-C++ transformer engine + macOS chat app on MLX. Export gate:
22
+ `--verify-bundle` (bundle reload == in-memory quantization, argmax-identical).
23
+
24
+ ## Download
25
+
26
+ ```sh
27
+ hf download apocryphx/gemma-4-E2B-it-q4-apml --local-dir gemma-4-E2B-it-q4.apml
28
+ ```
29
+
30
+ ## License
31
+
32
+ Gemma is provided under and subject to the
33
+ [Gemma Terms of Use](https://ai.google.dev/gemma/terms). By downloading this
34
+ model you agree to those terms, including the
35
+ [Gemma Prohibited Use Policy](https://ai.google.dev/gemma/prohibited_use_policy).
36
+ This repository redistributes a quantized **Model Derivative** of
37
+ `google/gemma-4-E2B-it`; the same terms and use restrictions apply to it and
38
+ to any further derivatives.