apocryphx's picture
Upload README.md with huggingface_hub
b4a15d8 verified
|
Raw History Blame Contribute Delete
1.14 kB
---
license: gemma
base_model: google/gemma-4-31B-it-qat-q4_0-unquantized
tags:
- apertura
- apml
- mlx
- quantized
- gemma4
---
# gemma-4-31b-it-qat-q4.apml — Apertura model bundle
Gemma 4 31B instruction-tuned, Google's QAT-Q4 weights converted to the
**Apertura `.apml` bundle format**: MLX-affine 4-bit quantization
(group size 64, 8-bit embeddings), `tokenizer.json` +
`chat_template.jinja` included, runtime `mlx`, architecture `gemma4`.
Consumed by [Apertura](https://github.com/apocryphx/Apertura) — an
Objective-C++ transformer engine + macOS chat app on MLX.
## Download
```sh
hf download apocryphx/gemma-4-31b-it-qat-q4-apml --local-dir gemma-4-31b-it-qat-q4.apml
```
## License
Gemma is provided under and subject to the
[Gemma Terms of Use](https://ai.google.dev/gemma/terms). By downloading this
model you agree to those terms, including the
[Gemma Prohibited Use Policy](https://ai.google.dev/gemma/prohibited_use_policy).
This repository redistributes a quantized **Model Derivative** of
`google/gemma-4-31B-it-qat-q4_0-unquantized`; the same terms and use
restrictions apply to it and to any further derivatives.