Instructions to use apocryphx/gemma-4-31b-it-qat-q4-apml with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use apocryphx/gemma-4-31b-it-qat-q4-apml with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir gemma-4-31b-it-qat-q4-apml apocryphx/gemma-4-31b-it-qat-q4-apml
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
gemma-4-31b-it-qat-q4.apml — Apertura model bundle
Gemma 4 31B instruction-tuned, Google's QAT-Q4 weights converted to the
Apertura .apml bundle format: MLX-affine 4-bit quantization
(group size 64, 8-bit embeddings), tokenizer.json +
chat_template.jinja included, runtime mlx, architecture gemma4.
Consumed by Apertura — an Objective-C++ transformer engine + macOS chat app on MLX.
Download
hf download apocryphx/gemma-4-31b-it-qat-q4-apml --local-dir gemma-4-31b-it-qat-q4.apml
License
Gemma is provided under and subject to the
Gemma Terms of Use. By downloading this
model you agree to those terms, including the
Gemma Prohibited Use Policy.
This repository redistributes a quantized Model Derivative of
google/gemma-4-31B-it-qat-q4_0-unquantized; the same terms and use
restrictions apply to it and to any further derivatives.
- Downloads last month
- 19
Hardware compatibility
Log In to add your hardware
Quantized
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support
Model tree for apocryphx/gemma-4-31b-it-qat-q4-apml
Base model
google/gemma-4-31B Finetuned
google/gemma-4-31B-it