nicolasramos's picture
Upload Qwen3.8-4B-Distill-MLX-oQ4e-fp16-mtp via oMLX
40b2840 verified
|
Raw
History Blame
518 Bytes
---
library_name: mlx
tags:
- mlx
- oq
- quantized
---
> [!IMPORTANT]
> This quantization was uploaded on **2026-08-28** and replaces a previous version.
> If you downloaded this model before this date, please re-download for the updated weights.
# Qwen3.8-4B-Distill-MLX-oQ4e-fp16-mtp
This model was quantized using [oQ](https://github.com/jundot/omlx) (oMLX v0.6.3) mixed-precision quantization.
## Quantization details
- **Model type**: qwen3_5
- **Bits**: 4
- **Group size**: 64
- **Format**: MLX safetensors