--- license: mit library_name: fastflowlm tags: - q4nx - npu - amd - ryzen-ai - fastflowlm base_model: XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B # Link to the original model --- # MiMo-V2.6-Distill-Qwen-9B-NPU2 This repository contains a **Q4NX** quantization of [XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B](https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B), converted for hardware-accelerated inference with **FastFlowLM** on AMD Ryzen AI (XDNA2) NPUs. ## Model Details * **Format:** Q4NX (FastFlowLM's native packed-quantization format) * **Original Model:** Qwen 3.5 9B Fine-tune * **Runtime:** FastFlowLM * **Hardware:** AMD Ryzen AI (Strix Point / XDNA2)