Qwen 3.8 9B OpenVINO IR Models
Collection
Qwen 3.8 9B by Empero-AI. • 3 items • Updated • 1
This is an OpenVINO NF4 weight-only quantized (Compressed Weights) version of empero-ai/Qwen3.8-9B-Distill, optimized for Intel NPU acceleration and deployment via OpenVINO Model Server (OVMS).
For detailed architecture info, evaluation benchmarks, and original weights, refer to the base repository: **[empero-ai/Qwen3.8-9B-Distill](https://huggingface.co/empero-ai/Qwen3.8-9B-Distill
Reccomend Enviroment and techinical detail:My blog(※Japanese)