--- license: mit base_model: - ornith-ai/Ornith-1.0-9B base_model_relation: quantized quantized_by: Atomic-Germ pipeline_tag: image-text-to-text tags: - transformers - safetensors - qwen3_5 - image-text-to-text - text-generation - conversational - eval-results - flm - fastflowlm - q4nx - npu2 --- # *IF YOU USE COMMUNITY QWEN MODELS DO NOT UPGRADE TO FLM v1.0.2+* # Ornith-1.0-9B-NPU2 **FastFlowLM Q4NX conversion of [`ornith-ai/Ornith-1.0-9B`](https://huggingface.co/ornith-ai/Ornith-1.0-9B)** for AMD XDNA NPU inference. This repository contains a quantized **Q4NX** port of the model, compiled for the FastFlowLM (FLM) runtime. It is **not** a GGUF file. | Item | Value | |------|-------| | Source model | [`ornith-ai/Ornith-1.0-9B`](https://huggingface.co/ornith-ai/Ornith-1.0-9B) | | Weights | `model.q4nx` (7.11 GB) | | Modality | language / vision | | FLM version | `1.0.2` | | Converted | 2026-08-18 | ## Install and run This repository works with `flm-add`, a small installer that copies the model into the FastFlowLM user directory and registers the tag. It never modifies the system FastFlowLM install. `pip install flm-add` or `uv tool install flm-add` ```bash uv tool install flm-add flm-add Atomic-Germ/Ornith-1.0-9B-NPU2 --family qwen3.5 FLM_CONFIG_PATH="$HOME/.config/flm/model_list.json" FLM_XCLBIN_PATH="$HOME/.config/flm" flm run ornith:9b ``` ## Files | File | Description | |------|-------------| | `model.q4nx` | Quantized weights (Q8_0 / Q4_1 / BF16) | | `config.json` | FLM runtime configuration | | `tokenizer.json` | Tokenizer vocabulary | | `tokenizer_config.json` | Tokenizer configuration | | `chat_template.jinja` | Chat template | | `vision_weight.q4nx` | Vision model | --- ## Source model card See the original model card: [ornith-ai/Ornith-1.0-9B](https://huggingface.co/ornith-ai/Ornith-1.0-9B)