Ornith-1.0-9B-NPU2 / README.md
Atomic-Germ's picture
Update README.md
95a44a4 verified
|
Raw History Blame Contribute Delete
1.84 kB
---
license: mit
base_model:
- ornith-ai/Ornith-1.0-9B
base_model_relation: quantized
quantized_by: Atomic-Germ
pipeline_tag: image-text-to-text
tags:
- transformers
- safetensors
- qwen3_5
- image-text-to-text
- text-generation
- conversational
- eval-results
- flm
- fastflowlm
- q4nx
- npu2
---
# *IF YOU USE COMMUNITY QWEN MODELS DO NOT UPGRADE TO FLM v1.0.2+*
# Ornith-1.0-9B-NPU2
**FastFlowLM Q4NX conversion of [`ornith-ai/Ornith-1.0-9B`](https://huggingface.co/ornith-ai/Ornith-1.0-9B)** for AMD XDNA NPU inference.
This repository contains a quantized **Q4NX** port of the model, compiled for the FastFlowLM (FLM) runtime. It is **not** a GGUF file.
| Item | Value |
|------|-------|
| Source model | [`ornith-ai/Ornith-1.0-9B`](https://huggingface.co/ornith-ai/Ornith-1.0-9B) |
| Weights | `model.q4nx` (7.11 GB) |
| Modality | language / vision |
| FLM version | `1.0.2` |
| Converted | 2026-08-18 |
## Install and run
This repository works with `flm-add`, a small installer that copies the model
into the FastFlowLM user directory and registers the tag. It never
modifies the system FastFlowLM install.
`pip install flm-add` or `uv tool install flm-add`
```bash
uv tool install flm-add
flm-add Atomic-Germ/Ornith-1.0-9B-NPU2 --family qwen3.5
FLM_CONFIG_PATH="$HOME/.config/flm/model_list.json" FLM_XCLBIN_PATH="$HOME/.config/flm" flm run ornith:9b
```
## Files
| File | Description |
|------|-------------|
| `model.q4nx` | Quantized weights (Q8_0 / Q4_1 / BF16) |
| `config.json` | FLM runtime configuration |
| `tokenizer.json` | Tokenizer vocabulary |
| `tokenizer_config.json` | Tokenizer configuration |
| `chat_template.jinja` | Chat template |
| `vision_weight.q4nx` | Vision model |
---
## Source model card
See the original model card: [ornith-ai/Ornith-1.0-9B](https://huggingface.co/ornith-ai/Ornith-1.0-9B)