Make the hardware wording neutral
#1
by subin - opened
README.md
CHANGED
|
@@ -8,7 +8,6 @@ tags:
|
|
| 8 |
- classification
|
| 9 |
- qwen3_5
|
| 10 |
- pytorch
|
| 11 |
-
- rocm
|
| 12 |
---
|
| 13 |
|
| 14 |

|
|
@@ -75,7 +74,7 @@ This downloads the complete model release. The root `config.json` lists the back
|
|
| 75 |
|
| 76 |
## Serve with vLLM Semantic Router
|
| 77 |
|
| 78 |
-
This repository contains model data only. Use the vLLM Semantic Router Decision runtime to load `llm-semantic-router/Decision-1.0-Nox-4B` and serve Choice, Noul, and Score requests. The serving implementation and its dependencies live in vLLM Semantic Router; this release does not bundle executable model code. `transformers.AutoModel.from_pretrained` cannot load the custom Decision head directly.
|
| 79 |
|
| 80 |
After configuring a compatible Decision endpoint, send a [SystemOne request](https://docs.typesafe.ai/api) (replace the placeholder URL and key):
|
| 81 |
|
|
|
|
| 8 |
- classification
|
| 9 |
- qwen3_5
|
| 10 |
- pytorch
|
|
|
|
| 11 |
---
|
| 12 |
|
| 13 |

|
|
|
|
| 74 |
|
| 75 |
## Serve with vLLM Semantic Router
|
| 76 |
|
| 77 |
+
This repository contains model data only. Use the vLLM Semantic Router Decision runtime to load `llm-semantic-router/Decision-1.0-Nox-4B` and serve Choice, Noul, and Score requests. The serving implementation and its dependencies live in vLLM Semantic Router; this release does not bundle executable model code. The weights do not depend on a particular accelerator; which hardware can serve them is decided by the runtime. `transformers.AutoModel.from_pretrained` cannot load the custom Decision head directly.
|
| 78 |
|
| 79 |
After configuring a compatible Decision endpoint, send a [SystemOne request](https://docs.typesafe.ai/api) (replace the placeholder URL and key):
|
| 80 |
|