Feature Extraction
Transformers
Safetensors
decision2
decision-model
classification
system-one
custom_code
Instructions to use vllm-sr/Decision-2.0-Sol-2B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use vllm-sr/Decision-2.0-Sol-2B with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("feature-extraction", model="vllm-sr/Decision-2.0-Sol-2B", trust_remote_code=True)# pip install -U transformers accelerate # Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("vllm-sr/Decision-2.0-Sol-2B", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 4,124 Bytes
b2c5d7e 951e7f7 b2c5d7e 9f6ca45 b2c5d7e 73bb148 b2c5d7e 73bb148 b2c5d7e 951e7f7 b2c5d7e 9f6ca45 73bb148 b2c5d7e 73bb148 b2c5d7e 1b7c47e 6a62b31 73bb148 b2c5d7e 9f6ca45 b2c5d7e 9f6ca45 b2c5d7e 9f6ca45 b2c5d7e 951e7f7 b2c5d7e 73bb148 951e7f7 b2c5d7e 73bb148 1b7c47e 73bb148 b2c5d7e 73bb148 ffe291b 73bb148 ffe291b 73bb148 ffe291b 73bb148 9f6ca45 73bb148 ffe291b b42b6ff 9f6ca45 73bb148 9f6ca45 73bb148 9f6ca45 951e7f7 9f6ca45 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 59 60 61 62 63 64 65 66 67 68 69 70 71 72 73 74 75 76 77 78 79 80 81 82 83 84 85 86 87 88 89 90 91 92 93 94 95 96 97 98 99 100 101 102 103 104 105 106 107 108 109 110 111 112 113 114 115 116 117 118 | ---
license: apache-2.0
base_model: vllm-sr/Decision-1.0-Sol-2B
base_model_relation: finetune
library_name: transformers
tags:
- decision-model
- classification
- system-one
- safetensors
---

# Decision-2.0-Sol-2B
**Decision-2.0-Sol-2B** is the 2B model of [Decision 2.0](https://huggingface.co/collections/vllm-sr/decision-20-6ab7cf7bdfb506bf8269cb00), the decision models of [vLLM Semantic Router](https://github.com/vllm-project/semantic-router). Give it an input (text or JSON) and the questions you need answered: pick one of several options, say yes or no, or rate on a scale. It answers them all at once and returns a probability for every answer, without generating text.
| | |
| --- | --- |
| **Parameters** | 1.88B |
| **Context length** | 16,384 tokens |
| **Decision types** | Choice · Yes / No · Score |
| **License** | Apache-2.0 |
## Highlights
- **Top JevArena score of its size:** 52.1, ahead of the 4 other same-size models compared.
- **Ahead of Decision 1.0 Sol:** +6.3 on JevArena and +4.2 on the Jev Decision Index.
- **Speed:** a median of 7.2 ms per single-question request on a single GPU.
- **Many questions, one pass:** Choice, Yes / No and Score questions about the same input are answered together in one forward pass, with a probability for every option.
## Quickstart
```bash
pip install "transformers>=5.17" torch safetensors
```
```python
import json
from transformers import AutoModel
model = AutoModel.from_pretrained("vllm-sr/Decision-2.0-Sol-2B", trust_remote_code=True)
result = model.system_one(
state="The order arrived damaged yesterday. The customer has a receipt and asks for a replacement today.",
questions={
"route": {
"type": "choice",
"instructions": "Which team should handle this request?",
"criteria": {
"returns": "Refunds, replacements and damaged deliveries",
"billing": "Payments, invoices and charges",
"technical": "Product setup and faults"
}
},
"receipt": {
"type": "noul",
"instructions": "Does the customer have a receipt?"
},
"urgency": {
"type": "score",
"instructions": "How urgent is this request?",
"criteria": [
"Routine",
"Soon",
"Today"
]
}
},
)
print(json.dumps(result["answers"], indent=2))
# Or as a pipeline:
# transformers.pipeline("decision", model="vllm-sr/Decision-2.0-Sol-2B", trust_remote_code=True)(state=..., questions=...)
```
## Evaluation
| Model | JevArena ↑ | Human-labelled transfer ↑ | Jev Decision Index ↑ |
| --- | ---: | ---: | ---: |
| **Decision-2.0-Sol-2B** | **52.1** | **51.3** | **29.5** |
| Decider 2B | 49.5 | 42.0 | — |
| This-That 1.2 | 46.1 | 40.5 | — |
| Decision 1.0 Sol | 45.8 | 49.3 | 25.3 |
| Bosun v3.1 1.7B | 42.1 | 38.0 | — |
### JevArena


<sub>Every model answers the same frozen prompts, scored the same way; missing or invalid answers count as errors. Human-labelled transfer is the median macro-F1 over 15 human-labelled tasks (×100).</sub>
### Jev Decision Index


<sub>Decision 2.0: independent reproduction with the official 0.2.1 kit on the released weights; others: public board snapshot, 2026-09-28. Training data audited at row level against all Index test items.</sub>
## License
Apache-2.0 ([LICENSE](LICENSE)).
## Citation
```bibtex
@misc{decision_2_0_sol_2b_2026,
title = {{Decision-2.0-Sol-2B}: A Decision 2.0 Model for Structured Decisions},
author = {{vLLM Semantic Router Team}},
year = {2026},
howpublished = {\url{https://huggingface.co/vllm-sr/Decision-2.0-Sol-2B}}
}
```
|