Feature Extraction
Transformers
PyTorch
Safetensors
English
hastejev
jev
decision-engine
system-1
agent-routing
tool-routing
non-generative
pica
quantized
Instructions to use noffy/hastejev-10m with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use noffy/hastejev-10m with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("feature-extraction", model="noffy/hastejev-10m")# Load model directly from transformers import HasteJevEngine model = HasteJevEngine.from_pretrained("noffy/hastejev-10m", device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 2,281 Bytes
ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 ad3c1ae 0512730 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 59 60 61 62 63 64 65 66 67 68 69 70 71 72 73 74 75 76 77 78 79 80 81 82 83 84 85 86 | ---
language:
- en
library_name: transformers
license: apache-2.0
pipeline_tag: feature-extraction
tags:
- hastejev
- jev
- decision-engine
- system-1
- agent-routing
- tool-routing
- non-generative
- pica
- safetensors
- pytorch
- quantized
---
# Haste Jev 10m (Large)
**Role:** Richer multimodal state
Open-weights **System-1 decision engine** for software paths that need typed
decisions under a time budget (agent routing, tool routing, intent classification,
pre-flight guardrails, browser action selection). Not a chat model.
## Measured specification
| Field | Value |
|---|---|
| Total parameters | 10,002,275 |
| Trainable parameters | 6,856,675 |
| Hash-table buffers | 3,145,600 |
| d_model | 320 |
| Layers | 5 |
| Heads | 4 |
Parameter counts match the GitHub README table (verified with `verify_claims.py`).
## Honest claims
| Claim | Status |
|---|---|
| PICA option-order bias = 0.0% | **Verified** architecturally |
| Exact parameter table | **Verified** |
| FP32 ~0.4 MB for 100k weights | **Verified** (weight storage only) |
| p99 < 15ms / ECE < 0.009 / 99.4% arithmetic | **Not verified** — do not cite from this card |
| Published latency / accuracy on your workload | **Measure yourself** |
Weights may be lightly or untrained prototypes depending on export; treat behavioral
accuracy as unknown until you evaluate on labeled data.
## Quickstart
```python
from hastejev import HasteJevEngine
eng = HasteJevEngine.from_pretrained("noffy/hastejev-10m")
# eng = HasteJevEngine.from_pretrained("noffy/hastejev-10m", quantization="int4")
r = eng.choice(
"Request: reset password for user@corp.example",
["auth_self_service", "billing", "security_review"],
)
print(r.decision, r.confidence)
```
Install: `pip install git+https://github.com/racstan/hastejev.git`
## Files
| File | Contents |
|---|---|
| `model.safetensors` / `pytorch_model.bin` | FP32 state dict |
| `model_fp16.safetensors` | FP16 |
| `model_int8.safetensors` | True weight-only int8 (`weight_q`) when re-exported with ≥1.1.0 |
| `model_int4.safetensors` | True packed int4 (`weight_packed`) when re-exported with ≥1.1.0 |
Older revisions of `model_int8`/`model_int4` may be mislabeled FP32; re-export or
re-download after this commit.
## License
Apache-2.0
|