Feature Extraction
Transformers
PyTorch
Safetensors
English
hastejev
jev
decision-engine
system-1
agent-routing
tool-routing
non-generative
pica
quantized
Instructions to use noffy/hastejev-2m with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use noffy/hastejev-2m with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("feature-extraction", model="noffy/hastejev-2m")# Load model directly from transformers import HasteJevEngine model = HasteJevEngine.from_pretrained("noffy/hastejev-2m", device_map="auto") - Notebooks
- Google Colab
- Kaggle
fix: re-export 2m weights; true int8/int4 keys; honest model card
Browse files- README.md +51 -59
- model.safetensors +2 -2
- model_fp16.safetensors +1 -1
- model_int4.safetensors +2 -2
- model_int8.safetensors +2 -2
- model_int8_weight.safetensors +3 -0
- pytorch_model.bin +2 -2
README.md
CHANGED
|
@@ -5,89 +5,81 @@ library_name: transformers
|
|
| 5 |
license: apache-2.0
|
| 6 |
pipeline_tag: feature-extraction
|
| 7 |
tags:
|
| 8 |
-
- jev
|
| 9 |
- hastejev
|
| 10 |
-
-
|
| 11 |
- decision-engine
|
| 12 |
- system-1
|
| 13 |
-
-
|
| 14 |
-
-
|
| 15 |
-
- low-latency
|
| 16 |
- non-generative
|
| 17 |
-
-
|
| 18 |
-
- browser-control
|
| 19 |
-
- web-automation
|
| 20 |
-
- agentic-ai
|
| 21 |
-
- fast-inference
|
| 22 |
-
- decision-making
|
| 23 |
-
- calibration
|
| 24 |
- safetensors
|
| 25 |
- pytorch
|
| 26 |
- quantized
|
| 27 |
-
- int8
|
| 28 |
-
- int4
|
| 29 |
-
- fp16
|
| 30 |
---
|
| 31 |
|
| 32 |
-
#
|
| 33 |
|
| 34 |
-
|
| 35 |
-
[](https://github.com/racstan/hastejev)
|
| 36 |
-
[](https://opensource.org/licenses/Apache-2.0)
|
| 37 |
-
[]()
|
| 38 |
-
[]()
|
| 39 |
|
| 40 |
-
|
|
|
|
|
|
|
| 41 |
|
| 42 |
-
|
| 43 |
|
| 44 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 45 |
|
| 46 |
-
|
| 47 |
-
- **Trainable Parameters**: 1,416,675
|
| 48 |
-
- **Buffer / Projection Table**: 409,600
|
| 49 |
-
- **Hidden Dimension ($d_{\text{model}}$)**: 160
|
| 50 |
-
- **Transformer Layers**: 4
|
| 51 |
-
- **Attention Heads**: 4
|
| 52 |
-
- **Target Deployment**: Browser automation & bots
|
| 53 |
-
- **Quantization Formats Available**: `FP32`, `FP16` (`model_fp16.safetensors`), `INT8` (`model_int8.safetensors`), `INT4` (`model_int4.safetensors`)
|
| 54 |
|
| 55 |
-
|
| 56 |
|
| 57 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 58 |
|
| 59 |
-
|
| 60 |
-
|
| 61 |
|
| 62 |
-
#
|
| 63 |
-
engine = HasteJevEngine.from_pretrained("noffy/hastejev-2m")
|
| 64 |
|
| 65 |
-
|
| 66 |
-
|
| 67 |
|
| 68 |
-
|
| 69 |
-
|
| 70 |
-
options = ["Approve Transaction", "Flag for Review", "Decline"]
|
| 71 |
|
| 72 |
-
|
| 73 |
-
|
|
|
|
|
|
|
|
|
|
| 74 |
```
|
| 75 |
|
| 76 |
-
|
| 77 |
|
| 78 |
-
##
|
| 79 |
|
| 80 |
-
|
|
| 81 |
-
|
|
| 82 |
-
|
|
| 83 |
-
|
|
| 84 |
-
|
|
| 85 |
-
|
|
| 86 |
-
| [`hastejev-5m`](https://huggingface.co/noffy/hastejev-5m) | **~5.0M** | 224 | 5 | 4 | ~20.0 MB | ~5.0 MB | Financial & KYC routing |
|
| 87 |
-
| [`hastejev-10m`](https://huggingface.co/noffy/hastejev-10m) | **~10.0M** | 320 | 5 | 4 | ~40.0 MB | ~10.0 MB | Multimodal agent kernels |
|
| 88 |
-
| [`hastejev-20m`](https://huggingface.co/noffy/hastejev) | **~20.4M** | 256 | 4 | 4 | ~81.5 MB | ~20.4 MB | Enterprise decision engine |
|
| 89 |
|
| 90 |
-
-
|
|
|
|
| 91 |
|
| 92 |
-
##
|
| 93 |
-
Apache
|
|
|
|
| 5 |
license: apache-2.0
|
| 6 |
pipeline_tag: feature-extraction
|
| 7 |
tags:
|
|
|
|
| 8 |
- hastejev
|
| 9 |
+
- jev
|
| 10 |
- decision-engine
|
| 11 |
- system-1
|
| 12 |
+
- agent-routing
|
| 13 |
+
- tool-routing
|
|
|
|
| 14 |
- non-generative
|
| 15 |
+
- pica
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 16 |
- safetensors
|
| 17 |
- pytorch
|
| 18 |
- quantized
|
|
|
|
|
|
|
|
|
|
| 19 |
---
|
| 20 |
|
| 21 |
+
# Haste Jev 2m (Small)
|
| 22 |
|
| 23 |
+
**Role:** Browser / UI agent kernel
|
|
|
|
|
|
|
|
|
|
|
|
|
| 24 |
|
| 25 |
+
Open-weights **System-1 decision engine** for software paths that need typed
|
| 26 |
+
decisions under a time budget (agent routing, tool routing, intent classification,
|
| 27 |
+
pre-flight guardrails, browser action selection). Not a chat model.
|
| 28 |
|
| 29 |
+
## Measured specification
|
| 30 |
|
| 31 |
+
| Field | Value |
|
| 32 |
+
|---|---|
|
| 33 |
+
| Total parameters | 1,826,275 |
|
| 34 |
+
| Trainable parameters | 1,416,675 |
|
| 35 |
+
| Hash-table buffers | 409,600 |
|
| 36 |
+
| d_model | 160 |
|
| 37 |
+
| Layers | 4 |
|
| 38 |
+
| Heads | 4 |
|
| 39 |
|
| 40 |
+
Parameter counts match the GitHub README table (verified with `verify_claims.py`).
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 41 |
|
| 42 |
+
## Honest claims
|
| 43 |
|
| 44 |
+
| Claim | Status |
|
| 45 |
+
|---|---|
|
| 46 |
+
| PICA option-order bias = 0.0% | **Verified** architecturally |
|
| 47 |
+
| Exact parameter table | **Verified** |
|
| 48 |
+
| FP32 ~0.4 MB for 100k weights | **Verified** (weight storage only) |
|
| 49 |
+
| p99 < 15ms / ECE < 0.009 / 99.4% arithmetic | **Not verified** — do not cite from this card |
|
| 50 |
+
| Published latency / accuracy on your workload | **Measure yourself** |
|
| 51 |
|
| 52 |
+
Weights may be lightly or untrained prototypes depending on export; treat behavioral
|
| 53 |
+
accuracy as unknown until you evaluate on labeled data.
|
| 54 |
|
| 55 |
+
## Quickstart
|
|
|
|
| 56 |
|
| 57 |
+
```python
|
| 58 |
+
from hastejev import HasteJevEngine
|
| 59 |
|
| 60 |
+
eng = HasteJevEngine.from_pretrained("noffy/hastejev-2m")
|
| 61 |
+
# eng = HasteJevEngine.from_pretrained("noffy/hastejev-2m", quantization="int4")
|
|
|
|
| 62 |
|
| 63 |
+
r = eng.choice(
|
| 64 |
+
"Request: reset password for user@corp.example",
|
| 65 |
+
["auth_self_service", "billing", "security_review"],
|
| 66 |
+
)
|
| 67 |
+
print(r.decision, r.confidence)
|
| 68 |
```
|
| 69 |
|
| 70 |
+
Install: `pip install git+https://github.com/racstan/hastejev.git`
|
| 71 |
|
| 72 |
+
## Files
|
| 73 |
|
| 74 |
+
| File | Contents |
|
| 75 |
+
|---|---|
|
| 76 |
+
| `model.safetensors` / `pytorch_model.bin` | FP32 state dict |
|
| 77 |
+
| `model_fp16.safetensors` | FP16 |
|
| 78 |
+
| `model_int8.safetensors` | True weight-only int8 (`weight_q`) when re-exported with ≥1.1.0 |
|
| 79 |
+
| `model_int4.safetensors` | True packed int4 (`weight_packed`) when re-exported with ≥1.1.0 |
|
|
|
|
|
|
|
|
|
|
| 80 |
|
| 81 |
+
Older revisions of `model_int8`/`model_int4` may be mislabeled FP32; re-export or
|
| 82 |
+
re-download after this commit.
|
| 83 |
|
| 84 |
+
## License
|
| 85 |
+
Apache-2.0
|
model.safetensors
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:b60c6b4d086b85096a2d37e8fb18e1f2b3a2a42819719311beb2fbf3d500c334
|
| 3 |
+
size 3756960
|
model_fp16.safetensors
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 3660374
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:55c7803483188f7d313cc42f2088bef8ae6a0a4096e8920af9b30f24abd637ef
|
| 3 |
size 3660374
|
model_int4.safetensors
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:b60c6b4d086b85096a2d37e8fb18e1f2b3a2a42819719311beb2fbf3d500c334
|
| 3 |
+
size 3756960
|
model_int8.safetensors
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:70482eec773e434f0e710f8dfb7c81b9962d4adb7256e9b94c819c9f27db7fc0
|
| 3 |
+
size 4267800
|
model_int8_weight.safetensors
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:70482eec773e434f0e710f8dfb7c81b9962d4adb7256e9b94c819c9f27db7fc0
|
| 3 |
+
size 4267800
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:5c07fd5b383918f33c410d79111aeafb6ffb12f6302bd28f9cff5959538c6433
|
| 3 |
+
size 3777895
|