BradNLP commited on
Commit
c88e186
·
verified ·
1 Parent(s): d296ce4

update quick start

Browse files
Files changed (1) hide show
  1. README.md +10 -1
README.md CHANGED
@@ -43,7 +43,7 @@ On one H200 with a 1280x720 screenshot, OneJev-4B answers 1 question in 64 ms an
43
  ## Quick start
44
 
45
  ```bash
46
- pip install git+https://github.com/OmniJev/OneJev.git
47
  qev serve --model OmniJev/OneJev-4B
48
  ```
49
 
@@ -62,6 +62,15 @@ r = Client("http://localhost:8000").system_one(
62
  The server speaks TypeSafe's System One API plus a `media` field for images and video. More examples, the latency
63
  benchmark and the code are on [GitHub](https://github.com/OmniJev/OneJev).
64
 
 
 
 
 
 
 
 
 
 
65
  ## All sizes
66
 
67
  | Model | Base | Weights |
 
43
  ## Quick start
44
 
45
  ```bash
46
+ pip install "qev[torch] @ git+https://github.com/OmniJev/OneJev.git"
47
  qev serve --model OmniJev/OneJev-4B
48
  ```
49
 
 
62
  The server speaks TypeSafe's System One API plus a `media` field for images and video. More examples, the latency
63
  benchmark and the code are on [GitHub](https://github.com/OmniJev/OneJev).
64
 
65
+ OneJev-4B also runs on llama.cpp from the [GGUF build](https://huggingface.co/mradermacher/OneJev-4B-GGUF) by mradermacher,
66
+ without PyTorch. Images work there; video needs the PyTorch server.
67
+
68
+ ```bash
69
+ brew install llama.cpp
70
+ pip install git+https://github.com/OmniJev/OneJev.git
71
+ qev serve --gguf mradermacher/OneJev-4B-GGUF:Q8_0
72
+ ```
73
+
74
  ## All sizes
75
 
76
  | Model | Base | Weights |