divyanshx11 commited on
Commit
b7bd874
·
verified ·
1 Parent(s): b356bcb

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -144,7 +144,7 @@ The default text adapter was fitted to public JevBench decision examples, starti
144
 
145
  JEVision currently has functional checks for typed responses, image routing, and long request acceptance, plus the recorded real-photo example above. The public JevBench examples were used to fit and select the text adapter, so accuracy on those examples is not an independent benchmark result. A separate 33-question grouped text test recorded 21 correct (63.6%) with a 4,096-token cap. This small test does not measure the visual route or long-context answer quality. The 80,000-token value is a service limit, and requests beyond it are rejected instead of silently truncated. The comparisons below describe these specific text panels; they are not a broad accuracy, latency, or cost claim.
146
 
147
- ![Text JevBench comparison of JEVision, Jev, and KEV on the public training panel and grouped test](assets/JevBench Results.png)
148
 
149
  The public-panel result for JEVision is 231/231 (100%), **in-sample** because those labels were used during training and selection. The grouped test is the separate 33-question check. Neither panel evaluates the visual route.
150
 
 
144
 
145
  JEVision currently has functional checks for typed responses, image routing, and long request acceptance, plus the recorded real-photo example above. The public JevBench examples were used to fit and select the text adapter, so accuracy on those examples is not an independent benchmark result. A separate 33-question grouped text test recorded 21 correct (63.6%) with a 4,096-token cap. This small test does not measure the visual route or long-context answer quality. The 80,000-token value is a service limit, and requests beyond it are rejected instead of silently truncated. The comparisons below describe these specific text panels; they are not a broad accuracy, latency, or cost claim.
146
 
147
+ ![Text JevBench comparison of JEVision, Jev, and KEV on the public training panel and grouped test](assets/JevBench-results.png)
148
 
149
  The public-panel result for JEVision is 231/231 (100%), **in-sample** because those labels were used during training and selection. The grouped test is the separate 33-question check. Neither panel evaluates the visual route.
150