File size: 4,098 Bytes
0355355
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
9dd2482
0355355
675200f
0355355
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
675200f
0355355
675200f
0355355
 
 
 
 
82c9c56
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
0355355
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
66
67
68
69
70
71
72
---
library_name: transformers
license: apache-2.0
base_model: Qwen/Qwen3.5-9B
base_model_relation: finetune
pipeline_tag: text-generation
language:
- en
- zh
tags:
- apus-openjev
- decision-model
- structured-output
- bf16
---
# APUS-OpenJev-v1-9B

[English](README.md) | [涓枃](README.zh-CN.md) 路 [Collection](https://huggingface.co/collections/apus-ailab/apus-openjev-v1-6ab1ee888eb002fcdd3a2825) 路 [Model family](https://huggingface.co/apus-ailab/APUS-OpenJev-v1) 路 [Technical Report](https://huggingface.co/apus-ailab/APUS-OpenJev-v1/blob/main/TECHNICAL_REPORT.md) 路 [Runtime](RUNTIME.md)

A Qwen3.5-based decision model for browser action selection, workflow routing, and natural-language principle judgments. This repository contains **9B checkpoint-3000 merged BF16 weights**, ready to download independently without a separate LoRA adapter.

## Highlights

- Score dynamic candidates supplied with each request and return their distribution.
- The included native runtime supports `effort=low` (16 layers) and `effort=high` (32 layers). Use high for text generation.
- Reuse Qwen language representations and vocabulary projection; application code can assemble decisions into structured workflow outputs.

## Quick start

```bash
python -m pip install huggingface_hub
hf download apus-ailab/APUS-OpenJev-v1-9B --local-dir ./APUS-OpenJev-v1-9B
cd APUS-OpenJev-v1-9B
python -m pip install -r requirements.txt
python examples.py . --device cuda:0 --effort high
```

## Evaluation and training

The merged model in this repository scores **68/80 (85.00%)** at full depth on the [Frozen80 development panel](https://huggingface.co/datasets/apus-ailab/APUS-OpenJev-Eval-Frozen80), covering Browser, HelpSteer3, BoolQ, MNLI, and attribute decisions. See [merged-evaluation.json](merged-evaluation.json) and [training.md](training.md). This reused engineering panel is not an independent blind benchmark or an end-to-end browser success rate.

Candidate probabilities express relative preference; calibration is required before interpreting them as correctness probabilities. BF16 merging changes some probabilities; see [runtime evidence and limitations](RUNTIME.md). The series 9B release selects checkpoint-3000, corresponding to the 85% development-panel result. Checkpoint-5949 remains available in the original family repository.

## Series and downloads

The 4B, 9B, and 35B-A3B models use independent repositories grouped in a Collection. Hugging Face displays downloads per model. The original family repository and legacy paths remain available. The 35B-A3B merged release is subject to its own validation and publication status.

<!-- OPENJEV-QUANT:START -->
## GGUF and MLX versions

Quantized versions of this model (Frozen80 68/80) for Ollama / llama.cpp / LM Studio and Apple Silicon Macs: [GGUF collection](https://huggingface.co/collections/apus-ailab/apus-openjev-v1-gguf-6ab39d5e724c4d8a1021198f) 路 [MLX collection](https://huggingface.co/collections/apus-ailab/apus-openjev-v1-mlx-6ab39d5fc988a1b1cb89dfc8).

| Version | Frozen80 | Decisions = BF16 |
|---|---:|---:|
| [GGUF Q8_0](https://huggingface.co/apus-ailab/APUS-OpenJev-v1-9B-GGUF) | 68/80 | 80/80 |
| [GGUF Q4_K_M](https://huggingface.co/apus-ailab/APUS-OpenJev-v1-9B-GGUF) | 68/80 | 78/80 |
| [MLX 8bit](https://huggingface.co/apus-ailab/APUS-OpenJev-v1-9B-MLX-8bit) | 69/80 | 79/80 |
| [MLX 4bit](https://huggingface.co/apus-ailab/APUS-OpenJev-v1-9B-MLX-4bit) | 68/80 | 75/80 |

```bash
ollama run hf.co/apus-ailab/APUS-OpenJev-v1-9B-GGUF:Q8_0 --think=false
```

Ollama needs thinking disabled (`--think=false`, or `"think": false` in the API). Per-question results and usage are in each repository.
<!-- OPENJEV-QUANT:END -->

## License and acknowledgments

We thank the [Qwen/Qwen3.5-9B](https://huggingface.co/Qwen/Qwen3.5-9B) team. See [LICENSE](LICENSE), [provenance](provenance/), and the pinned source and file identities in [release-manifest.json](release-manifest.json).

**Authors:** gumpcheng ([xDAN2099](https://huggingface.co/xDAN2099)), zhangxu, [APUS AI-LAB](https://github.com/APUS-AI-Lab).