--- license: apache-2.0 base_model: - apus-ailab/APUS-OpenJev-v1-4B tags: - text-generation-inference - llama-cpp - apus-openjev - decision-model - structured-output - bf16 language: - en - zh pipeline_tag: text-generation library_name: transformers --- # **APUS-OpenJev-v1-4B-GGUF** > **APUS-OpenJev-v1-4B** is a Qwen3.5-4B-based decision model from APUS AI-LAB, designed for browser action selection, workflow routing, and natural-language principle judgments rather than open-ended text generation — this repository provides the 4B checkpoint-5949 merged BF16 weights, downloadable independently without a separate LoRA adapter. It scores dynamic candidates supplied per request and returns their preference distribution, reusing Qwen's language representations and vocabulary projection so application code can assemble decisions into structured workflow outputs; its native runtime supports two effort levels — `low` (16 layers) and `high` (32 layers, recommended for text generation). On the internal Frozen80 development panel (covering Browser, HelpSteer3, BoolQ, MNLI, and attribute-decision tasks), the merged model scores 66/80 (82.50%), though the authors note this is a reused engineering panel rather than an independent blind benchmark or end-to-end browser success rate, and that candidate probabilities express relative preference rather than calibrated correctness (with BF16 merging further shifting some probabilities). It's part of a three-size family (4B, 9B, 35B-A3B) hosted in separate repositories under a shared collection, with the 9B release instead selecting checkpoint-3000 (85% on the same panel), and is released under Apache 2.0 as a finetune of the Qwen3.5-4B base. ## Model Files | File Name | Quant Type | File Size | File Link | Description | |-----------|------------|-----------|-----------|-------------| | APUS-OpenJev-v1-4B.BF16.gguf | BF16 | 8.42 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.BF16.gguf) | Full BF16 weights. Highest quality, largest file size. | | APUS-OpenJev-v1-4B.Q3_K_L.gguf | Q3_K_L | 2.42 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q3_K_L.gguf) | Lower quality but usable, good for low RAM availability. | | APUS-OpenJev-v1-4B.Q3_K_M.gguf | Q3_K_M | 2.26 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q3_K_M.gguf) | Low quality. | | APUS-OpenJev-v1-4B.Q4_K_M.gguf | Q4_K_M | 2.71 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q4_K_M.gguf) | Good quality, default size for most use cases, *recommended*. | | APUS-OpenJev-v1-4B.Q4_K_S.gguf | Q4_K_S | 2.56 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q4_K_S.gguf) | Slightly lower quality with more space savings, *recommended*. | | APUS-OpenJev-v1-4B.Q5_K_M.gguf | Q5_K_M | 3.07 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q5_K_M.gguf) | High quality, *recommended*. | | APUS-OpenJev-v1-4B.Q5_K_S.gguf | Q5_K_S | 2.99 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q5_K_S.gguf) | High quality, *recommended*. | | APUS-OpenJev-v1-4B.Q6_K.gguf | Q6_K | 3.46 GB | [Link](https://huggingface.co/prithivMLmods/APUS-OpenJev-v1-4B-GGUF/blob/main/APUS-OpenJev-v1-4B.Q6_K.gguf) | Very high quality, near perfect, *recommended*. | ## llama.cpp LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp