APUS-OpenJev-v1-4B-GGUF

APUS-OpenJev-v1-4B is a Qwen3.5-4B-based decision model from APUS AI-LAB, designed for browser action selection, workflow routing, and natural-language principle judgments rather than open-ended text generation — this repository provides the 4B checkpoint-5949 merged BF16 weights, downloadable independently without a separate LoRA adapter. It scores dynamic candidates supplied per request and returns their preference distribution, reusing Qwen's language representations and vocabulary projection so application code can assemble decisions into structured workflow outputs; its native runtime supports two effort levels — low (16 layers) and high (32 layers, recommended for text generation). On the internal Frozen80 development panel (covering Browser, HelpSteer3, BoolQ, MNLI, and attribute-decision tasks), the merged model scores 66/80 (82.50%), though the authors note this is a reused engineering panel rather than an independent blind benchmark or end-to-end browser success rate, and that candidate probabilities express relative preference rather than calibrated correctness (with BF16 merging further shifting some probabilities). It's part of a three-size family (4B, 9B, 35B-A3B) hosted in separate repositories under a shared collection, with the 9B release instead selecting checkpoint-3000 (85% on the same panel), and is released under Apache 2.0 as a finetune of the Qwen3.5-4B base.

Model Files

File Name Quant Type File Size File Link Description
APUS-OpenJev-v1-4B.BF16.gguf BF16 8.42 GB Link Full BF16 weights. Highest quality, largest file size.
APUS-OpenJev-v1-4B.Q3_K_L.gguf Q3_K_L 2.42 GB Link Lower quality but usable, good for low RAM availability.
APUS-OpenJev-v1-4B.Q3_K_M.gguf Q3_K_M 2.26 GB Link Low quality.
APUS-OpenJev-v1-4B.Q4_K_M.gguf Q4_K_M 2.71 GB Link Good quality, default size for most use cases, recommended.
APUS-OpenJev-v1-4B.Q4_K_S.gguf Q4_K_S 2.56 GB Link Slightly lower quality with more space savings, recommended.
APUS-OpenJev-v1-4B.Q5_K_M.gguf Q5_K_M 3.07 GB Link High quality, recommended.
APUS-OpenJev-v1-4B.Q5_K_S.gguf Q5_K_S 2.99 GB Link High quality, recommended.
APUS-OpenJev-v1-4B.Q6_K.gguf Q6_K 3.46 GB Link Very high quality, near perfect, recommended.

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Downloads last month
824
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

6-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for prithivMLmods/APUS-OpenJev-v1-4B-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(5)
this model

Collections including prithivMLmods/APUS-OpenJev-v1-4B-GGUF