prithivMLmods commited on
Commit
f736198
·
verified ·
1 Parent(s): d5e8024

Create README.md

Browse files
Files changed (1) hide show
  1. README.md +42 -0
README.md ADDED
@@ -0,0 +1,42 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ tags:
3
+ - text-generation-inference
4
+ - llama-cpp
5
+ - onejev
6
+ - system-one
7
+ - decision-model
8
+ - calibration
9
+ - multimodal
10
+ - gui-agent
11
+ - video
12
+ base_model:
13
+ - OmniJev/OneJev-9B
14
+ license: apache-2.0
15
+ language:
16
+ - en
17
+ pipeline_tag: image-text-to-text
18
+ library_name: transformers
19
+ ---
20
+
21
+ # **OneJev-9B-GGUF**
22
+
23
+ > **OneJev-9B** is the 9-billion-parameter model of OmniJev's OneJev family, a multimodal "System One" decision model built as a full fine-tune of Qwen3.5-9B on 99,193 questions drawn from real agent runs, videos, and images. It accepts a screenshot, photo, video, or plain text along with a few typed questions, such as `Noul` yes/no or `Choice` over labeled options, and returns a calibrated probability for every option in a single forward pass, with no text generation. It is served with the `qev` package, which speaks TypeSafe's System One API and adds a `media` field for images and video. On one H200 with a 1280x720 screenshot, it answers a single question in 81 ms, or 10 questions in one request in 131 ms (13.1 ms per question), which is somewhat slower than the 4B sibling's 64 ms and 104 ms. Its results are reported against Jev 1.13, Jev-Omni 12B, and Qwen3.8-27B in thinking mode on a held-out OneJev test set, but the card shows them only as a graphic, so the individual scores are not readable from the text. The card notes only that Jev 1.13's figures are its published ones and that it reads text only. OneJev-9B (18.8 GB of weights) sits in the middle of a family that runs from OneJev-0.8B to OneJev-27B and an FP8 build of it, and it is released under Apache 2.0.
24
+
25
+ ## Model Files
26
+
27
+ | File Name | Quant Type | File Size | File Link | Description |
28
+ |-----------|------------|-----------|-----------|-------------|
29
+ | OneJev-9B.BF16.gguf | BF16 | 17.9 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.BF16.gguf) | Full BF16 weights. Highest quality, largest file size. |
30
+ | OneJev-9B.Q3_K_L.gguf | Q3_K_L | 4.93 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q3_K_L.gguf) | Lower quality but usable, good for low RAM availability. |
31
+ | OneJev-9B.Q3_K_M.gguf | Q3_K_M | 4.62 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q3_K_M.gguf) | Low quality. |
32
+ | OneJev-9B.Q4_K_M.gguf | Q4_K_M | 5.63 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q4_K_M.gguf) | Good quality, default size for most use cases, *recommended*. |
33
+ | OneJev-9B.Q4_K_S.gguf | Q4_K_S | 5.35 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q4_K_S.gguf) | Slightly lower quality with more space savings, *recommended*. |
34
+ | OneJev-9B.Q5_K_M.gguf | Q5_K_M | 6.47 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q5_K_M.gguf) | High quality, *recommended*. |
35
+ | OneJev-9B.Q5_K_S.gguf | Q5_K_S | 6.31 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q5_K_S.gguf) | High quality, *recommended*. |
36
+ | OneJev-9B.Q6_K.gguf | Q6_K | 7.36 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q6_K.gguf) | Very high quality, near perfect, *recommended*. |
37
+ | OneJev-9B.Q8_0.gguf | Q8_0 | 9.53 GB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.Q8_0.gguf) | Extremely high quality, generally unneeded but max available quant. |
38
+ | OneJev-9B.mmproj-bf16.gguf | mmproj-bf16 | 922 MB | [Link](https://huggingface.co/prithivMLmods/OneJev-9B-GGUF/blob/main/OneJev-9B.mmproj-bf16.gguf) | Multimodal projection file in BF16 format. Used for vision/language models. |
39
+
40
+ ## llama.cpp
41
+
42
+ LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp