jbrashear commited on
Commit
dc35cf5
·
verified ·
1 Parent(s): aea96d9

Update the Decision Index line and add curated-list listings

Browse files
Files changed (1) hide show
  1. README.md +2 -1
README.md CHANGED
@@ -76,7 +76,8 @@ Made in Texas.
76
  **Results and docs**
77
 
78
  - Project site, with every result and how to run the models: [jebadiah.ai](https://jebadiah.ai).
79
- - For the family: [Decision Index 0.2.1](https://huggingface.co/spaces/multimodalart/jev-decision-index): the 27B scores 54.67, #5 of 67 open models (as of 2026-09-26). This is the board's own number; the maintainer validated my run and put it on the leaderboard. [Run record](https://huggingface.co/datasets/frontier-infra/jebadiah-decision-index-results/tree/main/runs/jebadiah-27b-1c0d794f).
 
80
  - JevBench v1.4.2: on its 231 public items, run through its own harness, Jebadiah 27B scores 0.866, the same as Jev 1.13.0; Jebadiah 9B v2 scores 0.818. This is my own run on the public items, not the official board, which also uses sealed items. [Details and caveats](https://github.com/getainode/jebadiah#where-we-stand-on-jevbench). The 4B has not been run on either.
81
  - All sizes: the [Hugging Face collection](https://huggingface.co/collections/frontier-infra/jebadiah-open-system-one-decision-models-6ab80765ddd3fa0b3eba5213), mirrored on [ModelScope](https://www.modelscope.ai/profile/JasonBrashear).
82
 
 
76
  **Results and docs**
77
 
78
  - Project site, with every result and how to run the models: [jebadiah.ai](https://jebadiah.ai).
79
+ - For the family: [Decision Index 0.2.1](https://huggingface.co/spaces/multimodalart/jev-decision-index): the 27B scores 54.67, No. 5 of 70 open models (board data generated 2026-09-28). This is the board's own number; the maintainer validated my run and put it on the leaderboard. [Run record](https://huggingface.co/datasets/frontier-infra/jebadiah-decision-index-results/tree/main/runs/jebadiah-27b-1c0d794f).
80
+ - Listed in the [LLM Engineer Toolkit](https://github.com/KalyanKS-NLP/llm-engineer-toolkit) and [awesome-jev-tools](https://github.com/v-modal/awesome-jev-tools).
81
  - JevBench v1.4.2: on its 231 public items, run through its own harness, Jebadiah 27B scores 0.866, the same as Jev 1.13.0; Jebadiah 9B v2 scores 0.818. This is my own run on the public items, not the official board, which also uses sealed items. [Details and caveats](https://github.com/getainode/jebadiah#where-we-stand-on-jevbench). The 4B has not been run on either.
82
  - All sizes: the [Hugging Face collection](https://huggingface.co/collections/frontier-infra/jebadiah-open-system-one-decision-models-6ab80765ddd3fa0b3eba5213), mirrored on [ModelScope](https://www.modelscope.ai/profile/JasonBrashear).
83