Non-autoregressive System 1 decision engine. Typed choice, score and yes/no decisions over any text in a single forward pass, in 100+ languages, with a router that picks the right checkpoint per request.
jev-bench
Independent benchmark of TypeSafe Jev (System One) vs cheap and frontier LLMs: accuracy, calibration, latency, cost
Project facts
- Relationship to Jev
- Research
- Evidence
- Documented
- Language
- Python
- License
- Not reported; check the repository
- Origin
- Original repository
- Repository status
- Not archived
- Created
- 2026-09-23
- GitHub stars
- 0
- Evidence checked
- 2026-09-24T08:40:10.508Z
- Metadata checked
- 2026-09-24T08:40:10.508Z
- Check status
- current
Stars measure the whole repository, including work unrelated to Jev.
Evidence and scope
Documented records the linked documentation or source. JevHunt has not independently run or benchmarked this project.
An independent benchmark of [TypeSafe's Jev] (`jev-1.13`) against a cheap LLM (`openai/gpt-6-luna`) and a frontier LLM (`openai/gpt-6-astra`) on two public classification tasks, graded against ground-truth labels.
Evidence commit: 8f5694d53550d56c8fc4f8a2cd48efc43da6de2f
Discovered through: github-search.
How we review →