Non-autoregressive System 1 decision engine. Typed choice, score and yes/no decisions over any text in a single forward pass, in 100+ languages, with a router that picks the right checkpoint per request.
reflexbench
ReflexBench — open benchmark and evaluation harness for System One models and typed decision engines
Project facts
- Relationship to Jev
- Local alternative
- Evidence
- Documented
- Language
- Python
- License
- Apache-2.0
- Origin
- Original repository
- Repository status
- Not archived
- Created
- 2026-09-23
- GitHub stars
- 1
- Evidence checked
- 2026-09-24T08:40:10.508Z
- Metadata checked
- 2026-09-24T08:40:10.508Z
- Check status
- current
Stars measure the whole repository, including work unrelated to Jev.
Evidence and scope
Documented records the linked documentation or source. JevHunt has not independently run or benchmarked this project.
ReflexBench measures engines such as **TypeSafe Jev, Laya, Reflex/Qwen, Kev, jeff and Verdict** on typed **Choice, Score and Noul/Binary** decisions. It also measures a separate question: how much operational value can a small deterministic **Reflex Core** policy add without changing the model response?
Evidence commit: 0c3535de20124d278502c6d5e261916876ecfadd
Discovered through: github-search.
How we review →