watfile

jexp/watfile
Research

Text/PDF - File categorization and sorting with Typesafe AI Jev or local calibrated decision model

PythonDocumented

dsh-jev-decide

nanami-0713/dsh-jev-decide
Research

DSH plugin: register TypeSafe Jev (System One decision model) as an agent tool — jev_decide returns calibrated probabilities (noul/choice/score) for routing/triage/guardrail judgments, no text generation. 把 TypeSafe Jev 决策模型注册为 DSH agent 工具

JavaScriptDocumented

jev-chess

hemanth/jev-chess
Research

Chess moves, evaluations, persona opponents, and game classification with TypeSafe AI System One

TypeScriptCode reference

llm-prompt-techniques-on-jev

leepokai/llm-prompt-techniques-on-jev
Research

Chain-of-thought and self-refinement for TypeSafe's Jev: feed its typed answers back as state and ask again. Benchmarks vs TypeSafe's own cookbook numbers.

PythonDocumented

siftr

Bentlybro/siftr
Research

Fast, cheap judgment for AI coding agents: semantic search, focused reads and list picking in ~2s. CLI + MCP server on TypeSafe Jev. Benchmarked on SWE-bench.

PythonDocumented

typesafe-chess

Dimesio/typesafe-chess
Research

FUn little experiment with Typesafe AI Jev Model playing chess against stockfish :)

JavaScriptDocumented

jev-baselines-eval

ickma2311/jev-baselines-eval
Research

Pre-registered independent eval of TypeSafe Jev against a nano-class LLM, a frontier LLM, and a supervised encoder (Banking77 + CLINC150 zero-shot)

PythonDocumented

jev-evals

NicolasMontone/jev-evals
Research

Rubric-based eval harness cheap enough to run on every PR, powered by typesafe-ai/jev

TypeScriptDocumented

jev-mode

ddfeyes/jev-mode
Research

I kept watching coding agents burn context on decisions that aren't hard - triage 400 tickets, tag 600 files, route to one of six teams. jev-mode moves those verdicts to a typed-judgment model. I A/B'd it: 78% fewer tokens, 16x less work-attributable input, accuracy 96.1% vs 93.7%. Python, no deps, MIT.

PythonDocumented

LightJev

rongxinzy/LightJev
Research

Train lightweight language backbones for typed decisions and candidate probabilities. CE/Brier training, evaluation, and an offline end-to-end demo.

PythonDocumented

openevals

memovai/openevals
Research

Fast and cheap agent evals. jev as judge.

TypeScriptDocumented

polymarket-btc-5m-agent

BrunooMoniz/polymarket-btc-5m-agent
Research

Agente de trading para o mercado BTC Up/Down de 5 minutos da Polymarket: modelo em código, Jev (TypeSafe System One) como portão, ordens maker, calibração e shadows em paper

PythonDocumented

research_desk

0xnairb/research_desk
Research

TypeSafe Jev demonstration for new analyzation — experimenting with Jev for fast analysis of news and tickers

PythonDocumented

ruby_llm-providers-typesafe

javiergradiche/ruby_llm-providers-typesafe
Research

TypeSafe System One models (Jev) for RubyLLM: typed judgments, evaluations and reranking.

RubyDocumented

toolgate

RiskAverseTech/toolgate
Research

Open auto mode for AI agents — a calibrated tool-call firewall powered by TypeSafe Jev. Ships as a Claude Code hook

TypeScriptDocumented

zerosweep

sysadarsh/zerosweep
Research

Autonomous System-One Triage Engine & Benchmark powered by TypeSafe AI (Jev). 75ms inference, $0 output tokens, and RLCD epistemic safety gates.

TypeScriptDocumented

Janus

FirasSX914/Janus
Research

Measure when to use Jev and other models on your data, then route accordingly.

PythonDocumented

jev-agent-failure-benchmark

TokenTrim/jev-agent-failure-benchmark
Research

Benchmarking Jev (Typesafe.ai) against a strong LLM on the Who&When Pro agent-failure-attribution benchmark (text subset).

PythonDocumented

jev-mcp

BYK/jev-mcp
Research

An eval-first MCP server for TypeSafe's Jev, a System One model that returns typed judgments (noul, choice, score) with probabilities instead of generated text.

TypeScriptDocumented

jev-research-eval

jgridifier/jev-research-eval
Research

Reproducible Jev Ultrafast research-browser eval harness + field note (QC’d cases, suite runner, report generator). Not investment advice.

HTMLDocumented

jev-routing-experiment

TokenTrim/jev-routing-experiment
Research

Benchmarking TypeSafe's Jev decision model as a cost-efficient LLM router on RouterArena

PythonDocumented

jev-spam-eval

bitnovus/jev-spam-eval
Research

Zero-shot spam filtering with TypeSafe Jev Noul questions, compared with TF-IDF baselines

Jupyter NotebookDocumented

jev-starter

hamakyo/jev-starter
Research

Typed, policy-driven decision workflows on top of TypeSafe AI Jev: confidence routing, fallbacks, evaluation, and RAG patterns for TypeScript apps.

TypeScriptDocumented

jev-freeform

kesku/jev-freeform
Research

An observable raw-character chat experiment powered entirely by TypeSafe Jev Choice

JavaScriptDocumented