rspamd-jev

rioriost/rspamd-jev
Research

TypeSafe Jev shadow-evaluation plugin for Rspamd with optional GPT provider comparison

PythonDocumented

system1-fraud-interceptor-demo

ordepas/system1-fraud-interceptor-demo
Research

Demo de un interceptor de fraude simulado que compara en paralelo un modelo Sistema 1 (Jev, TypeSafe AI) con un LLM (Gemini) sobre transacciones sintéticas: velocidad, costo por llamada y decisiones. Proyecto personal de experimentación, no es un benchmark.

TypeScriptDocumented

an-email-classifier

jnorgren/an-email-classifier
Research

A small CLI that uses the TypeSafe API (Jev model) to evaluate emails

PythonDocumented

battleship-vs-jev

ickas/battleship-vs-jev
Research

A 60-game benchmark of TypeSafe's Jev evaluation model playing Battleship. The model matches plain code; it does not beat it.

TypeScriptDocumented

jev-access-day

andreaserradev-gbj/jev-access-day
Research

A learning scaffold for TypeSafe AI's System One models: eval harness plus a measured, plain-language comparison of the Jev decision model vs an LLM stand-in on 24 real operational decisions. All numbers reproducible from committed run files.

TypeScriptDocumented

jev-algorithms

HikaruEgashira/jev-algorithms
Research

Runtime-agnostic algorithms built on TypeSafe's Jev structured-evaluation model

JavaScriptDocumented

jev-city

Bravim-Ketan-Purohit/jev-city
Research

A traffic city where every car is driven by TypeSafe's Jev model, benchmarked against a rule-based driver.

TypeScriptDocumented

jev-classification-benchmark

rachit-srivastava-devx/jev-classification-benchmark
Research

Benchmarking TypeSafe Jev against 15 chat-model configurations on support-ticket classification: latency, tokens, cost, accuracy.

HTMLDocumented

jev-drive

eylexlive/jev-drive
Research

A 3D driving simulator where Jev, TypeSafe's decision model, chooses what the car does. Code eye or Gemini camera eye, code reflexes, an evaluation harness.

PythonDocumented

jev-enterprise-decision-fabric

ghubnab99/jev-enterprise-decision-fabric
Research

Architecture for running many semantic decisions through one validated path, with a labelled 111-case benchmark comparing TypeSafe Jev against a Claude baseline, and a dashboard for inspecting any single decision. Experimental, not production.

C#Documented

jev-eval

dshvimer/jev-eval
Research

Classification benchmark: Jev vs Claude Sonnet 5 vs Claude Opus 5 on AG News

PythonDocumented

jev-experiment

immanuelsavio/jev-experiment
Research

Benchmarking TypeSafe Jev against general-purpose LLMs on support-ticket routing, with a focus on latency, accuracy, and confidence.

JavaScriptDocumented

jev-go

nandansrikrishna/jev-go
Research

Standalone Go CLI and MCP server for TypeSafe Jev: typed judgments, JSONL evaluation, and resumable batches.

GoDocumented

jev-ja-eval

uesgugikouhei-oss/jev-ja-eval
Research

Evaluate Jev (TypeSafe AI) on Japanese customer-inquiry data: 150-item dataset + comparison script (Jev / LLM / rule-based)

PythonDocumented

jev-lab

dairui1/jev-lab
Research

Experiments with TypeSafe's Jev: triage benchmark vs LLM, and a source study of jev-ultrafast vs Cline's jev-browser

HTMLDocumented

jev-lab

moguone/jev-lab
Research

Small apps for evaluating TypeSafe AI's System One model (Jev). Unofficial.

JavaScriptDocumented

jev-latam-lead-triage

integralmarketingmx/jev-latam-lead-triage
Research

Triage de leads de WhatsApp/CRM con Jev (TypeSafe AI): ruteo por confianza, plantilla n8n y benchmark en español. Sin dependencias. No afiliado.

PythonDocumented

jev-measured

WallerChen/jev-measured
Research

Measured cost, latency and raw output from the live Jev API (TypeSafe AI System One model) across 8 use cases — reproducible

PythonDocumented

jev-orderby-bench

yodablocks/jev-orderby-bench
Local alternative

Does ORDER BY over a Jev probability put rows in a defensible order? Independent ranking, calibration and invariant measurements of TypeSafe AI's Jev: passes six pre-registered gates on 360 labeled rows, fails four of six on graded product relevance.

PythonDocumented

jev-spam-lab

loicrg/jev-spam-lab
Research

CLI for evaluating Microsoft Outlook email with TypeSafe AI’s Jev model.

TypeScriptDocumented

jev-test

souvikr/jev-test
Research

Test harness + benchmark for TypeSafe's Jev decision model (noul/choice/score) via OpenRouter's Decisions API

PythonDocumented

jev-tetris-benchmark

planstack-ai/jev-tetris-benchmark
Research

Reproducible Tetris decision benchmark comparing TypeSafe Jev with Claude Haiku

TypeScriptDocumented

jeval

vrash/jeval
Research

jeval: open-source evaluations for AI outputs and agents, judged by Jev

TypeScriptDocumented