modular-rag-mcp

bloodfel/modular-rag-mcp
Research

Pluggable, observable RAG MCP server — hybrid retrieval (BM25+dense+RRF), BEIR-benchmarked rerankers (BGE / TypeSafe Jev / LLM), Streamlit dashboard + Langfuse, multimodal ingestion

PythonDocumented

reflexgate

intelliDean/reflexgate
Research

Ultra-fast, sub-100ms API & webhook guardrail and triage gateway powered by TypeSafe AI System One (Jev). Parallel 7-dimension speculative evaluation, deterministic policy router, zero-dependency SQLite audit trail, and automated outbound dispatch.

TypeScriptDocumented

S1Rank

zaesho/S1Rank
Research

S1Rank: Can a System-One decision model (TypeSafe Jev) rerank? Benchmark, raw responses, and paper.

PythonDocumented

strands-system-one

florianbuetow/strands-system-one
Research

Strands agents that make small local LLMs (Qwen3 0.6B, MiniCPM5 2B, Qwen3.5 4B) answer like Jev, a System 1 model: typed yes/no, choice and score answers with probabilities. Benchmarked against Jev on the jevals suite, via logprob readout and written-out JSON probabilities.

PythonDocumented

verdict

Manavarya09/verdict
Local alternative

Small, fast, honest decision models. Open alternative to Jev: zero-shot, fit on your labels in seconds, calibrated with a coverage guarantee, Jev wire-compatible.

PythonDocumented

agent-evals

marianoberton/agent-evals
Research

Deterministic evaluation harness for LLM agents, with a calibrated judge (Jev) that can act as a CI gate.

TypeScriptDocumented

All-About-JEV

vardhantech123/All-About-JEV
Research

Jev is actually not a traditional LLM, it doesn’t generate text. It’s what the TypeSafe AI team calls a System One model: 📖 System One models are a class of AI models built to make fast, structured decisions that software can use directly. A System One model evaluates a state and returns typed answers and probabilities.

Documented

better-jev-for-all

asp616848/better-jev-for-all
Local alternative

Open, self-hostable, faster System One decision model — API-compatible alternative to TypeSafe's Jev

PythonDocumented

fab-evidence-gate

hongdroid94/fab-evidence-gate
Research

Evidence-aware semiconductor alert triage research demo with TypeSafe Jev, policy guards, and reproducible evaluation

TypeScriptDocumented

foreman-jev-evaluation

MahdiHedhli/foreman-jev-evaluation
Research

JEV evaluation (FM-JEV-01): replay-only, advisory-only evaluation of Foreman-style supervision with TypeSafe Jev. Measures how much safety comes from the model versus a deterministic evidence gate. No worker authority.

PythonDocumented

GitHub-Issue-Classification-Using-Jev

KalyanM45/GitHub-Issue-Classification-Using-Jev
Research

This repository contains a GitHub issue classifier built on Jev, TypeSafe AI's System One model. It labels every new issue with typed values and calibrated confidence in milliseconds, labelling what it is sure about and escalating what it is not. Three guardrail layers guard every write, and a frozen eval suite gates each deploy.

PythonDocumented

jev-ai-project

rathan-bk/jev-ai-project
Research

Alert triage with Jev (TypeSafe System One): proof of concept and benchmark against a rules-only baseline

PythonDocumented

jev-cli

shetautnetjer/jev-cli
Research

Clean-room CLI for TypeSafe System One / Jev Decision Contracts, search, and benchmarks.

PythonDocumented

jev-csat-korean-2026

KKodiac/jev-csat-korean-2026
Research

typesafe.ai Jev(jev-1.13.0) benchmarked against the 2026 수능 국어영역 — JSON passages/questions + scored results

Documented

jev-decisionops

gbesse/jev-decisionops
Research

Production evaluation and gateway tooling for Jev and System One-compatible decision models

TypeScriptDocumented

jev-demo

nadeem4/jev-demo
Research

Decision Arena: TypeSafe's Jev vs open-source Laya playing highway-env, Snake and Blackjack with zero training, plus benchmarks and a Claude Code watchdog

TypeScriptDocumented

jev-deterministic-benchmark

etsabary/jev-deterministic-benchmark
Research

1,000-decision behavioral benchmark of Jev across 25 deterministic reasoning families.

PythonDocumented

Jev-DocEval

TechGenDM/Jev-DocEval
Research

A hybrid CLI and SPA dashboard leveraging the TypeSafe System One (Jev) API to perform quantitative, rubric-based evaluations of markdown documentation.

PythonDocumented

jev-eval

Clementtang/jev-eval
Research

Stance test of TypeSafe Jev vs Claude on Taiwan sovereignty questions, in Traditional Chinese, Simplified Chinese and English

JavaScriptDocumented

jev-eval

gooooloo/jev-eval
Research

在自己的数据上评测 TypeSafe Jev 的准确率、概率校准和可用阈值

PythonDocumented

jev-gaokao-eval

kuaitoukuai/jev-gaokao-eval
Research

Evaluate Jev (TypeSafe System One decision model) on GAOKAO-Bench objective questions - accuracy, confidence calibration and latency over 1497 Chinese college-entrance-exam multiple-choice questions

PythonDocumented

jev-langgraph-router

all3n2601/jev-langgraph-router
Research

Typed, confidence-aware Jev routing for LangGraph.js with reproducible benchmarks

TypeScriptDocumented