Profile
Back to NewsBack
GitHub Trending 27 min
Reader Mode
valentynkit/awesome-jev-typesafe: Typed decisions with TypeSafe's Jev, the first System One model

valentynkit/awesome-jev-typesafe: Typed decisions with TypeSafe's Jev, the first System One model

10 hours ago

Awesome Jev: typed decisions, System One

Awesome Jev Awesome</a>

Lint</a> Links</a> !Entries Last commit</a> License</a>

Search the site · Categories · Contribute · Agent skill · llms.txt · 中文 · 日本語 · 한국어

[!TIP]
Agents can install this list as a skill: npx skills add valentynkit/awesome-jev-typesafe. It teaches them to fetch projects.json once and filter locally.

Typed decisions from TypeSafe's Jev, the first System One model: state in, calibrated probabilities out, no text to parse.

Jev does not write. You hand it some state and a list of typed questions, and it answers each one with a probability, a pick from options you defined, or a position on a scale you defined. One request, about 100 ms, $0.042 per million input tokens, output free, every question scored in parallel. This list is where people are putting that to work, sorted by what you would install.

Awesome Jev in seventeen seconds: the radar sweeping repos, a search reranked by Jev, the radar page, and the trending board

Seventeen seconds of the site; the full-size MP4 is in the repo.

jev-ultrafast booking a flight in 7.1 seconds
jev-ultrafast
SemIf running a System One replica in the browser, no waitlist
SemIf
kev playground: six typed questions answered in one forward pass on a laptop
kev
NanoJev solving a maze beside Jev and an untuned Qwen
NanoJev
jev-review dashboard scoring a pull request file by file
jev-review
hermes-jev-skills logging every routing decision with its confidence and latency
hermes-jev-skills

Contents

- Claude Code - Codex - Pi - Hermes - Agent Zero - Any agent - Skills for writing Jev code - Launch coverage - Independent measurements - Essays and threads

Jev on one screen

Copied from the vendor's pages on 2026-09-18; every page is linked under Start here.

| | Chat model | Jev | | ------------- | ------------------------------- | ---------------------------------------------------------------- | | Output | Text you parse | A probability, a pick, or a scale position, per question | | Latency | Seconds | 70 to 500 ms, vendor reported | | Price | Dollars per million tokens | $0.042 per million input tokens, output free | | Hallucination | Any string | Only values you defined; still confidently wrong at times | | Fits | Planning, writing, open answers | Routing, gating, ranking, judging, anything with a finite answer |

  • Endpoint: POST https://api.typesafe.ai/v1/systemone with a model, a state, and a map of questions.
  • Model alias: jev-latest, currently jev-1.13.0.
  • Questions: Choice picks one of up to 255 options you define, Score places the state on a scale you describe, Noul is a calibrated yes/no probability. All questions in a request are scored in parallel against the same state.
  • Input: text or JSON state. No images, audio, or video yet.
  • Price: $0.042 per million input tokens; output tokens are free.
  • Limits: 250,000 tokens per second, 1,200 requests per minute, 32k tokens per request, all subject to change during early access.
  • Latency: 70 to 500 ms end to end, vendor reported.
  • Training: RLCD, reinforcement learning for calibrated decisions. Weights and architecture are unpublished.
  • Also served by Vercel AI Gateway, Cloudflare Workers AI, OpenRouter (beta), and Netlify AI Gateway, none of which need the TypeSafe waitlist.

Know before you build

  • Type safe is not the same as correct. A schema-valid answer can still be confidently wrong. The "cannot hallucinate" claim means no out-of-schema output, nothing more.
  • On the vendor's own four-workflow eval Jev lands around 68 percent, close to mid-tier LLMs. Keep irreversible actions behind a threshold and a human.
  • It cannot count, do arithmetic, reason about dates, or produce a value that is not in your option list. Ask it to pick from a deck, never to name a card.
  • Accuracy drops as the state fills with unrelated content. Curating what you send is your job, and it is most of the work.
  • The vendor publishes its known failure modes on the model jaggedness page linked below. Read it before you pick your first threshold.
  • Access is a waitlist. Open replicas and third-party gateways exist below if you cannot wait, or would rather not depend on one vendor.

Start here

  • Introduction - The mental model in two pages: state plus typed questions in, typed answers with probabilities out.
  • Quick start - First request in Python, TypeScript, or curl.
  • Primitives - Choice, Score, and Noul, and when each one fits.
  • State - How to package what Jev judges, and why less is more.
  • API reference - The request and response contract.
  • Models - Aliases, current version, price, and rate limits.
  • Model jaggedness: jev-1.13 - Known failure modes, straight from the vendor.
  • System One - What the category means and how it differs from a chat model.
  • How to build with System One - Decompose a judgment into atomic questions and keep the control flow in code.
  • Confidence - What the confidence field means and how to turn it into act, review, or fall back.
  • Patterns - Speculative fan-out, confidence routing, composite scoring, intent routing.
  • Use-case map - The vendor's own catalogue of where Jev fits and where it does not.
  • Cookbooks - Worked recipes, starting with batching many questions into one request; the sidebar has the rest.
  • Workflow evals - The vendor's benchmark on four workflows, with the caveats printed on the page.
  • Manifesto - The product thesis, summed up as build prod, not god.
  • llms.txt - Every documentation page as plain Markdown, for feeding to an agent.
  • Console - Waitlist, API keys, and usage.
  • Jev on Vercel AI Gateway - Model id typesafe-ai/jev, billed through Vercel, no TypeSafe waitlist.
  • Jev on Cloudflare Workers AI - Call typesafe/jev from a Worker through env.AI.run.
  • Jev on OpenRouter - Beta listing on the general-purpose gateway, model id typesafe/jev-1.13, billed on your OpenRouter key.
  • Jev-verified cascade - OpenRouter cookbook: a cheap model answers, Jev checks the answer, only the failures escalate.
  • Jev on Netlify AI Gateway - Zero-config access from Netlify Functions, no separate TypeSafe key.
  • LiteLLM pass-through - Route the System One endpoint through a LiteLLM proxy for key management and cost tracking; no streaming, since TypeSafe has none.
  • Discord - Official server; builder demos live in the show-and-tell channel.

Official SDKs and framework support

  • typesafe-sdk-js - TypeScript and JavaScript client with answer types inferred from your questions.
  • typesafe-sdk-python - Python client, sync and async.
  • system-one-adapter-python - Same TypeSafeClient interface backed by an LLM API, so you can compare Jev against a chat model on identical questions.
  • skills - Agent skills for designing questions, building workflows, and evaluating them.
  • Agent skill - How to install the official skill in Claude Code, Cursor, and friends.
  • Vercel AI SDK provider - @ai-sdk/typesafe-ai exposes Jev through experimental_evaluate.
  • eve - Vercel's agent framework; Jev is the typed judge in its evaluate step.
  • ai-cli - The Vercel AI SDK in your terminal, with an evaluate path that runs on Jev.

Coding agents

pi-warden: rules in a Markdown file, judged by Jev on every write

Claude Code

  • fast-jev-compaction - Replaces the compaction summary with Jev decisions: every tool call and result scored in one request, stale ones dropped, everything kept stays verbatim.
  • jev-router - Routes each task to the cheapest Claude model that can handle it.
  • winnow - Judges every tool result before it enters context, so the window fills slower instead of being cleaned later.
  • yoshi - Context-pruning proxy for Claude Code and Codex, with the savings measured rather than claimed.
  • skillranker - Rust CLI and hooks that rank installed skills for the next step using live session context, with abstention.
  • jcm-router - Local proxy that picks model and effort per message and leaves the cached main chat alone.
  • jev-skillful - Per-prompt router over skills, MCP servers, agents, and commands, and it measures whether the injection helped.
  • limpet - A Stop hook that keeps the agent from stopping too early, judged against plain-language rules.
  • jevwire - MCP server, embeddable decision model, and an escalate-only plugin that can make the harness stricter but never looser.
  • jev-code - Command-line toolkit that coding agents hand judgment-heavy work to, one typed Jev workflow per request.
  • vexjoy-agent - Agent toolkit whose /d command picks the specialist agent, skill, and pipeline with one Jev call, plus an optional Jev auto-compact plugin.
  • save-token-jev - Compaction that asks Jev which tool calls still matter and keeps the rest verbatim, with adapters for Claude Code, Codex, OpenCode, and raw API transcripts.
  • jev-pruner - Trims long Bash output with Jev after the command runs and before the model sees it; short output, errors, and structured formats pass untouched.
  • jev-rules - Scores your standing rules against each prompt and delivers only the ones that apply, once per session.
  • jev-belay - Stop hook that blocks an unverified "done": reads the transcript for evidence and, only when files changed with no passing check since, spends one four-question Jev call; fails open on every error path.
  • jev-use - Hands the Claude Code, Codex and pi steps that need no text output to Jev, with a typed escalation contract for everything it should not decide.

Codex

  • jev-codex-router - Picks model, thinking depth, and speed mode for every Codex turn.
  • codex-jev-router - Uses Jev to select the model and reasoning effort for Codex subagents, with confidence gates and a Sol fallback.
  • Astra-Ares - Has Jev pick the reasoning effort and how long to hold it for a running Codex task, on a patched Codex CLI built from upstream source.

Pi

  • pi-jev by y0usaf - A measured tool-call gate plus a jev_ask tool for typed answers inside Pi.
  • pi-warden - Guardrails that steer instead of interrupt: irreversible calls, off-task calls, stuck loops, unverified done claims, about 250 ms each.
  • pi-jev-auto-mode - Auto-approves bash, write, and edit calls semantically and fails closed when it cannot decide.
  • pi-jev by TheoOliveira - Semantic tool routing and typed decisions as Pi tools.
  • pi-jev-router - Automatic model routing for Pi through the Vercel AI Gateway.
  • pi-fast-jev-compaction - The verbatim compaction idea, ported to Pi.
  • bicameral - Hybrid harness for Pi: an LLM writes the code, Jev reflexes gate every call as allow, confirm, block, warn, or steer.
  • pi-quiet-ask - Jev as the Pi coding agent's quiet decision layer.
  • pi-typesafe-jev - Pi extension exposing Jev judgments as five Pi tools.
  • pi-typesafe - Batched evaluation tool, terminal playground, and a typed API for Pi extension authors.
  • pi-heed - Checks every side-effecting tool call against what you said earlier in the session, so "review only" still holds after compaction.
  • pi-jev-sentinel - Checks Pi tool calls, tool outputs, and replies for risky actions and prompt injection, with user approvals, context re-checks, secret scrubbing, and optional task pinning.
  • pi-mcp-adapter - Opt-in typed evaluation and semantic search over MCP tool results, behind a per-server data-egress allowlist.

Hermes

  • typesafe-skill-router - Names the one skill worth loading before the model call; stdlib only, about a tenth of a cent per turn.
  • jev-agent-skill-router - Confidence-aware skill routing with an abstain path.
  • hermes-jev - Typed decisions, ranking, verification, and an opt-in tool gate.
  • ask-jev-skill - Lets Hermes and similar agents ask Jev directly.
  • hermes-jev-plugin - Four Hermes tools for atomic checks, routing, and rubric scoring; listed in the Hermes plugin catalog.
  • hermes-jev-approvals - Approves, denies, or escalates flagged shell commands before they run; vendor-reported speedups.

Agent Zero

Any agent

  • skillbox - Self-hosted, versioned skills library served over MCP, with Jev recommending which skill to load.
  • jev-mcp by jkudish - The first MCP server for Jev, and still the most linked.
  • typesafe-mcp - Go MCP connector.
  • jev-mcp by blakestone-x - Classify, score, check, match, and screen, with confidence on every answer.
  • Jevbridge - ACP and MCP adapter that pairs Jev with any LLM for computer use and typed decisions.
  • jev-eval-mcp - Eval-first MCP server: prototype a question, map it over many items, then measure variants against labeled examples with a threshold sweep.
  • azdaja - Recursive language model layer for Claude Code, Codex, Gemini, and OpenCode that keeps full sources in a local evaluator; Jev is an optional leaf for reranking, verification, classification, and semantic joins, with budgeted, checkpointed batches.

Skills for writing Jev code

  • building-with-jev-skill - Skill for writing and improving programs that call Jev.
  • jev-system-architect - Finds the fuzzy judgment in a system and turns it into small Choice, Score, and Noul primitives.
  • jev-judgment - Sends a coding agent's closed judgments to Jev instead of the chat model.
  • skills by fabricioctelles - Agent-skill directory that can score subjective evaluation criteria with Jev.
  • Augustus - Skill for deciding where a typed judgment belongs at all and what stays in code; a companion to the official skill, not a replacement.
  • hermes-jev-skills - Hands an agent's small decisions to Jev: which model answers the turn, which skills to load, which passages matter, which turns survive compaction.
  • Stanley - Coding CLI where Jev routes a plain-language request to one deterministic workflow, and that workflow asks Jev fixed-choice questions about the evidence it gathered.
  • JevHarness - Has an LLM write a task-specific harness that turns observations into Jev questions, then freezes it and improves it from rewards and full execution traces.

Browser and computer use

mobile-jev driving the Uber app on a real Android phone

  • jev-ultrafast - One request picks both the operation and the target element from an indexed DOM table; a small LLM only writes typed text. Zürich to London booked in 7.1 seconds.
  • typesafe-computer-use - OCR the screen, classify the next action, click; about $0.0002 a step on macOS.
  • mobile-jev - The same loop on a real Android phone; nine Uber actions in 21 seconds in the demo.
  • jev-browser by jkudish - The first community browser agent on Jev, with a demo GIF.
  • jev-voice-browser - Intent and target decided per spoken word in about 300 ms, often before the sentence ends.
  • jev-browser by Ying-Kai-Liao - An LLM plans, Jev decides; library, CLI, and MCP server.
  • jev-browser by tontoko - One grounded Jev and Playwright core behind a typed SDK, a persistent CLI, and an MCP server.
  • jev-mobile - Android sub-agent over USB running observe, normalize, decide, mutate, verify, with Jev deciding.
  • jev-ego - Browser agent that spends one Jev request per step to pick the action.
  • jev-browser-use - Codex skill and plugin where Jev handles navigation, clicks, and scrolling and Codex keeps typing and verification; reports browser steps 5 to 10 times faster.
  • Jev-cu - Codex computer use where Jev picks the element, action, completion, and risk from on-screen text, no screenshots sent; Chinese readme.
  • JevScout - Job-hunting skill that drives Chrome over CDP and has Jev score every link and listing.
  • Jev Social - Lets Jev choose each read-only social research step while socai runs it in a real Chrome session and streams inspectable Instagram, TikTok, or LinkedIn evidence into a report.
  • jev-use by savka777 - Voice and typed computer use for macOS: Jev picks the next on-screen action from the Accessibility tree, with no screenshots.
  • jev-mobile by xinwang-nwpu - Android automation where one Jev request picks both the action and the target element from the accessibility tree, executed over ADB.
  • jev-browser-use by imanshu03 - Runs browser tasks from a plain-language instruction over CDP or Vercel's agent-browser, with Jev selecting operations and targets and code checking confidence before it acts.
  • jev-gui-delegate - Runs a delegated GUI task for Codex through the real Chrome session or Windows UI Automation, with Jev picking the control at each step.

Open models and replicas

None of these ship TypeSafe's weights. They reproduce the interface, the parallel scoring trick, or both, on open models.

  • SemIf - Semantic ifs from open models on a single 3090; the most starred independent replica, formerly openjev.
  • jevlike - Open option scorer that reads candidate logits instead of generating JSON.
  • NanoJev - 0.6B replica with parallel decisions, dynamic candidates, and an end-to-end training pipeline.
  • openjev-sglang - Jev-compatible API endpoint on SGLang, prefill only.
  • jev-visual - Educational visual-inference variant on Apple Silicon: shared context, direct candidate scoring.
  • reflex - Small open decision model on Qwen3.5: state plus typed questions to calibrated probabilities.
  • decider - One-pass typed decisions fine-tuned from Qwen3.5-2B.
  • jevmlx - Parallel constrained decisions for any MLX model on Apple Silicon, one forward pass.
  • mini-jev - Preregistered experiment on a frozen Qwen3-4B: read the option letter's logits, skip the JSON.
  • system-one-open - Typed calibrated decisions in one forward pass on Gemma 4 E2B and Gemma 3 270M.
  • Verdict-open-jev - Non-autoregressive decision engine on ModernBERT with calibrated uncertainty and an in-browser WebGPU playground.
  • jevfire - Parallel decisions for CUDA LLMs through a vLLM API, with game-agent examples and benchmarks.
  • jevbetter - A stronger one-pass scorer with a head-to-head benchmark against the jevlike starter design.
  • open-alternative-jev - Typed, calibrated decisions from any open-weights model in one forward pass, on Hugging Face and vLLM.
  • openjev by zhihz - Bilingual local decisions from context, questions, and candidate answers.
  • jev-on-a-laptop - Study of Jev-style decisions on stock 1.5B to 8B models on a laptop, with a Hugging Face demo.
  • typesafe-ai-benchmark - LLM gateway that mimics the TypeSafe response shape, useful as a stand-in while you wait for a key.
  • Parallel constrained decoding - Hugging Face Space demonstrating RLCD-style parallel decoding on Qwen2.5-1B.
  • openvons - Open decision layer answering a finite option set with probabilities split into execute, confirm, and reject. Independent replica, not TypeSafe weights.
  • von - Non-autoregressive open decision model reporting under 15 ms locally, as a drop-in alternative to Jev.
  • litjev - Turns any off-the-shelf LLM into a Jev-style decision layer.
  • open-jev - Typed JSON inference with DiffusionGemma, benchmarked against Jev.
  • kev - LoRA adapter and readout head on Qwen2.5-0.5B that answers many typed questions in one prefill; trains in under two hours on a MacBook, held-out ECE 0.065, speaks the TypeSafe wire format.
  • simple-jev - Reads next-token logits from any Hugging Face model for choice, rubric, and support questions; public demo API with no key.
  • OpenJev by razorback16 - Jev-compatible decision server on DiffusionGemma 26B through vLLM, images included, hosted free on Codiv.
  • openjev by daseinlabs - Prefills once and scores every option in one padded pass on Gemma 3 4B with MLX; plays Doom from the terminal in the demo.
  • jeff by logan-markewich - Self-hosted System One API on the 400M GLiFormer, with benchmarks that say where it trails Jev.
  • JevForge - End-to-end stack for auditable data construction, Qwen3.5-0.8B training, fixed Mind2Web and OOD evaluation, local serving, and a preliminary RLCD baseline.
  • PlayJev - Plays ten browser games from the frame alone on a fine-tuned Qwen3.5-0.8B, one forward pass per move, open weights and a browser demo.
  • openJev-verdict-2.0 - Non-autoregressive 151M decision engine with a WebGPU browser runtime, reporting 77.1 percent top-1 and 1.44 percent calibration error on its own benchmark.
  • LLM2Jev - Adapts local language models to runtime-defined Choice, Score, and Noul questions and returns typed answers with probabilities.
  • local-jev - Offline server wire-compatible with the System One endpoint, checked against the official SDK; reports 80.5 percent on JevBench's public items against 86.6 percent published for the hosted API.
  • Jev-Compatible - Gateway that turns an existing SGLang or vLLM deployment into a Jev-compatible decision service by scoring candidate tokens, with no training and no model changes.
  • Rizzo Flow - Local-first take on the System One idea: unstructured state in, typed probabilistic decisions out, without generating a token.
  • metask-jev - Open-weight typed-decision models on Qwen3.5 with calibrated per-option probabilities in one forward pass; reports 80.1 percent on JevBench against 75.3 for Jev 1.13.
  • AnyJev - Turns an open Hugging Face model into a calibrated decision endpoint served through vLLM, with no fine-tuning at the first two levels.
  • ollaya - Pulls and serves open decision models such as Laya, kev, and JevK5 behind a TypeSafe-compatible local endpoint, the way Ollama serves LLMs.
  • AgentJev - Decision model on Qwen3-0.6B that answers without decoding tokens, reports 79.25 percent top-1 on the public Typed Decisions benchmark.
  • Valen - Multimodal decision model on Qwen3.5 that scores candidates against images and video as well as text, with training code and open weights.
  • tev1 - Data recipe and training code that fine-tune Qwen3.5-4B into an open-weight decision model on Together AI.
  • JevK5 - Qwen3.5 replica with distilled LoRA weights, ranked second of 76 systems and first among open ones on JevBench v1.4.
  • djev - Serves /v1/systemone on DiffusionGemma by reading typed answers and text spans off a seeded, pinned canvas, with no generation.

Code review and quality

  • software-factory - Runs several coding agents locally with Jev as judge and orchestrator.
  • jev-review by devagrawal09 - Staged code-review workflow with a local dashboard.
  • jev-review by NiazMorshed2007 - Local-first MCP plugin for continuous quality review by coding agents.
  • foreman - Supervises a software factory of agents, with Jev making the go and no-go calls.
  • supercov - Code quality and coverage signals for coding agents.
  • diffjury - PR risk router and review coach.
  • clean-code-review - Every file in a PR judged against Clean Code rules, then reviewed by an LLM.
  • JevLint - Configurable semantic linting with file-level Noul judgments.
  • commit-miner - Classifies commit diffs and messages: bug fixes, security fixes with CWEs, change types.
  • jev-review-action - GitHub Action for submission review and PR classification with Jev, no text-generation model in the loop.
  • jev-triage - Pulls large repositories and triages their issues with typed Jev questions.
  • perch - Semantic linting: rules in plain language, each file judged by Jev, run locally or in CI.
  • jeff by Alurith - Read-only Go CLI that checks files against coded rules such as hidden side effects and weak error handling.
  • jev-pref - Turns the preferences in your AGENTS.md into a linter that runs on code changes and reports back to the agent.
  • jev-commit - Pre-commit hook: one Jev call judges whether the commit message matches the staged diff, plus debug leftovers, unmentioned work, and a credential belt; warns except on a secret, which it blocks.
  • slop-grader - Grades text and markdown files for AI slop, grammar, and technical documentation quality, and guides an AI agent to auto-fix violations.
  • JevPR - GitHub App that asks Jev whether a pull request is safe to approve or needs a specialist, then maps the verdict to a check run.
  • jevopt - C/C++ compiler driver that asks Jev whether to inline each discretionary call site, from the LLVM IR and the original source.
  • jev-spec - Checks the code against the requirements in a Markdown spec on every commit and fails the build when the two drift apart.
  • pytest-jev - Pytest plugin that asks Jev whether plain-English claims about a test's text hold, all in one request, and fails the test with each claim's probability unless Jev is at least 80 percent sure.
  • jgrep by kyu1204 - --diff gates a PR in CI on a rule written in English, --tests lists the test files a diff can affect, and plain jgrep greps code by what it does; one Noul per chunk, 16 chunks per request.

Routing and gateways

Janus measuring on banking77 when routing to Jev beats a single model

  • tiershift - Shifts every LLM call to the cheapest model that can handle it, policy in YAML, decision in about 180 ms.
  • jev-router by prismhq - LLM router on top of LiteLLM.
  • agent-router - Picks Cursor, Claude Code, Codex, or OpenCode plus model and effort for a task, then launches it.
  • Janus - Measures on your data when Jev beats other models, then routes accordingly.
  • hono-jev-router - Route HTTP requests by meaning in Hono.
  • JevRouter - Models, subagents, skills, MCP tools, and CLIs as one candidate set; Jev picks, the router enforces permissions and risk; reports 44 percent first-five tool-call hits against 24 for DeepSeek on Toolathlon.
  • jev-gateway - Local gateway for Codex and Claude Code that sends the "which tool next" decision to Jev and everything else to your usual model.
  • jev-router by daviddl9 - Jev picks the worker tier for each step in OMP and Pi, keeping planning and review on a strong model and bounded work on cheaper ones.
  • Jevonian - Local OpenAI, Anthropic, and Responses-compatible proxy that serves one Jev call per turn to answer both the model route and the thinking level for its virtual model jevonian/auto, with code filtering candidates by protocol, context window, effort floor, and spent quota windows first, and pinned models or explicit routes skipping Jev entirely.
  • neurolink - TypeScript AI SDK where decide, via Jev, is a peer of generate and stream: one typed-judgment call routes model choice, prunes context, and picks MCP tools.

Search, reranking and RAG

  • jev-search - Source selection, query understanding, and relevance ranking for web search.
  • blink - Codebase search where Jev scores the candidates.
  • reranker - Jev as a calibrated reranker: one call, up to 30 documents, a probability per document.
  • llama-index-jev - LlamaIndex reranker and router, cheaper than an LLM judge.
  • jev-tree - Recursive choice over a taxonomy, past the 255-option cap.
  • jev-folio-recursive-classifier - Classifies OCR'd legal agreements through the FOLIO Document Types ontology with recursive Jev Choices, beam search, confidence-gated leaf stopping, and context-length benchmarking.
  • neo4jev - Walks a Neo4j graph by classifying neighbouring relationships.
  • jev-sift - MCP tool that scores a batch of files, URLs, or snippets for relevance so the agent opens only what matters.
  • jev-scout - Rust CLI and MCP server that finds real, maintained repos and crates for a plain-language request, with Jev scoring the candidates.
  • jev-semgrep - Greps by meaning instead of by regex: every line gets a probability from Jev, and meanings combine with AND, OR, and NOT.
  • jev-search by kylemclaren - A shadcn/ui registry block: keyword hits on the first keystroke, re-ranked by Jev a moment later, and keyword order stands if the call fails.
  • Senseek - Browser extension that searches the page you are reading by meaning, in a Ctrl+F style box, with your own key and no backend.
  • Milvus Search with Jev - Nine runnable notebooks combining Gemini embeddings and Milvus retrieval with Jev decisions for reranking, filtering, search stopping, routing, cache reuse, curation, guardrails, and evaluation.
  • JevPDF - Searches a PDF by meaning in the browser: pdf.js extracts the lines locally, Jev answers one Noul per line in batches of 16, and matching lines light up page by page, ranked by probability.

Data and ops

jev-reviewer: every answer is a quote with its file and place

  • pg-jev - PostgreSQL extension that answers plain-language questions about your tables.
  • vgi-typesafe - DuckDB worker that exposes choice, noul, and score as lateral-joinable table functions in SQL.
  • jevsql - SQL with natural-language predicates over SQLite: filter, rank, and classify rows by meaning, batched and cost-guarded.
  • sqlite-jev - Adds Jev Noul, Choice, and Score judgments to SQLite through a loadable C extension and Python wrapper, with scalar functions and batched virtual-table queries.
  • jevlogs - Scores OpenTelemetry log signal before paying for LLM analysis.
  • jev-curate - Sifts Parquet and JSONL training data at more than 1,500 rows a second.
  • HA-Jev - Home Assistant integration: ask a question about your house, get a probability, choice, or score as an entity.
  • typesafe-migration-guard - Reviews database migrations for safety before they run.
  • jev-for-engineers - Eight small examples from mechanical and electrical engineering: CAD routing, FEM triage, DFM screening, BOM alignment.
  • jlink - Links records across two datasets from a match rule written in plain English, from Python, the shell, Stata, or R, and reports F1 0.73 against 0.69 for tuned string matching on NBER patent assignees to Compustat.
  • pg_typesafe - Pre-alpha PostgreSQL extension for categorical classification with Jev.
  • jev-mode - Ticket triage and file tagging on a typed-judgment model; reports 78 percent fewer tokens and 96.1 percent accuracy against a 93.7 percent baseline.
  • jev-reviewer - Asks a clinical trial report for systematic-review data by voice, text, or a questions file; every answer is a verbatim quote with its file and place.
  • tax-doc-classifier - One request per page picks among 261 IRS forms and seven page kinds; reports 100 percent on its corpus at $0.001 a page, 34 times cheaper than the LLM pipeline it replaced.
  • jevql - Semantic SQL for PostgreSQL, with Jev answering the predicates.
  • duckdb-jev - DuckDB extension that asks a question of every row and returns a real SQL type.
  • invalidate - Gives every stored agent memory a lease and asks Jev whether new evidence ends it; live demo.
  • jevgraph - Parses PDF, DOCX, PPTX, or text locally, then asks Jev one closed-set relation question per candidate entity pair and exports a graph with per-edge probabilities and page evidence to JSON, CSV, or Neo4j.
  • jev-ultralightspeed - Packs 32 items into one request for bulk classification and calibrates the confidence cut that sends the least-sure rows to a person, reporting 533 items a second at 89.2 percent agreement with human labels.
  • jev-seo by AgriciDaniel - Crawls a site, checks it against 52 SEO rules, has Jev judge every page, and writes PDF, spreadsheet, and Markdown reports.

Safety, moderation and verification

jev-guard: every tool call risk-scored as allow, ask, or deny before it runs

  • jev-shield - Semantic MCP firewall that screens every tool call, result, and description; reports 94 percent block recall at about $0.00002 a check.
  • jev-guard - Auto mode for Claude Code, Codex, Cursor, Gemini CLI, Pi, and OpenCode: risk-scores each tool call as deny, ask, or allow and flags prompt injection in results.
  • Jev-Moderation-Bot - Chat moderation with editable rules.
  • citation-verifier - Does the cited paper support the sentence citing it? Claude finds the quote, Jev scores it, a human decides.
  • human-compiler - Paste text, get diagnostics, like a compiler for prose.
  • snifftest - Prose linter for AI writing tells: countable rules plus one judgment model.
  • riff - Ruff-style rule codes for writing.
  • jev-secret-detection - Measures how well Jev spots real credentials in file snippets, with the hard config-shaped cases scored separately.
  • jev-audio-beeper - Low-latency audio censorship proof of concept: Jev typed decisions drive ffmpeg.
  • is-malicious - Scans a codebase for hidden or data-stealing behavior before you run it; a clean report is not proof, and it says so.
  • tripwire - AI SDK middleware and proxy that runs seven Jev checks on every LLM response before the user sees it; no accuracy numbers yet, and it says so.
  • SkillCheck - Sends an agent skill's text to Jev before installation and returns a verdict with category scores, without loading the skill into the agent's context; the submitted text is not redacted for secrets.
  • StopSpam - Telegram bot that removes spam and scam messages from group chats on calibrated-confidence classification.
  • Fake / Real - Checks claims against cited excerpts with Jev in an English and Romanian fact-checking site, with a public integration example and a closed-source full application.

Applications and extensions

jevibe-check labelling the tone of every Bluesky post in the feed

  • unclutter - Browser extension that removes page clutter with reusable template rules.
  • typesafe-adblock - Chrome extension that asks "is this element an ad?" per DOM node; a toy, and it says so.
  • vibecheck - Vibe-check your X post before you hit publish.
  • xtags - Labels every post in your X timeline with what it wants you to do.
  • jevibe-check - Live tone labels for Bluesky posts and drafts.
  • jevmeter - Puts a live meter on any video: every sentence scored on five questions, rendered as a 16:9 edit, a whole debate for about two cents.
  • killmyidea - Describe your startup idea; Jev says kill it, fix it, or ship it.
  • notra - Turns work into content, with Jev deciding what is worth posting.
  • slidepilot - Voice-driven auto-advance for Slidev on Cloudflare Agents.
  • should-ai-kill-us-all - Asks Jev the question every ten minutes, using the actual headlines.
  • Privacy Facts - Turns privacy policies into nutrition-style labels with plain-language answers, Jev confidence scores, and suggested source clauses.
  • jev-voice-control - Menu-bar Swift app turning spoken commands into Jev typed decisions and macOS actions.
  • jev-got - Game of Thrones roleplay where a story model writes each scene and Jev answers five typed questions that drive the header, soundtrack, art, and next prompt.
  • typesafe-jev - Jev experiments starting with a local CV-screening workbench, each with its own measured results.
  • super-jev - Small harness connecting evidence, Jev judgments, permitted actions, and verified outcomes.
  • safer-with-jev - Neon Function proxy for the Neon AI Gateway with Jev routing in front.
  • Sponsor Skip - Chrome extension that finds sponsor reads from the transcript or live audio and jumps past them; code owns every timestamp, under a cent an hour in transcript mode.
  • jev-seo by AkashPriyadarshii - Rust CLI and MCP server for SEO and GEO checks over DuckDuckGo results, scored by Jev.
  • [j
... (README truncated for length)
Chat with me