Awesome Jev

Review queue

These repositories fell between the gates: Jev was confident enough not to drop them, but not confident enough to list them. The policy leaves them here rather than guessing. A human reading one repository for thirty seconds usually settles it.

If you can tell which way a row should go — or if you spot a wrong call anywhere in the index — open an issue with the Jev got it wrong template. Corrections become overrides in the data, and overrides feed the calibration set.

79 of 579 judged repositories are awaiting review.

  • Use Jev (TypeSafe's System One model) as a calibrated reranker: one call, up to 30 documents, a probability per document. Apache-2.0.

    held because: category confidence 0.34 below 0.5

    SDKs and ClientsRankingPython
    genuine0.98Full judgment
  • No description provided.

    held because: category confidence 0.46 below 0.5

    ApplicationsControl LoopTypeScript
    genuine0.96Full judgment
  • No description provided.

    held because: category confidence 0.45 below 0.5

    Agent and Dev ToolingControl LoopJavaScript
    genuine0.96Full judgment
  • A playground for experiments around Jev, TypeSafe's System One model.

    held because: category confidence 0.49 below 0.5

    LearningFeature ScoringRust
    genuine0.96Full judgment
  • Sidecar: grade a cloudforge student writeup with TypeSafe Jev. Not part of the OSS product.

    held because: category confidence 0.47 below 0.5

    ApplicationsFeature ScoringPython
    genuine0.96Full judgment
  • hl/jen

    1

    No description provided.

    held because: category confidence 0.44 below 0.5

    Agent and Dev ToolingVerificationGo
    genuine0.95Full judgment
  • macOS computer use driven by Jev (TypeSafe System One) as the decision maker

    held because: category confidence 0.44 below 0.5

    ApplicationsControl LoopGo
    genuine0.95Full judgment
  • No description provided.

    held because: category confidence 0.24 below 0.5

    OtherInfraJavaScript
    genuine0.95Full judgment
  • Bounded exploratory browser testing with Jev, deterministic assertions, and replayable evidence.

    held because: category confidence 0.39 below 0.5

    SDKs and ClientsControl LoopTypeScript
    genuine0.94Full judgment
  • I tortured Jev into being a RISC-V CPU.

    held because: category confidence 0.49 below 0.5

    Games and SimulationControl LoopPython
    genuine0.94Full judgment
  • Put your code comments on trial. Powered by Jev.

    held because: category confidence 0.49 below 0.5

    SDKs and ClientsFeature ScoringTypeScript
    genuine0.94Full judgment
  • No description provided.

    held because: category confidence 0.43 below 0.5

    Research and EvalsControl LoopTypeScript
    genuine0.93Full judgment
  • No description provided.

    held because: category confidence 0.43 below 0.5

    Agent and Dev ToolingRankingPython
    genuine0.92Full judgment
  • Jev (TypeSafe) 性能評価プロジェクト — 日本郵便 KEN_ALL をマスタに、AI SDK 経由の Jev が住所のあいまい一致にどこまで使えるかを検証

    held because: category confidence 0.43 below 0.5

    ApplicationsExtractionTypeScript
    genuine0.90Full judgment
  • Pre-install security gate for npm lifecycle scripts using TypeSafe System One.

    held because: category confidence 0.40 below 0.5

    SDKs and ClientsVerificationJavaScript
    genuine0.90Full judgment
  • No description provided.

    held because: category confidence 0.49 below 0.5

    ApplicationsRoutingPython
    genuine0.89Full judgment
  • An observable raw-character chat experiment powered entirely by TypeSafe Jev Choice

    held because: category confidence 0.41 below 0.5

    LearningControl LoopJavaScript
    genuine0.88Full judgment
  • Experimental multi-horizon BTC signal generator using TypeSafe Jev probabilities and Binance market data.

    held because: category confidence 0.43 below 0.5

    Research and EvalsControl LoopTypeScript
    genuine0.87Full judgment
  • Reliability and evaluation harness for financial agent workflows: agent harness + evaluation harness, release gates, evidence packs, DSPy programs vs typed tool-use (Jev). Reference prototype, synthetic and public benchmark data.

    held because: category confidence 0.44 below 0.5

    Agent and Dev ToolingRoutingPython
    genuine0.86Full judgment
  • Typesafe.ai System One Model Jev navigating a Neo4j graph by using a classifier over neighbouring relationships

    held because: category confidence 0.42 below 0.5

    LearningControl LoopJupyter Notebook
    genuine0.85Full judgment
  • No description provided.

    held because: category confidence 0.30 below 0.5

    LearningRoutingTypeScript
    genuine0.84Full judgment
  • High-speed recursive AI Elo tournament engine powered by Jev and Swiss matchmaking

    held because: category confidence 0.43 below 0.5

    ApplicationsRankingPython
    genuine0.83Full judgment
  • Hyprland/NixOS fork of typesafe-computer-use — grim + hyprctl + ydotool, TypeSafe Jev decisions

    held because: category confidence 0.31 below 0.5

    IntegrationsControl LoopPython
    genuine0.83Full judgment
  • One grounded Jev/Playwright core: typed SDK, persistent CLI, and MCP server with native browser operations and deterministic assertions.

    held because: category confidence 0.36 below 0.5

    Agent and Dev ToolingControl LoopTypeScript
    genuine0.82Full judgment
  • Classify Git commit diffs and messages with Jev. Bug fixes, security fixes/CWEs, and change types.

    held because: category confidence 0.49 below 0.5

    ApplicationsExtractionRust
    genuine0.81Full judgment
  • ACME live support-call scoring demo with TypeSafe AI, Effect, SQLite, React, Vite, and Turborepo

    held because: category confidence 0.44 below 0.5

    ApplicationsFeature ScoringTypeScript
    genuine0.79Full judgment
  • No description provided.

    held because: category confidence 0.47 below 0.5

    Agent and Dev ToolingRoutingTypeScript
    genuine0.78Full judgment
  • Open-source Jev log triage for OpenTelemetry. Score the signal before expensive LLM analysis.

    held because: category confidence 0.39 below 0.5

    ApplicationsFeature ScoringTypeScript
    genuine0.72Full judgment
  • Automated database migration safety reviewer powered by TypeSafe AI (Jev System One model)

    held because: category confidence 0.44 below 0.5

    Agent and Dev ToolingVerificationTypeScript
    genuine0.61Full judgment
  • my stuff for pi

    held because: genuine 0.59 below 0.6

    Agent and Dev ToolingVerificationTypeScript
    genuine0.59Full judgment
  • Autonomous PlayStation 2 AI Agent with real-time visual telemetry HUD powered by TypeSafe Jev System One

    held because: genuine 0.59 below 0.6

    Games and SimulationControl LoopPython
    genuine0.59Full judgment
  • Self-improving context compiler for CoreWeave Hacks: Agent Loops 2026

    held because: genuine 0.58 below 0.6

    Agent and Dev ToolingUnclearPython
    genuine0.58Full judgment
  • Vibe coding experiment

    held because: genuine 0.58 below 0.6

    Games and SimulationControl LoopPython
    genuine0.58Full judgment
  • System-architecture skill for TypeSafe AI Jev/System One — find fuzzy semantic judgment and turn it into small Choice/Score/Noul primitives.

    held because: genuine 0.58 below 0.6

    Agent and Dev ToolingInfra
    genuine0.58Full judgment
  • openvons (open-Jev): 有限選択肢に確率で答える判断層 — テキスト / 画像 / 日本語音声コマンド

    held because: genuine 0.56 below 0.6

    Research and EvalsControl LoopPython
    genuine0.56Full judgment
  • SO-101 robot-arm agent workbench: Bun/Effect coordinator, React workbench, Python LeRobot motor owner

    held because: genuine 0.55 below 0.6

    Games and SimulationControl LoopTypeScript
    genuine0.55Full judgment
  • Benchmarking TypeSafe's Jev decision model as a cost-efficient LLM router on RouterArena

    held because: genuine 0.55 below 0.6

    Research and EvalsRouting
    genuine0.55Full judgment
  • A simple way to deal with streaming text, tools, images from LLMs.

    held because: genuine 0.54 below 0.6

    SDKs and ClientsInfraGo
    genuine0.54Full judgment
  • Agent tool/MCP call gate — allow / ask_human / deny via TypeSafe Jev

    held because: genuine 0.54 below 0.6

    Agent and Dev ToolingVerificationTypeScript
    genuine0.54Full judgment
  • 346 reusable character attributes for AI agent personas — compose, measure, and port system-prompt personas. Build once. Keep personality everywhere.

    held because: genuine 0.52 below 0.6

    Agent and Dev ToolingRoutingPython
    genuine0.52Full judgment
  • Evidence-aware local RAG and HandoffProof: controlled causal testing for operational handovers, with optional TypeSafe/Jev evidence governance.

    held because: genuine 0.51 below 0.6

    ApplicationsVerificationPython
    genuine0.51Full judgment
  • No description provided.

    held because: genuine 0.51 below 0.6

    Agent and Dev ToolingUnclearZig
    genuine0.51Full judgment
  • Jev (TypeSafe System One) × ASReview SYNERGY abstract screening demo — Choice/Noul vs gold labels

    held because: genuine 0.51 below 0.6

    Research and EvalsRanking
    genuine0.51Full judgment
  • An experimental protocol for evidence-aware agent handoffs, bounded worker continuation, and TypeSafe/Jev-assisted review, with reproducible evaluation.

    held because: genuine 0.50 below 0.6

    Agent and Dev ToolingVerificationPython
    genuine0.50Full judgment
  • TypeSafe Jev research + 5 product specs

    held because: genuine 0.50 below 0.6

    OtherUnclear
    genuine0.50Full judgment
  • An evaluation of typesafe AI chess. As it turns out, the AI isn't doing really well even though chess is not a particularly open-ended game. Still, it's only a prototype and this probably wasn't optimzied for games.

    held because: genuine 0.49 below 0.6

    Research and EvalsUnclearPython
    genuine0.49Full judgment
  • Workload-aware Pareto frontiers and dynamic model selection for agentic systems

    held because: genuine 0.47 below 0.6

    Research and EvalsRoutingPython
    genuine0.47Full judgment
  • Generate anything from your terminal

    held because: genuine 0.46 below 0.6

    SDKs and ClientsRoutingTypeScript
    genuine0.46Full judgment
  • Jev-style parallel constrained decisions for any MLX model on Apple Silicon. Typed, schema-valid JSON in one forward pass.

    held because: genuine 0.46 below 0.6

    Research and EvalsExtractionPython
    genuine0.46Full judgment
  • Judge agent steps — ok / retry / escalate / stop via TypeSafe Jev

    held because: genuine 0.46 below 0.6

    Agent and Dev ToolingUnclearTypeScript
    genuine0.46Full judgment
  • PulseLane — clinic triage decisions via TypeSafe Jev

    held because: genuine 0.46 below 0.6

    ApplicationsUnclearTypeScript
    genuine0.46Full judgment
  • No description provided.

    held because: genuine 0.45 below 0.6

    Agent and Dev ToolingInfraPython
    genuine0.45Full judgment
  • A typed decision-routing prototype for turning medical learning material into deterministic study actions, designed to evaluate TypeSafe Jev.

    held because: genuine 0.44 below 0.6

    ApplicationsRouting
    genuine0.44Full judgment
  • No description provided.

    held because: genuine 0.43 below 0.6

    LearningUnclearPython
    genuine0.43Full judgment
  • Agent budget brake — continue / downgrade_model / stop via TypeSafe Jev

    held because: genuine 0.42 below 0.6

    ApplicationsUnclearTypeScript
    genuine0.42Full judgment
  • Self-correcting audiobook TTS generator (CoreWeave Hacks) — TypeSafe tagging + Google TTS/STT + ARIA/Weave eval-correction loop

    held because: genuine 0.42 below 0.6

    ApplicationsControl LoopPython
    genuine0.42Full judgment
  • Local bilingual probability decisions from context, questions, and candidate answers. Independent research preview inspired by TypeSafe Jev.

    held because: genuine 0.41 below 0.6

    Research and EvalsInfraPython
    genuine0.41Full judgment
  • No description provided.

    held because: genuine 0.41 below 0.6

    Agent and Dev ToolingVerificationGo
    genuine0.41Full judgment
  • Typed judgment layer for coding agents — gates from PRD to ship. Jev-ready, provider-agnostic.

    held because: genuine 0.40 below 0.6

    Agent and Dev ToolingVerificationTypeScript
    genuine0.40Full judgment
  • Match user goals to MCP catalog (two-stage) via TypeSafe Jev

    held because: genuine 0.39 below 0.6

    ApplicationsRankingTypeScript
    genuine0.39Full judgment
  • TypeSafe Jev MCP decision layer for coding agents and CI

    held because: genuine 0.39 below 0.6

    Agent and Dev ToolingUnclear
    genuine0.39Full judgment
  • No description provided.

    held because: genuine 0.39 below 0.6

    LearningUnclearTypeScript
    genuine0.39Full judgment
  • The open-source email platform — self-host on your own AWS SES, or use the cloud. Resend-compatible API.

    held because: genuine 0.38 below 0.6

    ApplicationsFeature ScoringTypeScript
    genuine0.38Full judgment
  • Filters an agent's memories to fit a token budget. Local, HTTP, MCP, Docker.

    held because: genuine 0.38 below 0.6

    Agent and Dev ToolingRankingPython
    genuine0.38Full judgment
  • Route tasks to research/code/browser/support/writer agents via TypeSafe Jev

    held because: genuine 0.38 below 0.6

    Agent and Dev ToolingRoutingTypeScript
    genuine0.38Full judgment
  • Typesafe AI SDK in Elixir using Req

    held because: genuine 0.38 below 0.6

    SDKs and ClientsInfra
    genuine0.38Full judgment
  • 3D chess powered by TypeSafe AI (Jev). AI vs AI by default, or play either side. Multiple difficulty levels.

    held because: genuine 0.37 below 0.6

    Games and SimulationUnclearTypeScript
    genuine0.37Full judgment
  • A polished OpenAI + TypeSafe Jev terminal interface for answers with transparent decision reports

    held because: genuine 0.37 below 0.6

    SDKs and ClientsUnclear
    genuine0.37Full judgment
  • Jev-style parallel constrained decisions for any MLX model on Apple Silicon. Typed, schema-valid JSON in one forward pass.

    held because: genuine 0.36 below 0.6

    Research and EvalsExtractionPython
    genuine0.36Full judgment
  • typesafe.ai model jev finance benchmark

    held because: genuine 0.36 below 0.6

    Research and EvalsUnclear
    genuine0.36Full judgment
  • U-NSGA-III multi-objective evolutionary optimization for .NET (Seada & Deb) — ZDT/DTLZ, IGD oracle vs pymoo

    held because: genuine 0.35 below 0.6

    OtherRankingC#
    genuine0.35Full judgment
  • LaneBreak — support ticket priority+routing via TypeSafe Jev

    held because: genuine 0.35 below 0.6

    ApplicationsUnclearTypeScript
    genuine0.35Full judgment
  • Typed-decision benchmark from PadFlow (land development SaaS): schemas, anonymized labeled rows, and a runner for confidence-calibrated models like TypeSafe Jev.

    held because: genuine 0.33 below 0.6

    Research and EvalsRoutingPython
    genuine0.33Full judgment
  • Autonomous prospecting for freelancers.

    held because: genuine 0.33 below 0.6

    ApplicationsUnclearHTML
    genuine0.33Full judgment
  • HireSignal — resume first-pass fit+interview via TypeSafe Jev

    held because: genuine 0.33 below 0.6

    ApplicationsFeature ScoringTypeScript
    genuine0.33Full judgment
  • NPCs of River Oaks Houston, Texas using Jev to power NPCs

    held because: genuine 0.33 below 0.6

    Games and SimulationUnclear
    genuine0.33Full judgment
  • A Jev-inspired decision interface for existing LLMs. Explicit choices, scores, calibration, and review thresholds.

    held because: genuine 0.32 below 0.6

    IntegrationsRoutingHTML
    genuine0.32Full judgment
  • Offline evaluation for document extraction

    held because: genuine 0.32 below 0.6

    OtherVerificationPython
    genuine0.32Full judgment
  • A test project based on Jev AI, the goal is to build a search function for a blog/article website that has 100s of articles to search from, So the user can actually use the search as chat to question anything and find related answers/articles

    held because: genuine 0.31 below 0.6

    ApplicationsUnclearJavaScript
    genuine0.31Full judgment