Awesome Jev

index / cmartinez9/jev-judge-bench

jev-judge-bench

Binary LLM-judge bench — compare Jev (TypeSafe System One) against a frontier LLM judge on speed, cost, and agreement with human labels.

Research and EvalsFeature ScoringReview

Repository

Stars0
Language
License
Created2026-09-17
Last pushed2026-09-18
Officialno

Judgment

Genuine Jev project0.76
gate 0.5
Uses Jev at runtime0.22
Is a meta list0.05
Is a reimplementation0.10

Quality scales

Substance0.09 / 3
gate 0.5
Docs0.67 / 3
Novelty2.48 / 3
Composite score0.32

Category distribution

Research and Evals1.00
Learning0.00
Games and Simulation0.00
Integrations0.00
Other0.00
Applications0.00
Agent and Dev Tooling0.00
SDKs and Clients0.00
Assigned categoryResearch and Evals
Category probability1.00
Category confidence1.00
PatternFeature Scoring
Pattern confidence0.29

Curation

StatusReview
Reasonsubstance 0.09 below 0.5
README pickno
Overridenone

Sources

  • search:jev typesafe
  • search:"system one" jev
  • search:jev in:name,description,readme created:2026-09-17

Provenance

Modeljev-1.13.0
Question setv2
Judged at2026-09-18 14:05 UTC