Back to benchmarks & research
jevalyzer project preview
Benchmarks & research

jevalyzer

Grade the agent sessions already on your disk. Claude Code, Codex, opencode, Gemini CLI and Antigravity, scored with Jev for cents.

How it uses Jev

The scoring runs on Jev , TypeSafe AI's System One model, through the Vercel AI Gateway. Jev doesn't write text: you give it a state and a set of typed questions, and it returns booleans, choices and scores as calibrated probabilities in one round trip. That is what makes grading a whole history cheap enough to bother with — a typical user pays one to thirty cents for their entire archive.

Languages
TypeScript 99% · JavaScript 1%
License
MIT
Last activity
Sep 2026
Added
2026-09-23
Stars at snapshot
2
Status
Community listing
ai-agentsbunclaude-codeclicodexgemini-cli

Source: community catalog. Details and metrics reflect the published community snapshot and are not a performance endorsement. Documentation details extracted from GitHub · the README · the project website.

Listed in Jev Library

If this is your project, the badge below links back to this page. Paste it near the top of your README.

[![Listed in Jev Library](https://jevlibrary.dev/api/badge/killerz3-jevalyzer)](https://jevlibrary.dev/projects/killerz3-jevalyzer)

Something wrong on this page, or would you rather not be listed? Tell us and it will be changed.