Back to benchmarks & research
jevals project preview
Benchmarks & research

jevals

Agent evals and guardrails in one request. Built on Jev, Kev and Laya.

How it uses Jev

Evals and guardrails for agents, using Jev-style decision models instead of an LLM judge. All the evals for a trace go out as one request that costs a few thousandths of a cent and comes back in a few hundred milliseconds, so you can run them on every trace and inside the agent loop.

Language
Python 100%
License
MIT
Latest release
jevals 0.1.1
Last activity
Sep 2026
Added
2026-09-21
Stars at snapshot
12
Status
Community listing
agentsevalsguardrailsjevllmllm-evaluation

Source: community catalog. Details and metrics reflect the published community snapshot and are not a performance endorsement. Documentation details extracted from GitHub · the README · the project website.