Back to benchmarks & research
Benchmarks & research

jev-reward-model-evaluation

Jev 1.13 reward-model evaluation across 8 benchmark tracks, with an interactive report and 54-row SOTA comparison

Added
2026-09-20
Language
Python
Stars at snapshot
1

Source: community catalog. Details and metrics reflect the published community snapshot and are not a performance endorsement.