Back to benchmarks & research
Benchmarks & research

jev-rerank-bench

Can a decision model beat dedicated rerankers? TypeSafe Jev vs Cohere Rerank 4 vs ZeroEntropy zerank-2 vs a chat-model baseline: 14 datasets, every raw API response, bootstrap ranges on every gap.

Added
2026-09-17
Language
Python
Stars at snapshot
2

Source: community catalog. Details and metrics reflect the published community snapshot and are not a performance endorsement.