jev-rerank-bench
anessbelbati
Compares Jev, dedicated rerankers and chat models on the same retrieved passages.
- License
- MIT
- GitHub Stars
- 1
- Source reviewed
- 2026-09-19
Where Jev makes a decision
Ranks candidate passages with Choice, Noul and rubric scores, then computes retrieval metrics.
What this project offers
Publishes raw responses, scoring code and per-dataset results for inspection.
Review scope
The author reports equal-dataset nDCG@10 of 0.692 for Jev and 0.691 for Cohere on eight English datasets, without establishing a winner. Query weighting changes the comparison. README and integration source reviewed at a fixed commit; not independently run or benchmarked by this site.
Sources and implementation
Related projects
jev-review
NiazMorshed2007 · Evaluation & Observability
A local MCP code-quality reviewer returning structured scores to coding Agents.
jev-benchmarks
AbdelStark · Evaluation & Observability
A benchmark comparing Jev and GLiNER on text classification, probability calibration and selective automation.
jev-playground
mizchi · Evaluation & Observability
A MoonBit and TypeScript Jev playground covering games, browsers, command risk and small languages.
jev-lm
y0usaf · Evaluation & Observability
A word-level generation experiment that asks Jev to select words or verify locally drafted continuations.