jevcal
abhixhek
A toolkit for evaluating Jev probabilities on labeled data, selecting confidence thresholds and checking model drift.
- License
- MIT
- GitHub Stars
- 5
- Source reviewed
- 2026-09-19
Where Jev makes a decision
Runs fixed questions and measures accuracy, calibration, coverage and escalation rates.
What this project offers
Connects threshold selection and model-change checks to reports and CI.
Review scope
The README demo table uses a simulator, not measured Jev results; production thresholds need workload-specific validation. Not run here.
Sources and implementation
Related projects
jev-review
NiazMorshed2007 · Evaluation & Observability
A local MCP code-quality reviewer returning structured scores to coding Agents.
jev-benchmarks
AbdelStark · Evaluation & Observability
A benchmark comparing Jev and GLiNER on text classification, probability calibration and selective automation.
jev-playground
mizchi · Evaluation & Observability
A MoonBit and TypeScript Jev playground covering games, browsers, command risk and small languages.
jev-lm
y0usaf · Evaluation & Observability
A word-level generation experiment that asks Jev to select words or verify locally drafted continuations.