jev-no-enem
patryckalves
Reproducible benchmark evaluating TypeSafe AI's Jev (System One paradigm) on Brazil's ENEM 2025 standardized exam. Evaluates typed decision-making, domain-specific accuracy, and RLCD uncertainty calibration against open LLM baselines with an interactive GitHub Pages dashboard.
- License
- Not declared
- GitHub Stars
- 1
- Source reviewed
- —
Where Jev makes a decision
Jev returns a structured decision for the local program; consult the source for the exact decision policy.
What this project offers
Adds structured choices or scores to the workflow; performance and cost benefits have not been independently verified.
Review scope
Based on repository metadata and README with rule-based classification; pending human review, with no independent runtime or performance verification.
Sources and implementation
Related projects
jev-playground
Little-Planet-Labs · Decision Tools
A web playground for entering state and decision questions, then inspecting Jev answers and probability distributions.
jev-predict-skill
DanielKillenberger · Decision Tools
An Agent skill recipe that predicts another skill’s closed-set outcome from its rules and evidence.
jevify
altryne · Decision Tools
An Agent Skill for finding suitable Jev decision points and designing questions and comparison experiments.
jevchat
kt3k · Decision Tools
A chat-style Jev demo whose answers are selected from predefined or custom options rather than generated prose.