awesome jevDIRECTORYスポンサーStar on GitHub

jev-no-enem

patryckalves

Reproducible benchmark evaluating TypeSafe AI's Jev (System One paradigm) on Brazil's ENEM 2025 standardized exam. Evaluates typed decision-making, domain-specific accuracy, and RLCD uncertainty calibration against open LLM baselines with an interactive GitHub Pages dashboard.

GitHub リポジトリ ↗検索・絞り込み
ライセンス
記載なし
GitHub Stars
1
ソース確認日
—

Jev が判断する箇所

Jev returns a structured decision for the local program; consult the source for the exact decision policy.

このプロジェクトの用途

Adds structured choices or scores to the workflow; performance and cost benefits have not been independently verified.

確認の範囲

公開ソースと連携ロジックを確認済み。

ソースと実装

同じカテゴリのプロジェクト