すべてのプロジェクト

jev-lab

llt22/jev-lab

Hands-on research lab for TypeSafe's Jev (System One model): reproducible benchmarks of Noul/Choice/Score primitives, confidence gating, fan-out latency, agent control — plus a living audit of the Jev ecosystem.

研究と評価Python
スター
1
フォーク
0

審査時の参照元

参照元を見る

トピック

jevai-agentsbenchmarkconfidence-calibrationevalsllmllm-evaluationllm-observability