Issue 2026-09-21 · Industry · 开源 · 安全 · 研究
jevals: replacing LLM judges with typed Jev decisions for agent evals

jevals is an open-source tool that replaces LLM judges with typed Jev decision models for agent evals and guardrails. All evals for a trace are sent as one HTTP request costing a few thousandths of a cent and returning in a few hundred milliseconds, suitable for every trace or inside the agent loop. It supports Jev, Kev, Laya, and regular chat LLMs.
Read original ↗