jevals: replacing LLM judges with typed Jev decisions for agent evals
jevals is an open-source tool that replaces LLM judges with typed Jev decision models for agent evals and guardrails. All evals for a trace are sent as one HTTP request costing a few thousandths of a cent and returning in a few hundred milliseconds, suitable for every trace or inside the agent loop. It supports Jev, Kev, Laya, and regular chat LLMs.