Jev in production › Evaluation and testing
LangChain tested Jev as an agent-eval judge against GPT-5.6 and Claude Sonnet 4.6 judges on identical traces, then added Jev-as-a-judge evaluators to LangSmith Evals. Jev cost $0.00035/call (company).
For the project's own README, linking back here: