Fast, typed, calibrated evaluations for LLM and agent outputs, powered by Jev — with simple, framework-agnostic Python APIs
ai evaluation evaluation-metrics evaluation-framework ai-agents rag llm langchain evals llm-evaluation crewai microsoft-agent-framework jev system-one-models typed-evals
-
Updated
Sep 20, 2026 - Python