Evidently
An open-source framework for evaluating and monitoring AI systems, with more than a hundred built-in metrics plus LLM-as-judge evaluators and drift detection, rendered as interactive reports.
Image from the business's own website.
Services
- 100+ built-in metrics
- LLM-as-judge evaluators
- Drift monitoring
- Interactive reports
Highlights
Source: Official GitHub repository. Last checked 2026-09-18. Spot an error? Tell us.