Skip to content
#

ai-quality

Here are 91 public repositories matching this topic...

Evaluate your LLM apps with one function call. Hallucination detection, RAG scoring, and agent evals for OpenAI, Anthropic, and more. 14 evaluators, pytest plugin, composite trust scores.

  • Updated Apr 3, 2026
  • Python

Add this topic to your repo

To associate your repository with the ai-quality topic, visit your repo's landing page and select "manage topics."

Learn more