AI AGENT ADDONS

Phoenix Evals

arize-ai/phoenix

Phoenix Evals helps you build tests for your AI applications. These tests check if your app responds correctly. You can use your own code or a language model to judge answers. Then you can compare results with human ratings to see how well your app works.

You can start with built in test types or create your own. The tool works with Python and TypeScript. It also lets you run experiments on large sets of data. This makes it easy to find and fix problems in your AI app.

Many teams use Phoenix Evals to improve their apps before going live. It helps catch mistakes and ensures your app is reliable and accurate.

Add Phoenix Evals skill to your workflow

Global

mkdir -p ~/.claude/skills/phoenix-evals

Project

mkdir -p .claude/skills/phoenix-evals

Source Repository

Stars
10,307
Forks
946
Watchers
10,307
License
NOASSERTION
Last Push
23 days ago
Created
3 years ago