AI AGENT ADDONS
AI & Agent Building
1,469installs

Online evaluations help you check how well your AI responds to users. You attach judges to your AI configs to automatically score each answer. Judges use an LLM to give a score between 0.0 and 1.0.

You can use built-in judges for accuracy, relevance, and safety. You can also create your own judges for specific needs. This helps you monitor quality and improve your AI over time.

Anyone managing AI responses can set up evaluations with just a few steps. No complicated setup needed. Just a LaunchDarkly account and API token.

Add Online Evals skill to your workflow

Global

mkdir -p ~/.claude/skills/online-evals

Project

mkdir -p .claude/skills/online-evals

Source Repository

Stars
17
Forks
5
Watchers
17
License
NOASSERTION
Last Push
1 month ago
Created
5 months ago