AI AGENT ADDONS

Evaluation Methodology

wshobson/agents

Plugin quality is measured with a clear and detailed system. Three different layers check how well a skill works. Static analysis looks at the structure and writing. An AI judge reviews the skill by running example tasks. A simulation tests the skill with many real prompts. Each layer gives a score for ten dimensions like triggering accuracy and output quality. These scores combine into one final score. Badges show if a skill is excellent or needs work. This helps you understand why a score is low and how to make your skill better.

Add Evaluation Methodology skill to your workflow

Global

mkdir -p ~/.claude/skills/evaluation-methodology

Project

mkdir -p .claude/skills/evaluation-methodology

Source Repository

Stars
37,285
Forks
4,009
Watchers
37,285
License
MIT
Last Push
1 month ago
Created
1 year ago