AI AGENT ADDONS

Promptfoo Evaluation

daymade/claude-code-skills

Promptfoo helps you test and compare different language models side by side. You can set up automated checks to see which model gives the best answer. This tool makes it easy to run LLM evaluations without writing a lot of code.

Anyone working with language models can use this skill. It guides you through creating configuration files, writing custom tests, and using LLM-as-judge to rate responses. You can also add few-shot examples to your prompts for better results.

With Promptfoo Evaluation, you get clear reports on model performance. It saves time and helps you pick the right model for your task.

Add Promptfoo Evaluation skill to your workflow

Global

mkdir -p ~/.claude/skills/promptfoo-evaluation

Project

mkdir -p .claude/skills/promptfoo-evaluation

Source Repository

Stars
1,223
Forks
199
Watchers
1,223
License
MIT
Last Push
23 days ago
Created
9 months ago