Promptfoo helps you test and compare different language models side by side. You can set up automated checks to see which model gives the best answer. This tool makes it easy to run LLM evaluations without writing a lot of code.
Anyone working with language models can use this skill. It guides you through creating configuration files, writing custom tests, and using LLM-as-judge to rate responses. You can also add few-shot examples to your prompts for better results.
With Promptfoo Evaluation, you get clear reports on model performance. It saves time and helps you pick the right model for your task.
Global
mkdir -p ~/.claude/skills/promptfoo-evaluationProject
mkdir -p .claude/skills/promptfoo-evaluationSource Repository
Find Skillsvercel-labs/skills
Find and install the perfect skill to extend your AI agent
Microsoft Foundrymicrosoft/azure-skills
Build, deploy, and improve AI agents on Microsoft Foundry from start to finish
Azure Aimicrosoft/azure-skills
Search, transcribe, and analyze with Azure AI tools for smarter apps
Azure Hosted Copilot Sdkmicrosoft/azure-skills
Build, deploy, and manage your Copilot SDK apps on Azure with ease
Triagemattpocock/skills
Triage issues with a state machine driven by clear roles and agent briefs
Handoffmattpocock/skills
Hand off your work to another AI agent with a clear summary
Image Editagentspace-so/runcomfy-agent-skills
Smart router picks the best AI model for your image editing needs
Agentspaceagentspace-so/skills
See your AI agent's live folder from any browser instantly