Ai Evaluator banner
Productfculty-aipm Productfculty-aipm

Ai Evaluator

Data & AI community

Description

--- name: ai-evaluator description: > Designs and runs AI product evaluation frameworks: error analysis, eval suite design, LLM-as-judge pipelines, human eval protocols, regression testing plans, and improvement flywheels. Use this agent when the user is building an AI-powered feature and needs to define how to measure quality, catch regressions, or systematically improve model outputs. <example> Context: User shipped an AI feature and is seeing quality complaints but can't quanti

Installation

Installs to ~/.claude/agents/productfculty-aipm-pm-copilot-by-product-faculty-ai-evaluator.md

Terminal
mkdir -p ~/.claude/agents && curl -fsSL https://raw.githubusercontent.com/Productfculty-aipm/PM-Copilot-by-Product-Faculty/HEAD/agents/ai-evaluator.md -o ~/.claude/agents/productfculty-aipm-pm-copilot-by-product-faculty-ai-evaluator.md

Restart Claude Code, or start a new session, for it to be picked up.

Full documentation available on GitHub

View Source Repository