Add Eval
Description
--- name: add-eval description: Add automated eval (LLM-as-judge or deterministic) for an AI-powered feature user-invocable: true --- Add eval for: $ARGUMENTS ## What Are Evals? Evals are automated tests that measure the quality of AI-powered outputs. Unlike unit tests (pass/fail on exact output), evals measure: - Accuracy (is the answer factually correct?) - Relevance (does it answer what was asked?) - Format compliance (does it follow the output schema?) - Safety (does it refuse harmful req
Installation
Installs to ~/.claude/commands/srikanthvemulapally-ai-native-boilerplate-add-eval.md
mkdir -p ~/.claude/commands && curl -fsSL https://raw.githubusercontent.com/SrikanthVemulapally/ai-native-boilerplate/HEAD/.claude/commands/add-eval.md -o ~/.claude/commands/srikanthvemulapally-ai-native-boilerplate-add-eval.md Restart Claude Code, or start a new session, for it to be picked up.
Full documentation available on GitHub
View Source RepositoryRelated Skills
Agency Agents
A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy inject
AI Awesome Llm Apps
100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
AI Firecrawl
🔥 The API to search, scrape, and interact with the web for AI
AI Artifacts Builder
Suite of tools for creating elaborate, multi-component claude.ai HTML artifacts using modern frontend web tech
AI Headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agen
AI CrewAI
Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewA
AI