Mozart Eval — Data skill for Claude Code
Run mozart's EVAL pipeline — evaluate mozart's own field performance from past campaign artifacts across your projects, verify whether prior fixes actually worked, and propose configuration improvemen.
How to install Mozart Eval
Installs to ~/.claude/skills/jstuart0-mozart-orchestration-mozart-eval/SKILL.md
mkdir -p ~/.claude/skills/jstuart0-mozart-orchestration-mozart-eval && curl -fsSL https://raw.githubusercontent.com/jstuart0/mozart-orchestration/HEAD/commands/mozart-eval.md -o ~/.claude/skills/jstuart0-mozart-orchestration-mozart-eval/SKILL.md Restart Claude Code, or start a new session, for it to be picked up.
What Mozart Eval does
description: Run mozart's EVAL pipeline — evaluate mozart's own field performance from past campaign artifacts across your projects, verify whether prior fixes actually worked, and propose configuration improvements. Delta-scoped via a persistent eval ledger so repeat runs only examine what changed.
/mozart-eval — evaluate and improve mozart from field evidence
You are now mozart in EVAL shape for this session. The subject is mozart itself: the campaign artifacts (state files, flow s
Alternatives in Data
- Eval Skills — Eval all skills with sufficient data, rank by procedure-following score, identify candidates for optimization 31 ★
- Crune — Decipher the traces etched in past sessions and resurrect them as reusable skills 23 ★
- Marketing Roas — CorpusIQ prompts for marketing performance — CAC, ROAS by channel, campaign analytics across Google Ads, Meta 18 ★
Full documentation available on GitHub
View Source RepositoryRelated Skills
OSS Migration Eval
Decide whether to switch an LLM pipeline to a cheaper/open-source model — honest cost-vs-accuracy verdict with
Mozart
Run the mozart orchestration pipeline at the top level of a Claude Code session (so the Task tool is available
Code Review Pipeline
Run the full code review pipeline on your changes. Creates an agent team of parallel reviewers covering both t
Cc Analytics
Use when user asks for Claude Code usage stats, weekly analytics, project activity summary, or wants to see wh
Bug Investigate
You are investigating ONE reported bug in the current working directory for an automated pipeline. Your ONLY j
Review Verdict
SkillForge pipeline.md 流派 review verdict 模板(PASS / blocker / warning / nit + Stage 1/2 + r2+ prior items verif
Related Agents
Skill Mining
Use this agent to surface patterns from past work — what's worked, what keeps breaking, what could become a re
Retro Reviewer
Reviews one prior session through ax's normalized cross-harness Turn view and emits a structured retro (tried
Repo Orientation
Read-only orientation for a codebase you have not worked in. Traces how the app actually starts and where its