Benchmark
Description
--- name: benchmark description: "Compare skill scores against ideal benchmarks" usage: "/benchmark [skill-name]" --- # /benchmark — Compare Against Ideal Benchmarks Compare a skill's historical performance against defined benchmark standards. ## Arguments - `skill-name` (required): The skill to benchmark. ## What to Do 1. Read benchmark standards from `skills/judge/references/benchmark-standards.md` 2. Read historical scores for the specified skill from `skills/judge/scores/` 3. Compute a
Installation
Installs to ~/.claude/skills/walllmat-verdict-benchmark/SKILL.md
mkdir -p ~/.claude/skills/walllmat-verdict-benchmark && curl -fsSL https://raw.githubusercontent.com/Walllmat/verdict/HEAD/commands/benchmark.md -o ~/.claude/skills/walllmat-verdict-benchmark/SKILL.md Restart Claude Code, or start a new session, for it to be picked up.
Full documentation available on GitHub
View Source RepositoryRelated Skills
Awesome Go
A curated list of awesome Go frameworks, libraries and software
Development next.js
| The React Framework | 138360 | 1503 | 1 |
Development sharing-skills
skill for guidance.
Development root-cause-tracing
Use when errors occur deep in execution and you need to trace back to find the original trigger.
Development Template Skill
Minimal skeleton for a new skill project structure.
Development Third-party Notices
THE FOLLOWING SETS FORTH ATTRIBUTION NOTICES FOR THIRD PARTY SOFTWARE THAT MAY BE CONTAINED IN PORTIONS OF THI
Development