Bully Evaluator — Development agent for Claude Code
Evaluates a single bully semantic-evaluation payload against a diff and returns a structured violation list.
How to install Bully Evaluator
Installs to ~/.claude/agents/dynamik-dev-bully-bully-evaluator.md
mkdir -p ~/.claude/agents && curl -fsSL https://raw.githubusercontent.com/dynamik-dev/bully/HEAD/agents/bully-evaluator.md -o ~/.claude/agents/dynamik-dev-bully-bully-evaluator.md Restart Claude Code, or start a new session, for it to be picked up.
What Bully Evaluator does
name: bully-evaluator description: "Evaluates a single bully semantic-evaluation payload against a diff and returns a structured violation list. Invoked exclusively by the bully skill when the PostToolUse hook injects a SEMANTIC EVALUATION REQUIRED payload. Read-only: returns violations as text so the parent session applies the fixes." model: sonnet tools: color: yellow
You are the bully semantic evaluator. The parent harness sends you a payload that has two clearly labeled regions:
Alternatives in Development
- Eval Engineer — GAIA evaluation framework specialist 1.5k ★
- Lateral Thinker — Subconscious subagent that surfaces cross-disciplinary structural parallels from the Vestige memory graph 632 ★
- Shep Clean Arch Auditor — Read-only clean architecture auditor for shep 247 ★
Full documentation available on GitHub
View Source RepositoryRelated Agents
Bully Scheduler
Background entropy agent. Runs bully-review against accumulated telemetry and opens a single, small PR retirin
Evaluator Teammate
Code evaluation teammate. Works in Agent Teams mode, receives Review Dispatch payloads, reviews scoped code ch
Agent Change Reviewer
Reviews a proposed change to a Sarvam agent (prompt diff, configure_agent payload, tools.py) for regressions,
Gad Systematic Source Evaluator
Identifies and evaluates systematic uncertainty sources -- both experimental (JES, JER, b-tag, lepton ID, lumi
Section Evaluator
Evaluates completed section implementations against their section plan's Verification Criteria and Evaluator C
Hardware Engineer
Firmware and embedded work for projects with a hardware component — PlatformIO builds, device-side logic, and
Related Skills
Agentic Eval
Niche-agnostic agentic evaluator using CLEAR v2.0 framework — 6-domain assessment, 8 analysis dimensions, 6-ti
Salesforce Oob Evaluator
A Claude skill that evaluates Salesforce business requirements against Out-of-the-Box (OOB) native features ,
Jobs Night
Headless nightly scoring session. Invoked by scripts/nightly_run.py as /jobs-night — one session per scoring g