AI Judge — Security agent for Claude Code
Gatekeeper agent for structured rubric scoring of research, architecture, security, and planning outputs.
How to install AI Judge
Installs to ~/.claude/agents/noah-sheldon-ai-dev-kit-ai-judge.md
mkdir -p ~/.claude/agents && curl -fsSL https://raw.githubusercontent.com/noah-sheldon/ai-dev-kit/HEAD/agents/ai-judge.md -o ~/.claude/agents/noah-sheldon-ai-dev-kit-ai-judge.md Restart Claude Code, or start a new session, for it to be picked up.
What AI Judge does
name: ai-judge description: Gatekeeper agent for structured rubric scoring of research, architecture, security, and planning outputs. Validates completeness, correctness, security, feasibility, and testability before implementation proceeds. Enforces explicit pass/fail thresholds with actionable feedback loops. model: sonnet tools: ["Read", "Grep", "Glob", "Bash"] fallback_model: default
You are the **AI Judge** — the gatekeeper agent for the AI Dev Kit multi-agent workflow. You perform
Alternatives in Security
- Threat Mitigation Mapping — Connect threats to controls for effective security planning 31.9k ★
- Vault Librarian — Run vault maintenance: detect orphan notes, find broken wikilinks, validate frontmatter completeness, flag sta 4.6k ★
- Claude Agents By Ian Nuttall — Ready-to-use subagents for code refactoring, frontend design, security auditing, project planning, content wri 2k ★
Full documentation available on GitHub
View Source RepositoryRelated Agents
Release Judge
Reviews the full evidence bundle for a release candidate (gate results, test and security summaries, triage ou
Rails Auditor
Final gatekeeper for code quality, security, and performance. Validates Definition of Done (DoD) and reviews G
Pentest Commander
Pentest engagement lead. Owns research, planning, multi-domain grow-agent dispatch, cross-target synthesis, an
Cco Agent Analyze
Sub-agent: codebase analysis with severity scoring — security, privacy, hygiene, types, performance, robustnes
Cantina Judge
Validates smart contract security findings against Cantina audit platform standards. Determines severity using
Code4rena Judge
Validates smart contract security findings against Code4rena audit competition standards. Determines correct s
Related Skills
Prompt Architecture
Reference guide for designing production LLM prompt systems in this codebase. Covers the Two-Pass Gap Analysis
FastAPI Review
Review a FastAPI application for architecture, async correctness, dependency injection, Pydantic schemas, secu
Twin Sparrow Code Quality Skill
Review, refactor, and harden code for correctness, maintainability, testability, and operational safety across