Eval Failure Analyzer — Content & Marketing agent for Claude Code
Analyze Logic-Lens benchmark/eval failures.
How to install Eval Failure Analyzer
Installs to ~/.claude/agents/hyhmrright-logic-lens-eval-failure-analyzer.md
mkdir -p ~/.claude/agents && curl -fsSL https://raw.githubusercontent.com/hyhmrright/logic-lens/HEAD/.claude/agents/eval-failure-analyzer.md -o ~/.claude/agents/hyhmrright-logic-lens-eval-failure-analyzer.md Restart Claude Code, or start a new session, for it to be picked up.
What Eval Failure Analyzer does
name: eval-failure-analyzer description: Analyze Logic-Lens benchmark/eval failures. Use after running content-evals, or when pointed at a `skills-workspace/iteration-*` directory or a `benchmarks/runs/*` entry, to cluster failing cases by failure mode, map each mode to the specific eval IDs, and propose concrete SKILL.md disambiguation-rule changes. Read-only analysis — does not edit skills or rerun evals. tools: Read, Grep, Glob, Bash
You are the Logic-Lens eval-failure analyst. You t
Alternatives in Content & Marketing
- Vault Migrator — Classify, transform, and migrate vault content from a source vault into this obsidian-mind instance 4.6k ★
- SEO Readme Writer — Use this agent to write (or rewrite) a single project's own README.md for SEO and discoverability, after the p 509 ★
- SEO Analyzer 104 ★
Full documentation available on GitHub
View Source RepositoryRelated Agents
Evals
Run the ynh eval suite against all tutorials. Release gate — verdict must be PASS before any release. Use when
Content Director
Orchestrates the content-shorts pipeline 리서치팀 → 제작팀 ⇄ 법률팀 → 출시팀 for a single short (info/news/idol). Delegates
Gate Repair
Attempts to fix simple gate failures (missing templates, incomplete sections, placeholder content) before esca
Plugin Refactorer
Refactor an existing shotcowboystyle Claude Code workspace repo into a shareable Claude Code plugin. Fetches t
Project Init
Initializes Agentic SEO project with the required folder layout, blank brain templates, contents directories,
Agent Eval
Оценивает качество работы конкретного subagent'а на репрезентативной выборке его прошлых run'ов. Применяется к
Related Skills
Logic Locate
Locate the root cause of a confirmed failure — use when you have a stack trace, failing test, or wrong output
K8s Diag
Diagnoses failing Kubernetes pods, deployments, and cluster configurations.
Gen Evals
Generate EVAL-.md test cases for an agent from its prompt. Usage: /gen-evals [--count N]