Jev Score — Development skill for Claude Code
Local-first document evaluation workspaces powered by Jev.
How to install Jev Score
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open a-Fig/jev-score and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Jev Score does
Local-first document evaluation workspaces powered by Jev.
Alternatives in Development
- IPolloWork — Enterprise-grade, local-first Agent Workbench for people and agent teams 4.9k ★
- Claude Prism — An offline-first scientific writing workspace powered by Claude 1.1k ★
- Alethe Agents — A local-first desktop workspace for running, organizing, and resuming multiple coding agents and shells with r 511 ★
README
Jev Score
[](https://github.com/a-Fig/jev-score/actions/workflows/ci.yml) [](https://github.com/a-Fig/jev-score/releases/latest) [](https://nodejs.org/) [](./LICENSE)
Give your coding agent a scoreboard
Most AI writing loops run on vibes: draft, rewrite, hope. **Jev Score is a test suite for writing.** It scores every revision against the context the document must satisfy and the questions that define "good", and keeps the history so you can see whether draft 8 actually beats draft 3.
It is a local CLI and web app for developers who already let Codex or Claude Code edit their text: a resume against a job posting, an essay against its prompt, a spec against requirements, a landing page against its positioning. [TypeSafe's Jev model](https://www.typesafe.ai/) does the scoring through OpenRouter; drafts and scores stay in local SQLite.
- Numbers per question, not vibes. On the bundled resume example, one agent revision aimed at the two weakest questions moved the median score from 83.2 to 96.6 (the run).
- Every run is kept, shown as medians and ranges, because Jev is probabilistic.
- CLI for your agent, UI for you. Both read one database: JSON out of the CLI; in the browser, a progress chart and side-by-side or diff comparison of any two drafts.

**Jump to:** [Install](#install) · [First score](#your-first-score) · [Concepts](#concepts) · [Agent loop](#the-agent-loop) · [Limits](#honest-limits)
Evidence: this project's own README
That screenshot is a real workspace holding this project's README: 3 drafts, 14 evaluation
Related Skills
Agent Blackbox
Local-first flight recorder for coding agents : replay every run as a live session map, score the context bill
Jev Gates
Three gates for any coding agent: a deterministic approval gate before irreversible actions, a completion gate
Itr Agent
Local-first MCP server for Indian income tax: deterministic tax computation, regime comparison, advance tax pl
Quick Eval
Quick job evaluation. Paste a JD and get a score plus one-paragraph summary. Faster than a full evaluate. Use
Research Codebase Nt
Document codebase as-is without evaluation or recommendations
Naive Harness Kit
Prompt-first starter kit for Codex and Claude Code workspaces, with bootstrap, upkeep, and archive workflows.
Related Agents
Claude Code Slack Bot
This is a TypeScript-based Slack bot that integrates with the Claude Code SDK to provide AI-powered coding ass
Skill Health Observer
Analyze skill outcomes, score per-skill health, write SKILL_HEALTH.md. Haiku-powered, ~$0.02/run.
AI Product Designer
The AI Product Designer designs LLM-, agent-, and ML-powered features inside the app: prompt UX, guardrails, l