Shipworthy — Testing skill for Claude Code
AI-agent-driven end-to-end & frontend testing + backend-symptom analysis — free, read-only Claude Code/Codex skill.
How to install Shipworthy
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open NeuraCerebra-AI/Shipworthy and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Shipworthy does
AI-agent-driven end-to-end & frontend testing + backend-symptom analysis — free, read-only Claude Code/Codex skill.
Alternatives in Testing
- AI Contribution Skills (Engram) — branch-pr: clean branch and PR workflow 2.5k ★
- Playwright Automation — git clone Browser automation skill for end-to-end testing and web interaction 2.3k ★
- Spec Kitty — Spec-Driven Development for serious software developers 1.6k ★
README
Automated frontend & end-to-end (E2E) testing + backend-symptom analysis — an AI agent that proves your app is ready to ship.
It walks your whole product like your most paranoid senior engineer — every safe, discoverable screen and path, plus the backend underneath — then proves whether you're ready to ship.
Point it at your app, ask **"Are we shipworthy?"**, and get a proof-backed ship-or-don't verdict.
[](https://github.com/NeuraCerebra-AI/shipworthy) [](LICENSE) [](#-the-four-skills) [](https://github.com/NeuraCerebra-AI/shipworthy/releases)
**Read-only · self-contained markdown · no telemetry · no credential access · no auto-update**
**Shipworthy** is a free, open-source **AI testing agent** that runs **automated frontend and end-to-end (E2E) tests** on your web app — and analyzes the **backend symptoms** behind each path. It walks every safe user path like a real user (browser, Playwright, or Computer Use), catches failures the UI hides (a 500 behind a "success" screen, data that never saved), installs as a **Claude Code or Codex skill**, runs **read-only**, and returns a proof-backed *ship-or-don't* verdict: **ready, conditionally ready, not ready, or cannot determine**. It works on normal apps and on AI-built ("vibe-coded") apps and AI agents — and it never overclaims.
😱 What silently breaks without this
Most "it works on my machine
Related Skills
Onecommand
Build a complete, production-ready software system from a single prompt. Orchestrates Claude and Codex through
Twin Sparrow Expert Engineer
Multi-domain engineering skill covering backend, frontend, TypeScript, debugging, architecture, performance, t
Test Verify Backend
Explores back-end testing for a use case
Performance Test
Run performance profiling and load testing using the performance-engineer agent — API response times, Node.js
Dev Engineering Super Skill
Comprehensive full-stack development and engineering skill merging Perplexity Computer's dev tools with Claude
Doesitwork
Does it work? Quick health check for any feature - validates API endpoints, frontend-backend integration, test
Related Agents
Integration Tester
Specializes in end-to-end integration testing, validating that frontend and backend work together correctly. U
Worker Codex Implementation
Implements a ticket end-to-end inside an isolated worktree by orchestrating codex exec for code-gen. Commander
Fullstack Feature
Orchestrates backend-engineer + frontend-engineer for a single end-to-end feature. Use when the user says "add