Governed Pass
Description
A governed protocol for multi-day autonomous AI research: one prompt, five days, and an agent that refused to certify itself. Includes the verbatim specification, a reusable template, and six integrity findings.
Installation
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open the source below and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
README
A Governed Pass: One Prompt, Five Days, and an Agent That Refused to Certify Itself
*From the author:* I built this over months of failed attempts with AI agents. Every rule in the specification below exists because its absence broke something. I loaded the prompt into an agent on a Sunday; it worked alone for four and a half days and ended by correctly refusing to certify its own work. The write-up is an independent audit by a different model — the specification is mine, verbatim. The next pass gets published whether it works or not.
**A working protocol for multi-day autonomous AI research, with the exact specification that produced it and six empirical integrity findings.**
**Status:** Pass 2 (R2) is complete. It terminated after **4 days 8 hours 53 minutes** of continuous autonomous work with the disposition **`METHOD_NEEDS_REPAIR`** — the specification's designated honest-failure outcome — after its own adversarial review round blocked method freeze with 3 HIGH and 4 MEDIUM findings. No scientific result was run or claimed. Pass 3 is planned; its disposition and runtime will be appended to [UPDATES.md](UPDATES.md) regardless of direction.
**Evidence repository:** currently private. Commit hashes, artifact SHA-256s, and file paths are cited throughout so the mechanics are concrete now and independently checkable later: the evidence repository will be opened when the experiment's method freeze binds pinned model identities, closing the contamination window that early publication would create for the sealed scientific run. Publishing the hashes before the repository opens is deliberate — when it opens, readers can verify nothing was retrofitted in the interim.
Summary of claims
This document makes five claims, each bound to the evidence chain described below, and deliberately makes no others.
**Claim 1 — the run.** A single specification prompt (reproduced verbatim in [ORIGINAL_PROMPT.verbatim.md](ORIGINAL_PROMPT.verbatim.md)), loaded once into a
Related Skills
Agency Agents
A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy inject
AI Awesome Llm Apps
100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
AI Firecrawl
🔥 The API to search, scrape, and interact with the web for AI
AI Artifacts Builder
Suite of tools for creating elaborate, multi-component claude.ai HTML artifacts using modern frontend web tech
AI Headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agen
AI CrewAI
Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewA
AI