Researchforge banner
forger-labs-hq forger-labs-hq

Researchforge

Development community

Description

A lab protocol for coding agents. Freeze baselines, run hypotheses in isolated worktrees, reject failures and validate improvements from Claude Code, Cursor or a direct API key (Gemini/Anthropic/openAI).

Installation

This entry records only its repository, not the path inside it, so there is no exact command to give. Open the source below and copy the folder into ~/.claude/skills/, or the file into ~/.claude/agents/.

README

ResearchForge

ResearchForge

Your AI generates ideas. ResearchForge tests which ones actually hold up.

Claude Code or Cursor is the researcher. ResearchForge is the lab protocol.


From papers to proof.

PyPI Python 3.12+ Apache 2.0 CI


▶ Watch the intro

Product introduction

ResearchForge turns a research question — or an "improve my repository" goal — into reproducible, traceable evidence. The agent proposes; ResearchForge freezes, executes, measures, rejects, and validates. It finds relevant papers, generates testable hypotheses, benchmarks competing implementations against a frozen baseline in isolated local workspaces, and delivers the strongest supported result as a clean branch, an engineering report, or a research package.

Works with **Claude Code** (slash-command skills), **Cursor** (MDC rules), or **standalone** with any AI API key — install once for your whole machine.

Inspired by Andrej Karpathy's [autoresearch](https://github.com/karpathy/autoresearch), where an agent autonomously runs training experime