AI Gateway Testkit — Testing skill for Claude Code
Open-source conformance and regression testkit for Anthropic- and OpenAI-compatible AI gateways, with stable test IDs, compatibility profiles, SDK checks, agent workflows, and shareable reports.
How to install AI Gateway Testkit
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open trungdlp/ai-gateway-testkit and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What AI Gateway Testkit does
Open-source conformance and regression testkit for Anthropic- and OpenAI-compatible AI gateways, with stable test IDs, compatibility profiles, SDK checks, agent workflows, and shareable reports.
Alternatives in Testing
- Ext Apps — Official repo for spec & SDK of MCP Apps protocol - standard for UIs embedded AI chatbots, served by MCP serve 1.9k ★
- Robotics Agent Skills — Agent skills that make AI coding assistants write production-grade robotics software 170 ★
- Opencode Skill Creator — OpenCode skill for creating, testing, and optimizing other OpenCode skills 157 ★
README
AI Gateway Testkit
[](https://github.com/trungdlp/ai-gateway-testkit/actions/workflows/ci.yml) [](go.mod) [](LICENSE)
**Know what your AI gateway actually supports before your users find out.**
AI Gateway Testkit (`agtk`) is an open-source, black-box conformance and regression runner for Anthropic- and OpenAI-compatible gateways. It turns compatibility claims into repeatable test profiles, stable assertion IDs, CI-friendly verdicts, and canonical reports you can compare or safely share.

gateway target + compatibility profile
|
v
versioned test catalog
|
v
protocol / SDK / agent checks
|
v
PASS / FAIL / INDETERMINATE
|
v
comparable canonical report
Why `agtk`?
Claims of OpenAI or Anthropic compatibility can mean anything from basic authentication to full tool-calling and SDK interoperability. A successful request alone does not prove that a gateway is ready for production workloads.
`agtk` gives teams evidence they can act on:
- Precise failures: permanent scenario and assertion IDs such as
OAI-TOOL-001/A04, instead of fragile log messages. - Explicit compatibility claims: versioned profiles for core APIs, tool calling, official Go SDKs, behavioral diagnostics, Codex, and Claude Code.
- Honest verdicts:
FAILmeans observed incompatibility; missing or unavailable evidence becomesINDETERMINATE, never a false failure or pass. - **Regression-r
Related Skills
Turnpike
A minimal local gateway that proxies Claude/Anthropic and OpenAI-spec clients to any OpenAI- or Anthropic-comp
FreeComputerUse
Local-first AI browser automation with Playwright. Use OpenAI, Grok, DeepSeek, Anthropic and compatible APIs—o
Spec To Tickets
Break a capability spec (or the shaping conversation that produced it, or a parent issue) into tracer-bullet t
Clawsouls
ClawSouls — an open-spec platform for shareable AI agent personas (souls). 80+ curated souls, one-command inst
Cfn Share
Publish a plan, spec, or any project markdown doc as a private shareable web page so a non-terminal colleague
Gemma 4 31B MTP MLX
Local MLX gateway for running Gemma 4 31B with the Gemma 4 MTP assistant drafter. This is a text-only local ga
Related Agents
Provider Debugger
Diagnose live-provider compatibility failures for Kimi, GLM, and other OpenAI-compatible endpoints
AI ML
AI/ML 통합 전문가 + LLM API 최신 모델/SDK 코딩 가이드. RAG 시스템, 문서 분석, OpenAI/Anthropic/Gemini/Ollama 최신 API 보장. "AI integra
Fable5 Apfel Engineer
Heavy-lifting Fable-5 engineer for the apfel project (Apple on-device FoundationModels CLI + OpenAI-compatible