Model Market Comparison — DevOps skill for Claude Code
Compare open-source & frontier LLM prices across providers (OpenRouter, AWS Bedrock, Azure Foundry, GitHub Copilot, Claude Code) with ArtificialAnalysis & DesignArena benchmarks.
How to install Model Market Comparison
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open fstandhartinger/model-market-comparison and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Model Market Comparison does
Compare open-source & frontier LLM prices across providers (OpenRouter, AWS Bedrock, Azure Foundry, GitHub Copilot, Claude Code) with ArtificialAnalysis & DesignArena benchmarks.
Alternatives in DevOps
- Claude Code Router — Use Claude Code as the foundation for coding infrastructure, allowing you to decide how to interact with the m 30.1k ★
- Claudable — Claudable is an open-source web builder that leverages local CLI agents, such as Claude Code, Codex, Gemini CL 4k ★
- Proxy — route Claude Code requests through multiple upstream providers (OpenCode Go, OpenCode Zen, and AWS Bedrock) wi 957 ★
README

Every benchmark result for every model — and what each one actually costs.
Benchmark Heaven (formerly Model Market Comparison) has its primary base URL at **[benchmarkheaven.com](https://benchmarkheaven.com)** (2026-09-11). The previous host **[model-market-comparison.app.mintapis.com](https://model-market-comparison.app.mintapis.com) remains valid** and serves the same app and API endpoints. The GitHub repo, file paths, data schema, IDs, units, Composite and settings keys did not move. Browser preferences are stored per hostname and are not automatically transferred to the new domain.
The [daily refresh](ops/daily/README.md) stages source changes and requires a cheap worker plus a different-family critic before publication. Each observation keeps its source and date; unavailable or contested candidates do not erase accepted data.
Compare **open-source and frontier LLMs** by capability and price in one place. Capability comes from [ArtificialAnalysis](https://artificialanalysis.ai) (Coding, Coding Agent & Intelligence indices) and [Intelligence.ai / DesignArena](https://intelligence.ai) (Agentic Web Dev Frontend & Full-Stack Elo). Prices are aggregated across **OpenRouter** inference providers, **AWS Bedrock**, **Azure AI Foundry**, **Google Vertex AI**, **Nebius**, **Inceptron**, **TensorX**, **Scaleway**, **IONOS**, **Mistral**, **Chutes**, **OVHcloud**, **STACKIT**, **T-Systems LLM Hub**, **TrustedTokens**, **GitHub Copilot**, and the **Anthropic / Claude Code** list price — normalized to USD per 1M tokens.
📋 Product spec: [PRD.md](PRD.md) · 🗂️ Data schema: [data/SCHEMA.md](data/SCHEMA.md) · 🔄 Data collection & refresh: [data/SCRAPING.md](data/SCRAPING.md) · 🔌 Public API: [API.md](API.md) · 🚀 Deploy / migrate / env vars: [DEPLOYMENT.md](DEPLOYMENT.md) · 📜 Data/URL changes for consumers: [CHANGELOG.md](CHANGELOG.md)
Benchmark exploration
Explore [benchmark rankings](https://benchmark
Related Skills
Claude Overlay
Manage project-level Claude Code configuration for custom model providers (Databricks, Bedrock, OpenRouter, Li
AI Contributing
Open-source multi-agent governance for running multiple AI coding agents in parallel on one GitHub repo: issue
Swobu
Provider router and compatibility proxy for Claude Code, Codex and AI coding agents. Switch and fail over acro
Clawsync
ClawSync, OpenClaw for the cloud. Deploy an open source personal AI agent with chat UI, skills system, MCP sup
Env Compare
Compare Terraform configurations across environments (dev, staging, prod)
LibreChat
Enhanced ChatGPT Clone: Features OpenAI, Assistants API, Azure, Groq, GPT-4 Vision, Mistral, Bing, Anthropic,
Related Agents
AWS AI Terraform
Write, extend, and review the Terraform in this repo for AWS AI/ML services on the AI Practitioner (AIF-C01) e
Family Researcher
General local-research agent for one assigned angle — activities, providers, childcare options, restaurants, c
Evolve Behavior Compare
Behavior comparison agent for the Evolve Loop (Evaluate archetype). The advisor INSERTS this phase on refactor