fstandhartinger

Model Market Comparison — DevOps skill for Claude Code

DevOps community

Compare open-source & frontier LLM prices across providers (OpenRouter, AWS Bedrock, Azure Foundry, GitHub Copilot, Claude Code) with ArtificialAnalysis & DesignArena benchmarks.

How to install Model Market Comparison

This entry records only its repository, not the path inside it, so there is no exact command to give. Open fstandhartinger/model-market-comparison and copy the folder into ~/.claude/skills/, or the file into ~/.claude/agents/.

What Model Market Comparison does

Compare open-source & frontier LLM prices across providers (OpenRouter, AWS Bedrock, Azure Foundry, GitHub Copilot, Claude Code) with ArtificialAnalysis & DesignArena benchmarks.

Alternatives in DevOps

  • Claude Code Router — Use Claude Code as the foundation for coding infrastructure, allowing you to decide how to interact with the m 30.1k ★
  • Claudable — Claudable is an open-source web builder that leverages local CLI agents, such as Claude Code, Codex, Gemini CL 4k ★
  • Proxy — route Claude Code requests through multiple upstream providers (OpenCode Go, OpenCode Zen, and AWS Bedrock) wi 957 ★

README

![Benchmark Heaven](public/brand/wordmark.svg)

Every benchmark result for every model — and what each one actually costs.

Benchmark Heaven (formerly Model Market Comparison) has its primary base URL at **[benchmarkheaven.com](https://benchmarkheaven.com)** (2026-09-11). The previous host **[model-market-comparison.app.mintapis.com](https://model-market-comparison.app.mintapis.com) remains valid** and serves the same app and API endpoints. The GitHub repo, file paths, data schema, IDs, units, Composite and settings keys did not move. Browser preferences are stored per hostname and are not automatically transferred to the new domain.

The [daily refresh](ops/daily/README.md) stages source changes and requires a cheap worker plus a different-family critic before publication. Each observation keeps its source and date; unavailable or contested candidates do not erase accepted data.

Compare **open-source and frontier LLMs** by capability and price in one place. Capability comes from [ArtificialAnalysis](https://artificialanalysis.ai) (Coding, Coding Agent & Intelligence indices) and [Intelligence.ai / DesignArena](https://intelligence.ai) (Agentic Web Dev Frontend & Full-Stack Elo). Prices are aggregated across **OpenRouter** inference providers, **AWS Bedrock**, **Azure AI Foundry**, **Google Vertex AI**, **Nebius**, **Inceptron**, **TensorX**, **Scaleway**, **IONOS**, **Mistral**, **Chutes**, **OVHcloud**, **STACKIT**, **T-Systems LLM Hub**, **TrustedTokens**, **GitHub Copilot**, and the **Anthropic / Claude Code** list price — normalized to USD per 1M tokens.

📋 Product spec: [PRD.md](PRD.md) · 🗂️ Data schema: [data/SCHEMA.md](data/SCHEMA.md) · 🔄 Data collection & refresh: [data/SCRAPING.md](data/SCRAPING.md) · 🔌 Public API: [API.md](API.md) · 🚀 Deploy / migrate / env vars: [DEPLOYMENT.md](DEPLOYMENT.md) · 📜 Data/URL changes for consumers: [CHANGELOG.md](CHANGELOG.md)

Benchmark exploration

Explore [benchmark rankings](https://benchmark