Arbiter — AI skill for Claude Code
A local proxy that lets Claude Code use NVIDIA NIM models instead of Anthropic's API.
How to install Arbiter
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open Ujwal397/Arbiter and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Arbiter does
A local proxy that lets Claude Code use NVIDIA NIM models instead of Anthropic's API. Features task-aware model routing, three-tier fallback, and an interactive model selection menu.
Alternatives in AI
- Model Selection — Model Selection & Task Tool 1.1k ★
- Memory Forge Rs — Stop resetting satisfying AI chats — edit the memory instead 474 ★
- Askimo — AI Client for chat, RAG, Skills, MCP tools, and agents 470 ★
README
⚖️ Arbiter
Run Claude Code on NVIDIA's free cloud models — zero cost, no hardware needed.
[](LICENSE) [](https://python.org) [](https://fastapi.tiangolo.com) [](https://github.com/Ujwal397/Arbiter)
Claude Code ──▶ Arbiter (localhost:4005) ──▶ NVIDIA NIM Cloud ──▶ Free Models
Claude Code only works with Anthropic's paid API by default. Arbiter breaks that lock — it sits between Claude Code and NVIDIA's free cloud API, transparently routing every request to models like **Kimi K2 0905**, **Mistral Large 3**, and **Llama 3.3 70B** at **zero cost**.
[**Get started →**](#-quick-start) · [Model availability](#-known-model-availability) · [Configuration](#️-configuration)
✦ Why Arbiter?
| Without Arbiter | With Arbiter |
|---|---|
| Claude Code only works with Anthropic's paid API | Routes through NVIDIA's free cloud models |
| You pay per token — costs stack up fast | Completely free with an NVIDIA API key |
| Limited to models Anthropic offers | Access models like Kimi K2, Mistral Large 3, Llama 3.3 |
| Large models require enterprise hardware locally | Runs in the cloud — nothing to install beyond Python |
**NVIDIA NIM is free.** Get an API key at [build.nvidia.com](https://build.nvidia.com) — no credit card, no per-token charges.
✦ Features
🧭 Task-aware model routingArbiter reads your last two messages and classifies the task before dispatching. Heavy work goes to the most capable model; quick questions go to the fastest one — automatically.
| Task type | Triggers on | Model tier | |:---|:---
Related Skills
OmniPilot
Cross-platform LLM computer-use agent — OpenAI, Anthropic, Gemini, NVIDIA NIM and more control your PC behind
First Pass
Stage 1 first pass - best-source selection, ONE AI upscale (IllustrationJaNai via .venv-upscale, else realesrg
Agent Observability Stack
Self-hostable observability for LLM agents + a Linux host (Intel iGPU/NPU + NVIDIA RTX 3090 eGPU aware): Prome
Ast Outline
ast-outline lets AI coding agents pull exactly the code context they need — a repo skeleton, a file outline, o
MindBank
MindBank gives your AI Hermes or claude code a permanent, searchable, relationship-aware memory that persists
Presentation Kit
Build interactive, personalized web presentations and proposals with your AI agent — a live web deck instead o
Related Agents
Fleet Explorer
Topology-aware code explorer for a fleet of multi-repo Spring Boot microservices. Traces a request or feature
Glm Agent
Runs Z.AI GLM on the user's GLM Coding Plan. Three modes - consult (direct API call, 1M context, for inputs to
Omd Codex Image
Channel-aware image materializer. Reads spec blocks in HTML/MD/JSX and materializes them through Codex's nativ