Token Station
Description
Local routing control plane for AI agents and LLM providers, with a loopback-only gateway, smart and quota-aware routing, and desktop + CLI apps.
Installation
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open the source below and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
README
Token Station
**One local gateway for every AI agent.**
Connect Claude Code, Codex, Gemini CLI, Cursor, and other agents to the models you control. Pin a provider, route by task, or use quota before it resets.
[](https://github.com/ballast-ai/token-station/releases/latest) [](https://github.com/ballast-ai/token-station/actions/workflows/ci.yml) [](LICENSE)
[Download](https://github.com/ballast-ai/token-station/releases/latest) · [Quick start](#quick-start) · [Docs](docs/README.md) · [Issues](https://github.com/ballast-ai/token-station/issues) · [简体中文](README.zh-CN.md)
Highlights
- Local by default. The Rust gateway listens on
127.0.0.1:8787and requires authentication. Agent request traffic leaves the device only when you route it to a cloud provider. - Three routes. Direct pins one provider and model. Smart tiers picks High, Mid, or Low in a single decision. Quota first spends buckets that reset sooner.
- Enterprise-managed routing. Enter an enterprise Base URL and credential once. The enterprise service keeps control of its real models and routing policy.
- Your providers. Start from 40+ editable presets, add a custom OpenAI-compatible endpoint, or keep work on a local runtime such as Ollama.
- Desktop and CLI. Both share the same Rust core. Usage, latency, cost estimates, and request logs st
Related Skills
Agency Agents
A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy inject
AI Awesome Llm Apps
100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
AI Firecrawl
🔥 The API to search, scrape, and interact with the web for AI
AI Artifacts Builder
Suite of tools for creating elaborate, multi-component claude.ai HTML artifacts using modern frontend web tech
AI Headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agen
AI CrewAI
Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewA
AI