Llmix — AI skill for Claude Code
Production LLM call layer for AI agents and tools: keep OpenAI/Anthropic/AI SDK/LiteLLM, hot-swap models with MDA presets, and add cache, retries, circuit breakers, key rotation, singleflight, and Pyt.
How to install Llmix
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open sno-ai/llmix and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Llmix does
Production LLM call layer for AI agents and tools: keep OpenAI/Anthropic/AI SDK/LiteLLM, hot-swap models with MDA presets, and add cache, retries, circuit breakers, key rotation, singleflight, and Python/TypeScript/Rust parity.
Alternatives in AI
- Ralph For Claude Code — by Frank Bria - An autonomous AI development framework that enables Claude Code to work iteratively on project 8k ★
- Cc Mirror — Create multiple isolated Claude Code variants with custom providers (Z.ai, MiniMax, OpenRouter, LiteLLM) 2.1k ★
- Meme Rush — Track real-time meme token lists from launchpads (Pump.fun, Four.meme) and AI-powered hot market topics ranked 483 ★
README
LLMix
[](https://www.npmjs.com/package/@snoai/llmix) [](https://pypi.org/project/sno-llmix/) [](https://crates.io/crates/llmix-rs) [](https://www.python.org/downloads/) [](https://www.typescriptlang.org/) [](https://www.rust-lang.org/) [](LICENSE)
Read in other languages: **English** · [中文](docs/llmix/readme/README.zh-CN.md) · [Deutsch](docs/llmix/readme/README.de.md) · [Español](docs/llmix/readme/README.es.md) · [Français](docs/llmix/readme/README.fr.md) · [Русский](docs/llmix/readme/README.ru.md) · [한국어](docs/llmix/readme/README.ko.md) · [日本語](docs/llmix/readme/README.ja.md) · [हिन्दी](docs/llmix/readme/README.hi.md)
Config-driven LLM calls for Python, TypeScript, and Rust. Keep your SDK. Move model behavior into MDA presets. Put cache, retries, key rotation, and rollout control around the call.
LLMix is the layer between your product and the provider SDK.
It does not ask you to rewrite your OpenAI, Anthropic, Gemini, LiteLLM, AI SDK, or custom client code. It wraps the call. The boring parts go around it: response cache, circuit breaker, key pools, singleflight, retry policy, adaptive concurrency, provider kwargs, and MDA config loading.
The model stops being a hard-coded string buried in application code. It becomes data. Change a preset, publish a compiled registry release, reload the servi
Related Skills
Canopus
Canopus keeps long-running AI coding agents locked on their goal: frozen acceptance envelopes, deterministic c
Cmc Proxy
Reverse-proxy a GOAT/commandcode AI subscription to localhost for Claude Code & Codex — OpenAI/Anthropic/Respo
Vibestrate
Open-source supervised flow for AI coding. Run Claude Code, Codex, Gemini, Aider or local models as one crew,
Ghost Showcase
Invisible AI assistant for Windows — real-time STT + multi-provider LLM (Claude/GPT/Ollama/LM Studio) + RAG, s
Gradient Starter Kit
A starter kit for building with DigitalOcean Gradient AI — includes serverless inference examples, OpenAI SDK
Claude Code Trace Visualizer
Trace and visualize what Claude Code does during a run: tool calls, timing, file access, retries/failures, and
Related Agents
Lens Backend
Backend architecture lens of the production readiness audit. Judges what breaks at 10x traffic and what breaks
AI ML
AI/ML 통합 전문가 + LLM API 최신 모델/SDK 코딩 가이드. RAG 시스템, 문서 분석, OpenAI/Anthropic/Gemini/Ollama 최신 API 보장. "AI integra
AIOps Manager
Owns AIOps — anomaly detection and correlation over observability telemetry, alert-noise reduction (grouping,