Topsha — AI skill for Claude Code
Local Topsha 🐧 AI Agent for simple PC tasks - focused on local LLM (GPT-OSS, Qwen, GLM).
How to install Topsha
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open vakovalskii/topsha and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Topsha does
Local Topsha 🐧 AI Agent for simple PC tasks - focused on local LLM (GPT-OSS, Qwen, GLM).
Alternatives in AI
- AI Dev Tasks — A simple task management system for managing AI dev agents 7.8k ★
- Claude Squad — by smtg-ai - Claude Squad is a terminal app that manages multiple Claude Code, Codex (and other local agents i 6.5k ★
- WindsurfAPI — Turn Windsurf / Devin Desktop's 100+ AI models (Claude, GPT, Gemini, DeepSeek, Kimi, GLM, SWE) into OpenAI-, A 3k ★
README
🐧 Topsha Local AI Agent for simple every day tasks
**AI Agent Framework for Self-Hosted LLMs — deploy on your infrastructure, keep data private.**
🎯 **Built for companies and developers who need:**
- 100% on-premise AI agents (no data leaves your network)
- Any OpenAI-compatible LLM (vLLM, Ollama, llama.cpp, text-generation-webui)
- Production-ready security (battle-tested by 1500+ hackers)
- Simple deployment (
docker compose upand you're done)
Why LocalTopSH?
🏠 100% Self-Hosted
Unlike cloud-dependent solutions, LocalTopSH runs entirely on your infrastructure:
| Problem | Cloud Solutions | LocalTopSH |
|---|---|---|
| Data Privacy | Data sent to external APIs | ✅ Everything stays on-premise |
| Compliance | Hard to audit | ✅ Full control, easy audit |
| API Access | Need OpenAI/Anthropic account | ✅ Any OpenAI-compatible endpoint |
| Sanctions/Restrictions | Blocked in some regions | ✅ Works anywhere |
| Cost at Scale | $0.01-0.03 per 1K tokens | ✅ Only electricity costs |
🤖 Supported LLM Backends
| Backend | Example Models | Setup |
|---|---|---|
| vLLM | gpt-oss-120b, Qwen-72B, Llama-3-70B | vllm serve model --api-key dummy |
| Ollama | Llama 3, Mistral, Qwen, 100+ models | ollama serve |
| llama.cpp | Any GGUF model | llama-server -m model.gguf |
| text-generation-webui | Any HuggingFace model | Enable OpenAI API extension |
| LocalAI | Multiple backends | Docker compose included |
| LM Studio | Desktop-friendly | Built-in server mode |
💰 Cost Comparison (1M tokens/day)
| Solution | Daily Cost | Monthly Cost |
|---|---|---|
| OpenAI GPT-4 | ~$30 | ~$900 |
| Anthropic Claude | ~$15 | ~$450 |
| Self-hosted (LocalTopSH) | Electricity only | ~$50-100 (GPU power) |
🌍 Works Everywhere
- ✅ Russia, Belarus, Iran — sanctions don't apply to self-hosted
- ✅ China — no Great Firew
Related Skills
Local LLM Benchmarks
Measured llama.cpp benchmarks on AMD Radeon RDNA4 with ROCm: RX 9070 XT + Radeon AI PRO R9700 (48 GB). Qwen3.8
Claude Code Gemini Vertex
Run Claude Code on Google Gemini (or DeepSeek/Qwen/GPT-OSS) via Vertex AI, billed to a Google Cloud $300 credi
OmniCopilot
1200+ AI models in your GitHub Copilot Chat — free & forever free. VS Code extension powered by OmniRoute: 340
Devo
Model-neutral agent desktop/runtime for private, enterprise, and OpenAI-compatible / Anthropic-compatible mode
Evren Fleet
Claude Code skill: EVREN LLM API (DeepSeek, GLM, Qwen...) as parallel workers and opencode agent lanes
Ag Local Bridge
Local OpenAI, Anthropic & Gemini API bridge for Antigravity — use Claude Sonnet/Opus 4.6, Gemini 3.7/3.6/3.5,
Related Agents
External LLM
When a request mentions external LLM model names (Kimi, K2, Grok, GLM, Gemini, GPT-5)
Delegate
Expert LLM delegation specialist that seamlessly connects to external language models including GPT-4, GPT-3.5
ChatGPT On Wechat
CowAgent是基于大模型的超级AI助理,能主动思考和任务规划、访问操作系统和外部资源、创造和执行Skills、拥有长期记忆并不断成长,比OpenClaw更轻量和便捷。同时支持微信、飞书、钉钉、企微、QQ、公众号、网页