Samzong — DevOps skill for Claude Code
Inference Serving Stack: Scheduling and routing across heterogeneous models — the production path for vLLM and llm-d workloads on Kubernetes.
How to install Samzong
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open samzong/samzong and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Samzong does
- Inference Serving Stack: Scheduling and routing across heterogeneous models — the production path for vLLM and llm-d workloads on Kubernetes.
Alternatives in DevOps
- Skypilot — Run, manage, and scale AI workloads on any AI infrastructure 9.7k ★
- Skill Claw — OpenClaw instance administration — manage hosts across macOS, Ubuntu/Debian, Docker, OCI, and Proxmox 2.8k ★
- Cloudflare Skill — Comprehensive Cloudflare platform reference docs for AI/LLM consumption 709 ★
README
Hi, I'm Samzong (船长) 👋
What I build
- Inference Serving Stack: Scheduling and routing across heterogeneous models — the production path for vLLM and llm-d workloads on Kubernetes.
- Kubernetes-Native Workload Plumbing: Batch queueing, multi-cluster scheduling, and GPU sharing for AI/ML workloads — contributing upstream to Kueue, Karmada, and HAMi.
- Agent Harnesses: The supervision layer for long-running LLM agents — parallel sessions, multi-agent teams, artifacts, approval gates, scheduled dispatch. Because the work shouldn't collapse into chat.
- Agentic Developer Workflows: Multi-worktree dispatch, Claude Code skills, and commit/PR/task automation — the inner loop I live in daily.
Contributions
- Mosoo (Core Contributor): Open-source, Cloudflare-native agent runtime for Codex, Claude Agent SDK, and OpenCode. stars
- OpenClaw : Open-source agent infrastructure for long-running, multi-channel AI work. stars
- semantic-router : Defining the decision-making layer for multi-model LLM serving. stars
- llm-d: Cloud-native infrastructure for disaggregated LLM inference. stars
- HAMi & Kueue: Kubernetes-native batch scheduling and GPU virtualization. HAMi stars Kueue stars
- Istio: Traffic governance for the service mesh layer. stars
- **[Karmada](https://github.com/karmada-io/karm
Related Skills
GPU Server Setup
Agent skill that prepares a Linux server with NVIDIA GPUs for LLM workloads - driver and CUDA, Docker with NVI
Modelarts Vllm Ascend Deploy Skill
New Sample Request Repo name: modelarts-vllm-ascend-deploy-skill Description: An AI coding agent skill (Cursor
Kherep
Orchestration layer for Claude Code and Codex: shared rules, hooks, skills and agents, model and tool routing,
Claude Local Stack
Claude Code stack for Apple Silicon — local MLX + AWS Bedrock routing, token compression, plugin management
Deployment Guide
Production deployment guide for Claude Code Agent Monitor. This document covers every supported deployment pat
Ps Cka
All learner resources for Tim Warner's Pluralsight skill path for the CNCF Certified Kubernetes Administrator
Related Agents
Inference Architect
You are the system Inference Architect — a senior ML infrastructure engineer specializing in local LLM serving
Ndv Flow
Fleet orchestrator. Use when the work is too large for one agent — PRDs, epics, multi-task workloads, anything
Aiml Engineer
AI/ML integration specialist. Designs model pipelines (inference, fine-tuning, LoRA adapters), selects infrast