tachyon-beep

Deep Rl — Development skill for Claude Code

Development community

Deep reinforcement learning - DQN/Rainbow/R2D2/Agent57/BBF, PPO/TRPO/GRPO, SAC/TD3/REDQ/DroQ/CrossQ, DreamerV3/TD-MPC2/MuZero, CQL/IQL/TD3+BC/AWAC/Decision Transformer, MAPPO/IPPO multi-agent, Go-Expl.

How to install Deep Rl

Installs to ~/.claude/commands/tachyon-beep-skillpacks-deep-rl.md

Terminal
mkdir -p ~/.claude/commands && curl -fsSL https://raw.githubusercontent.com/tachyon-beep/skillpacks/HEAD/.claude/commands/deep-rl.md -o ~/.claude/commands/tachyon-beep-skillpacks-deep-rl.md

Restart Claude Code, or start a new session, for it to be picked up.

What Deep Rl does


description: Deep reinforcement learning - DQN/Rainbow/R2D2/Agent57/BBF, PPO/TRPO/GRPO, SAC/TD3/REDQ/DroQ/CrossQ, DreamerV3/TD-MPC2/MuZero, CQL/IQL/TD3+BC/AWAC/Decision Transformer, MAPPO/IPPO multi-agent, Go-Explore/NGU/BYOL-Explore, reward shaping, counterfactual reasoning (HER/OPE), debugging, evaluation. Routes to 13 specialist sheets, 3 commands, 2 SME agents.

Deep RL Routing

**Problem type determines algorithm family. RL is not one algorithm - action space (discrete vs continuo

Alternatives in Development

  • PaperSpine — PaperSpine is a motivation-driven skill for learning from strong academic papers, building a paper’s central a 5k ★
  • Claudeception — A Claude Code skill for autonomous skill extraction and continuous learning 2.2k ★
  • Cq — An open standard for shared agent learning 1.3k ★

Full documentation available on GitHub

View Source Repository