Deep Rl
Description
--- description: Deep reinforcement learning - DQN/Rainbow/R2D2/Agent57/BBF, PPO/TRPO/GRPO, SAC/TD3/REDQ/DroQ/CrossQ, DreamerV3/TD-MPC2/MuZero, CQL/IQL/TD3+BC/AWAC/Decision Transformer, MAPPO/IPPO multi-agent, Go-Explore/NGU/BYOL-Explore, reward shaping, counterfactual reasoning (HER/OPE), debugging, evaluation. Routes to 13 specialist sheets, 3 commands, 2 SME agents. --- # Deep RL Routing **Problem type determines algorithm family. RL is not one algorithm - action space (discrete vs continuo
Installation
Installs to ~/.claude/commands/tachyon-beep-skillpacks-deep-rl.md
mkdir -p ~/.claude/commands && curl -fsSL https://raw.githubusercontent.com/tachyon-beep/skillpacks/HEAD/.claude/commands/deep-rl.md -o ~/.claude/commands/tachyon-beep-skillpacks-deep-rl.md Restart Claude Code, or start a new session, for it to be picked up.
Full documentation available on GitHub
View Source RepositoryRelated Skills
Awesome Go
A curated list of awesome Go frameworks, libraries and software
Development next.js
| The React Framework | 138360 | 1503 | 1 |
Development sharing-skills
skill for guidance.
Development root-cause-tracing
Use when errors occur deep in execution and you need to trace back to find the original trigger.
Development Template Skill
Minimal skeleton for a new skill project structure.
Development Third-party Notices
THE FOLLOWING SETS FORTH ATTRIBUTION NOTICES FOR THIRD PARTY SOFTWARE THAT MAY BE CONTAINED IN PORTIONS OF THI
Development