A set of reliable implementations of reinforcement learning algorithms in PyTorch, including PPO, SAC, TD3, DQN, and more.
May 11, 2026·1 min read
Stable Baselines3 — Reliable Reinforcement Learning in PyTorch
A set of reliable implementations of reinforcement learning algorithms in PyTorch, including PPO, SAC, TD3, DQN, and more.
Agent ready
Ready-to-run agent install
This asset can be installed after the agent chooses its runtime, checks the plan, and runs the matching command.
Native · 98/100Policy: allow
Agent surface
Any MCP/CLI agent
Kind
Skill
Install
Single
Trust
Trust: Community
Entrypoint
Stable Baselines3 Overview
Direct install command
npx -y tokrepo@latest install 8dee1283-4cd0-11f1-9bc6-00163e2b0d79 --target codexRun after dry-run confirms the install plan.
Discussion
Sign in to join the discussion.
No comments yet. Be the first to share your thoughts.