Skip to main content
Manusで任意のスキルを実行
ワンクリックで

onpolicy-algorithms

スター18
フォーク4
更新日2026年7月22日 17:11

Implement, extend, and run on-policy RL algorithms in active-adaptation (PPO, symmetry augmentation, SPO, Muon). Use when adding or modifying algo files under learning/ppo, wiring Hydra algo configs, debugging train_ppo.py runs, GAE/advantage computation, or trust-region policy updates.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

ファイルエクスプローラー
2 ファイル
SKILL.md
readonly