Skip to main content

rl-training-iteration

Stars16
Forks6
UpdatedJuly 18, 2026 at 02:52

Use whenever the user wants autonomous or semi-autonomous iterative RL/policy training โ€” launching training runs, watching metrics, diagnosing plateaus or divergence, tweaking reward/env/hyperparameter config, and repeating until a target behavior emerges. Applies to robot locomotion, manipulation, game agents, or any Isaac Lab / Gym / RSL-RL / rl_games style training loop where the user says things like "keep iterating," "train until it works," "watch this run and adjust," or asks for autonomous tuning of a training pipeline. Not just for humanoid walking โ€” the loop generalizes to any sparse-reward, exploration-hard RL task.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

File Explorer
2 files
SKILL.md
readonly