#001yuuki-lab1 skills70updated 2026-07-18100% of creatorskilloccupationdescriptionupdatedrl-improvement-loopsoftware-developersRuns the Isaac Lab RL improvement loop for biped_ppo_walk: analyze prior training results (W&B, TensorBoard, logs), propose and implement one-hypothesis code changes, execute smoke/full training, evaluate with eval_biped_walk.py, and record each iteration in markdown until a stated goal is met. Use when improving Isaac Lab biped walking, analyzing W&B/TensorBoard logs, tuning rewards/termination/PPO in yuuki_isaac_lab, or when the user asks to iterate RL until a target (e.g. stable 5 m walk) is achieved.2026-07-18