Skip to main content

autoloop

Stars0
Forks0
UpdatedJune 9, 2026 at 18:32

把一个 role / prompt / 工作流对着你定义的 benchmark 可量化地越改越好:变异 → 评测 → 多裁判打分 → 赢则保留、输则回滚,在护栏内(变异上限 / 停止条件 / 证据门槛)迭代,可无人值守过夜跑。当想让某个 agent / 角色在能打分的任务上自我改进(而非只重试)时用;前提是先有可机器打分的 eval。

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly