Compare two eval runs and report what changed. Reads both runs' events, transcripts, and produced artifacts. Writes a short markdown summary classifying differences as regression, improvement, or neutral.
原文の言語: 英語
メニュー
SkillsMP は adam-s/agent-spec から 8 件の skill を収集しています。skill を開くとソースと詳細を確認できます。
収集済み skill 8 件中 8 件を表示しています。
Compare two eval runs and report what changed. Reads both runs' events, transcripts, and produced artifacts. Writes a short markdown summary classifying differences as regression, improvement, or neutral.
原文の言語: 英語
Generalized recursive iteration loop. Runs parallel sub-agents against a target, scores deterministically, diagnoses instruction gaps, applies fixes, and recurses until the stop condition is met or max depth is reached.
原文の言語: 英語
Run an evaluation against an eval with a specific config
原文の言語: 英語
Write a handoff document so a new chat can continue the work
原文の言語: 英語
Show evaluation results and comparisons
原文の言語: 英語
Test-driven development of a Hono/Bun WebSocket application. Read requirements, read tests, build server, verify, iterate until all tests pass.
原文の言語: 英語
Scaffold a new evaluation
原文の言語: 英語
Stop all agent-spec processes, clear ports, remove sandboxes, verify clean state.
原文の言語: 英語