ワンクリックで
loop-triage
Read AutoResearch run evidence and loop state, then propose one bounded next experiment or escalate. L1 is report-only.
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
メニュー
Read AutoResearch run evidence and loop state, then propose one bounded next experiment or escalate. L1 is report-only.
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
SOC 職業分類に基づく
Enforce AutoResearch loop cadence, attempt, token, time, sub-agent, remote-compute, and kill-switch limits.
Load and enforce AutoResearch loop constraints before triage, edits, remote work, or external writes.
从 Verl examples 脚本和示例 YAML 中维护“可控参数即特性”的中文情报库与 Excel 台账。Use when Codex needs to scan Verl examples/run scripts, extract all user-adjustable parameters and Hydra overrides, create or update per-parameter markdown explanation files under docs/verl/features/lists, use subagents to research unclear parameters from local explanation files, official Verl docs, and repo docs, and produce docs/verl/features/verl-example-parameters.xlsx with parameter name, category, Chinese explanation, common values, performance impact, accuracy impact, and example count.
Generate RMB task cost reports from Codex/OpenAI-style token usage, including total input tokens, cached input tokens, uncached input tokens, output tokens, and GPT plus DeepSeek API cost estimates. Use when the user asks to统计 token 消耗, API 费用, 人民币成本, cache hit 成本, GPT/DeepSeek 对比, or to produce an `RMB-Cost.md` report for a task, goal, session, or experiment. Use by default for any goal, experiment, remote run, long command, or automation whose expected or actual runtime exceeds 20 minutes.
Verl GRPO formal case adapter for AutoResearch. Use when building, running, or diagnosing autoresearch run verl-case; preparing Qwen/geo3k GRPO matrices; wiring Verl containers to model/data assets, W&B, Prometheus, reports, provenance, and numbered evidence bundles; or explaining val-only versus real GRPO training boundaries.
Generate, validate, inspect, and safely handle AutoResearch customer configuration files. Use when working on config init/show/validate commands, Pydantic config schema behavior, keyring or env secret placeholders, redacted display, or config/config.yaml templates.
| name | loop-triage |
| description | Read AutoResearch run evidence and loop state, then propose one bounded next experiment or escalate. L1 is report-only. |
| user_invocable | true |
Produce a concise, evidence-backed update for the AutoResearch experiment loop.
LOOP.md, loop-constraints.md, loop-budget.md, and STATE.md..planning/STATE.md for project blockers.Return exactly these sections:
High Priority — at most one problem/opportunity and why it matters.Evidence — run id, paths, metrics, failure signature, and confidence.Candidate Hypothesis — one falsifiable hypothesis and smallest validation.Risk / Budget / Authority — what is safe now and what needs approval.Verdict — KEEP, REJECT, or ESCALATE_HUMAN.State Updates — exact proposed edits for STATE.md and an append-only log entry.