一键导入
loop-triage
Read AutoResearch run evidence and loop state, then propose one bounded next experiment or escalate. L1 is report-only.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Read AutoResearch run evidence and loop state, then propose one bounded next experiment or escalate. L1 is report-only.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
Enforce AutoResearch loop cadence, attempt, token, time, sub-agent, remote-compute, and kill-switch limits.
Load and enforce AutoResearch loop constraints before triage, edits, remote work, or external writes.
从 Verl examples 脚本和示例 YAML 中维护“可控参数即特性”的中文情报库与 Excel 台账。Use when Codex needs to scan Verl examples/run scripts, extract all user-adjustable parameters and Hydra overrides, create or update per-parameter markdown explanation files under docs/verl/features/lists, use subagents to research unclear parameters from local explanation files, official Verl docs, and repo docs, and produce docs/verl/features/verl-example-parameters.xlsx with parameter name, category, Chinese explanation, common values, performance impact, accuracy impact, and example count.
Generate RMB task cost reports from Codex/OpenAI-style token usage, including total input tokens, cached input tokens, uncached input tokens, output tokens, and GPT plus DeepSeek API cost estimates. Use when the user asks to统计 token 消耗, API 费用, 人民币成本, cache hit 成本, GPT/DeepSeek 对比, or to produce an `RMB-Cost.md` report for a task, goal, session, or experiment. Use by default for any goal, experiment, remote run, long command, or automation whose expected or actual runtime exceeds 20 minutes.
Verl GRPO formal case adapter for AutoResearch. Use when building, running, or diagnosing autoresearch run verl-case; preparing Qwen/geo3k GRPO matrices; wiring Verl containers to model/data assets, W&B, Prometheus, reports, provenance, and numbered evidence bundles; or explaining val-only versus real GRPO training boundaries.
Generate, validate, inspect, and safely handle AutoResearch customer configuration files. Use when working on config init/show/validate commands, Pydantic config schema behavior, keyring or env secret placeholders, redacted display, or config/config.yaml templates.
| name | loop-triage |
| description | Read AutoResearch run evidence and loop state, then propose one bounded next experiment or escalate. L1 is report-only. |
| user_invocable | true |
Produce a concise, evidence-backed update for the AutoResearch experiment loop.
LOOP.md, loop-constraints.md, loop-budget.md, and STATE.md..planning/STATE.md for project blockers.Return exactly these sections:
High Priority — at most one problem/opportunity and why it matters.Evidence — run id, paths, metrics, failure signature, and confidence.Candidate Hypothesis — one falsifiable hypothesis and smallest validation.Risk / Budget / Authority — what is safe now and what needs approval.Verdict — KEEP, REJECT, or ESCALATE_HUMAN.State Updates — exact proposed edits for STATE.md and an append-only log entry.