一键导入
evidence-synthesis
Synthesize multiple issues, configs, seeds, metrics, and artifacts into conservative mechanism-level conclusions with caveats.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Synthesize multiple issues, configs, seeds, metrics, and artifacts into conservative mechanism-level conclusions with caveats.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
Open a conservative Robot SF PR with scope verification, freshness checks, and artifact discipline.
Use for an autonomous Robot SF issue-to-PR loop that selects eligible GitHub issues, implements one scoped issue at a time, validates, pushes, and opens PRs.
Guarded PR merger; merges merge-ready PRs after verifying label, CI status, branch protection, and preflight checks.
Use for an autonomous Robot SF PR review loop that fixes scoped review gaps, validates proof, resolves review threads, and applies merge-ready; not for merging.
Autonomous issue-to-PR workflow from next eligible issue to ready PR with consistent metadata handling.
Continuous goal autopilot; orchestrates implement, review, merge, and discover cycles with preflight validation and delegation failure recovery.
| name | evidence-synthesis |
| description | Synthesize multiple issues, configs, seeds, metrics, and artifacts into conservative mechanism-level conclusions with caveats. |
| category | benchmark-evidence |
| kind | analysis |
| phase | analysis |
| requires_write | true |
| requires_slurm | false |
| requires_benchmark_artifacts | true |
| delegates_to | ["artifact-provenance","benchmark-row-status","paper-facing-docs"] |
| output_schema | evidence_synthesis_summary.v1 |
Use this skill when multiple issues, campaigns, configs, seeds, metrics, and artifacts need one conservative conclusion.
benchmark-row-status when benchmark data is involved.artifact-provenance.diagnostic-only, smoke evidence,
nominal benchmark evidence, or paper-grade.For ordinary exploratory runs, prefer per-experiment hypothesis notes over a central ledger. The minimum reusable fields are hypothesis, variant/config, expected signal, result classification, artifact pointer or snapshot, and next decision. Create a central hypothesis ledger only when a research family has enough related runs that the synthesis question becomes "what do we believe now?" instead of "what should we run next?".
| Mechanism | Source issue | Evidence tier | Config | Seeds | Artifacts | Metrics | Verdict | Caveats |
|---|
paper-facing-docs before manuscript-support language is published.Use evidence_synthesis_summary.v1.