一键导入
orca-benchmark
Compare install, onboarding, spec, or orchestration quality across benchmark cases using inspectable scoring.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Compare install, onboarding, spec, or orchestration quality across benchmark cases using inspectable scoring.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
Create a host-native `/goal` prompt for bounded ORCA plan work, especially when the user wants an agent to keep working through ready ORCA plans, delegate blocker fixes to subagents, record true blockers, and move on without stalling.
Implement planned changes with scoped edits, verification, and respect for existing project patterns.
Create bounded delegation briefs and structured executor returns for ORCA Framework workflows.
Prepare release readiness checks, release notes, rollback guidance, and handoff.
Record time, retries, and optional token, cache, and cost metrics for ORCA Framework workflows without fabricating precision.
Provide lightweight, skill-sensitive coaching without slowing down normal ORCA Framework execution.
| name | orca-benchmark |
| description | Compare install, onboarding, spec, or orchestration quality across benchmark cases using inspectable scoring. |
A benchmarking workflow for install, onboarding, spec, and orchestration quality across realistic cases.
Use when validating install, onboarding, spec, or orchestration workflow quality, comparing framework versions, or reviewing whether a workflow change helped or hurt.
Do not use as a substitute for product QA or generic issue triage.
templates/benchmark-report.mdThe report should make clear whether workflow quality improved, regressed, or stayed mixed.
Works with orca-install, orca-onboard, orca-spec, orca-eval, and orca-observability.