원클릭으로
orca-benchmark
Compare install, onboarding, spec, or orchestration quality across benchmark cases using inspectable scoring.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
메뉴
Compare install, onboarding, spec, or orchestration quality across benchmark cases using inspectable scoring.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
SOC 직업 분류 기준
Create a host-native `/goal` prompt for bounded ORCA plan work, especially when the user wants an agent to keep working through ready ORCA plans, delegate blocker fixes to subagents, record true blockers, and move on without stalling.
Implement planned changes with scoped edits, verification, and respect for existing project patterns.
Create bounded delegation briefs and structured executor returns for ORCA Framework workflows.
Prepare release readiness checks, release notes, rollback guidance, and handoff.
Record time, retries, and optional token, cache, and cost metrics for ORCA Framework workflows without fabricating precision.
Provide lightweight, skill-sensitive coaching without slowing down normal ORCA Framework execution.
| name | orca-benchmark |
| description | Compare install, onboarding, spec, or orchestration quality across benchmark cases using inspectable scoring. |
A benchmarking workflow for install, onboarding, spec, and orchestration quality across realistic cases.
Use when validating install, onboarding, spec, or orchestration workflow quality, comparing framework versions, or reviewing whether a workflow change helped or hurt.
Do not use as a substitute for product QA or generic issue triage.
templates/benchmark-report.mdThe report should make clear whether workflow quality improved, regressed, or stayed mixed.
Works with orca-install, orca-onboard, orca-spec, orca-eval, and orca-observability.