一键导入
heavy-coding-eval
Design and scaffold evaluations comparing single-candidate Composer coding against adaptive Heavy Coder candidate teams.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Design and scaffold evaluations comparing single-candidate Composer coding against adaptive Heavy Coder candidate teams.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
Keep coordinator and leaf context lean-batch reads, cap logs, use execute_code to filter-so swarms stay fast and within token limits.
Use at the start of non-trivial software tasks. Gather repo ground truth with batched reads and searches before planning or delegate_task.
Obey Plan 1A hook session phases so mutating tools, delegate_task sizing, and re-planning do not waste turns or trigger blocks.
Use when spawning delegate_task leaves for coding work. Pack self-contained goals and context so parallel Composer workers ship verifiable patches without coordinator chat history.
Require structured candidate-result JSON from swarm leaves so critique_candidates.py and synthesis run on evidence, not prose self-reports.
Enrich slim hook-injected delegate_tasks with touch map, failure excerpts, and verification commands so parallel leaves produce verifiable patches without coordinator chat history.
| name | heavy-coding-eval |
| description | Design and scaffold evaluations comparing single-candidate Composer coding against adaptive Heavy Coder candidate teams. |
| version | 0.1.0 |
| author | CodeGraphTheory |
| license | MIT |
Use this skill when planning or running Heavy Coder evaluations.
Scaffolded. Benchmark execution is not implemented.
candidate_widths), blind comparative critic, reasoning-model synthesis when available, same verifier.