بنقرة واحدة
mm-contest-review
数学建模竞赛 V2.6 高分论文对标评审阶段。用于按高分论文标准审查模型、实验、图表、论文结构、结论追踪和可提交性,并给出修改清单。
التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.
القائمة
数学建模竞赛 V2.6 高分论文对标评审阶段。用于按高分论文标准审查模型、实验、图表、论文结构、结论追踪和可提交性,并给出修改清单。
التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.
استنادا إلى تصنيف SOC المهني
环境检查与安装向导。检查 MathModelAgent V2.6 工作流所需依赖是否已安装,对缺失项提供安装命令,并在用户确认后执行安装。手动触发。
Mathematical modeling contest V2.6 revision loop. Use after mm-contest-review when PAPER_SCORECARD or REVISION_ACTIONS contains BLOCKER, HIGH, MEDIUM, weak-claim, figure-audit, or method-implementation issues that must be repaired before final verification.
数学建模竞赛 V2.6 最终验收阶段。用于检查完整产物、图表插入、结果追踪、论文编译、代码可复现、高分论文评分和提交就绪状态。
数学建模竞赛 V2.6 总控入口。用于启动高分论文对标的 Skill + Codex 子代理混合工作流,创建持久化上下文、调度题面拆解、建模评审、代码实验、图表、论文写作和最终验收。
共享规范知识库。包含 V2.6 流程契约、Codex 子代理协议、评分标准、图表标准、本地 RAG 使用契约、题型路由、模型卡和评审协议等参考内容。其他 skills 在执行过程中按需读取;不要单独触发本 skill。
数学建模竞赛 V2 实验与可视化阶段。用于根据已确认建模路线编写可复现代码,执行 EDA、各子问题、敏感性分析,生成结果表、图表和 RESULTS_MANIFEST.json。
| name | mm-contest-review |
| description | 数学建模竞赛 V2.6 高分论文对标评审阶段。用于按高分论文标准审查模型、实验、图表、论文结构、结论追踪和可提交性,并给出修改清单。 |
| allowed-tools | Bash(*), Read, Write, Edit, Grep, Glob, Agent, WebSearch, WebFetch |
文件关系全貌请见 [[FILE_RELATIONSHIP_MAP]] · 上游: [[skills/mm-paper-build/SKILL|Phase 3 Paper]] · 下游: [[skills/mm-revision-integrator/SKILL|Phase 5 Revise]] · 共享规范: [[skills/_references/SKILL|_references]]
Read:
../_references/contest_score_rubric.md../_references/paper_benchmark_profile.md../_references/agent_review_protocol.md../_references/rag_usage_contract.md../_references/source_quality_policy.md../_references/judge_skim_review_protocol.md../_references/anti_template_review.md../_references/figure_evidence_map.md../_references/evaluator_optimizer_protocol.md../_references/figure_quality_standard.md../_references/nature_figure_integration_guide.md when optional nature-figure scientific plotting integration is available../_references/ars_v2_integration_guide.md when optional ARS editorial synthesis is availableReview all current artifacts, especially:
PROBLEM_BRIEF.mdDATA_AUDIT.mdreports/MODELING_DECISION.mdreports/RESULTS_REPORT.mdreports/FIGURE_PLAN.mdreports/FIGURE_AUDIT.mdreports/TEMPLATE_ADAPTATION_LOG.mdreports/REFINEMENT_LOG.mdreports/METHOD_IMPLEMENTATION_MATRIX.mdreports/CLAIM_TRACE.mdpaper/results/RESULTS_MANIFEST.jsonreports/PAPER_SCORECARD.mdreports/REVISION_ACTIONS.mdreports/FIGURE_AUDIT.mdWORKFLOW_STATE.mdUse independent subagents when possible:
contest-reviewer: scoring and judge perspectivedevils-advocate: weaknesses and likely objectionsvisualization-reviewer: figure/table adequacypaper-writer or final-integrator: structure and prose coherencemodel-reviewer or final-integrator: approved-route versus implemented-method consistencyAfter independent panels complete, optionally load <ARS_ROOT>/academic-paper-reviewer/agents/editorial_synthesizer_agent.md as a role prompt. Use only to synthesize panel findings. Build a synthesis matrix, resolve precedence by severity (BLOCKER > HIGH > MEDIUM > LOW), and append an Editorial Decision Letter plus Revision Roadmap to reports/PAPER_SCORECARD.md. The synthesizer must not fabricate findings; every actionable point must trace to a specific panel report and be represented in reports/REVISION_ACTIONS.md.
If subagents are unavailable, simulate the panels sequentially and record that in reports/AGENT_RUNS.md.
Whether panels are native Codex subagents or simulated roles, log the same metadata: goal, inputs, model/reasoning, permission scope, outputs, conclusion, and thread/id when available.
Before writing the final scorecard:
review_rubrics for scoring/fast-review cues and model_methods when a model-misuse question appears. Record only sourced hits using ../_references/rag_usage_contract.md and ../_references/source_quality_policy.md; B review notes are auxiliary only, and C/D hits can only create caution or review actions.../_references/judge_skim_review_protocol.md. Add the Judge Skim Review section to reports/PAPER_SCORECARD.md.../_references/anti_template_review.md against the final paper claims and implemented methods. Add an Anti-Template Review section to reports/PAPER_SCORECARD.md.reports/FIGURE_PLAN.md, results/RESULTS_MANIFEST.json, and paper figures against ../_references/figure_evidence_map.md. Judge whether each core figure supports the claim, not only whether it is visually clean.reports/TEMPLATE_ADAPTATION_LOG.md exists, verify field mapping, retained/deleted metrics, applicability, and residual template names. If code templates were used but the log is missing, create a HIGH action.BLOCKER, HIGH, or MEDIUM finding from RAG evidence, judge skim, anti-template review, figure evidence, or template adaptation must be represented in reports/REVISION_ACTIONS.md.Score these dimensions from 0 to 5:
When scoring visualization quality and claim evidence, use ../_references/figure_evidence_map.md: a beautiful figure that does not support a claim should score poorly.
For any dimension below 4, create an action item in reports/REVISION_ACTIONS.md and do not mark the review as full PASS.
Conclusion rules:
PASS: every dimension is 4 or 5, no BLOCKER/HIGH actions remain, figure audit has no failed inserted figures, automated audit has no HIGH/BLOCKER issue, and method implementation matrix has no not_implemented core rows.CONDITIONAL_PASS: no fatal correctness issue, but at least one MEDIUM or justified score weakness remains.FAIL: any hard failure, unimplemented approved core method, missing core evidence, or judge-facing defect that would likely cause major score loss.Write reports/REVISION_ACTIONS.md as a table with these fields:
# Revision Actions
| ID | Issue | Severity | Source Panel | Required Fix | Scope | Status |
| --- | --- | --- | --- | --- | --- | --- |
Severity must be BLOCKER, HIGH, MEDIUM, or LOW.
Default status is unresolved. mm-revision-integrator must later update resolution evidence in REVISION_STATUS.md.
Assign severity as follows:
BLOCKER: correctness, traceability, compilation, missing core paper, or impossible-to-submit issueHIGH: likely meaningful score loss, including weak core claims stated strongly, failed inserted figures, missing route diagram in a complex problem, missing required p/error/validation evidence, or approved method not implementedMEDIUM: clarity, justification, appendix coverage, or non-core robustness gapsLOW: polishWrite or update reports/FIGURE_AUDIT.md:
# Figure Audit
| Figure | Inserted | Opens | Readable Text | Labels/Units | Backend Match | Vector Export | Source Data Trace | Stats/Legend | Caption Supports Claim | Status | Required Fix |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
Rules:
FAIL and create a HIGH or BLOCKER action.FAIL or at least WARN depending on centrality; core claim figures should create HIGH actions.HIGH action.nature-figure is enabled, also audit figure contract presence, backend match, editable SVG/PDF text where applicable, source-data traceability, statistics/legend sufficiency, and export bundle completeness. Missing core source data, missing selected-backend script, or missing vector export is at least HIGH unless documented as not applicable.Pillow are at least HIGH when Nature is available. Pillow is acceptable only for non-data diagrams or raster annotations.HIGH.reports/FIGURE_AUDIT.md missing the extended figure-evidence columns is itself a HIGH issue when Nature rules are enabled.For formal contests with four or more subproblems:
HIGH risk unless PAPER_BUILD_REPORT.md documents short-report mode or unusually dense appendix coverage.BLOCKER for full PASS.Review reports/METHOD_IMPLEMENTATION_MATRIX.md against code, results, and paper wording.
not_implemented core row is at least HIGH.HIGH.Review reports/CLAIM_TRACE.md for weak or missing core claims.
missing core claims are hard failures.weak core claims stated as strong conclusions are HIGH.weak core claims stated cautiously may remain MEDIUM or lower depending on contest importance.Mark review as FAIL if: