用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
直接命令不会经过审查 Prompt;运行前请先检查来源。
npx skills add https://github.com/yogsoth-ai/de-anthropocentric-research-engine --skill robustness-design命令会保持在同一行。复制前请横向滚动并检查完整内容。
想先保存到本地?可下载 SkillsMP 当前能够提供的文件。
Strategy: Dialectic engine retuned for truth-seeking, not survival. A defender steelmans a claim into its MOST falsifiable form, a critic attacks to refute it, a judge classifies the exchange into BROKEN/CORROBORATED/UNFALSIFIABLE — the judge does NOT pick a winner or score persuasiveness. Methods: Irving debate (repurposed), Toulmin argumentation, Mayo severe testing.
Campaign: Logical extreme and boundary testing via reductio ad absurdum and edge-case analysis. Core question: Does this artifact collapse under logical limits and boundary conditions? Methods: Lakatos 1976, Dutilh Novaes 2016, BVA, Flyvbjerg Critical Case, Popper.
Campaign for mapping argument structures — extract claims, link evidence, assess strength, synthesize positions. Produces argument graphs in the wiki vault.
正在显示 SKILL.md
基于 SOC 职业分类
| name | robustness-design |
| description | Design experiments to identify failure boundaries and robustness limits |
| version | 1.0.0 |
| category | experiment-execution |
| type | strategy |
| sops | ["factor-identification","level-specification","baseline-selection","metric-specification","sample-size-estimation","design-matrix-construction"] |
| tactics | ["statistical-method-selection"] |
| dependencies | {"sops":["baseline-selection","design-matrix-construction","factor-identification","level-specification","metric-specification","sample-size-estimation"],"tactics":["statistical-method-selection"]} |
Question: Under what conditions does the method fail?
| Robustness Type | Conditions | Severities | Min Runs | Notes |
|---|---|---|---|---|
| Single perturbation | 1 | 3-5 | 3-5 | Quick sanity check |
| Multi-perturbation | 3-5 | 3 each | 9-15 | Standard robustness eval |
| Adversarial sweep | 1 attack | 5-10 epsilon | 5-10 | Adversarial robustness curve |
| Comprehensive | 5+ types | 3-5 each | 50+ | Publication-ready robustness |
| Cross-domain | N domains | 1 | N | Transfer evaluation |
Optional, no fixed order; the final leaf is always a sop.
| Tactic | When to use |
|---|
| statistical-method-selection | Select appropriate statistical methods for experiment analysis |
Optional, no fixed order; the final leaf is always a sop.
| SOP | When to use |
|---|---|
| baseline-selection | Select appropriate baselines for experimental comparison |
| design-matrix-construction | Build the experiment design matrix with proper orthogonality and balance |
| factor-identification | Identify independent, dependent, and control variables for an experiment |
| level-specification | Determine appropriate levels for each experimental factor |
| metric-specification | Define experiment metrics and significance standards |
| sample-size-estimation | SOP: power analysis and required experiment count estimation |