一键导入
run-benchmark
Runs SpatialVortex benchmarks (flux position accuracy, ELP accuracy, sacred boost verification, geometric reasoning, humanities final exam, performance benchmarks), compares to SOTA/baselines (GPT-4, Claude 3, BERT), generates markdown reports/tables with scores (e.g., 95% accuracy, 111.1% improvement, latency), and suggests optimizations. Automatically invoked for "run benchmarks", "benchmark SpatialVortex", "compare flux accuracy", "eval geometric reasoning", "check ASI progress", "submit results". Uses Rust benchmark suite (cargo run --bin run_benchmarks), saves JSON outputs, and prepares for GitHub PR or PapersWithCode upload.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。