Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

ab-harness

النجوم١
التفرعات٠
آخر تحديث٢٤ أبريل ٢٠٢٦ في ١٤:٤٥

Run a counterfactual A/B harness on Claude Code to measure whether the user's ~/.claude setup (memories, lessons, axioms, skills, hooks, plugins) actually helps on real tasks vs. a blank-canvas env. Use when: (1) User wants to "prove my setup works" or "quantify setup impact" to colleagues; (2) User asks "is my setup actually helping" beyond what ecosystem-audit's reference-count scan can show; (3) A project has the counterfactual measurement step of an audit → measure → clean pipeline queued. Covers the clean-env mechanism (CLAUDE_CONFIG_DIR), what it does and doesn't isolate, fair-comparison knobs (model pinning, permission mode, stdin), how to mine num_turns/tool_calls/pitfall-hits from the resulting JSONL transcripts, and the honest caveats any n=3 harness report must declare.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
5 ملفات
SKILL.md
readonly