ソース情報
- リポジトリ
- shenli/devtool-ax-kit
- ソースの最終更新活動
- 2026年8月27日 22:03
- 検出された SKILL.md の言語
- 英語
- スター
- 2
- フォーク
- 0
インストール方法
デフォルトでは、最初にソースを確認する Prompt が選択されています。直接コマンドに切り替えるか、ローカルコピーをダウンロードすることもできます。
ソースファイルを確認
インストールを決める前に、SKILL.md と SkillsMP に表示されている付属ファイルをお読みください。
メニュー
デフォルトでは、最初にソースを確認する Prompt が選択されています。直接コマンドに切り替えるか、ローカルコピーをダウンロードすることもできます。
インストールを決める前に、SKILL.md と SkillsMP に表示されている付属ファイルをお読みください。
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
直接コマンドでは確認用 Prompt が省略されます。実行前にソースを確認してください。
npx skills add https://github.com/shenli/devtool-ax-kit --skill ax-comparisonコマンドは1行のまま表示されます。コピー前に横へスクロールして全体を確認してください。
ローカルで確認しますか?SkillsMP が現在取得できるファイルをダウンロードできます。
Capture agent runs with checkpoints, effect receipts, and redacted trajectories.
Build independent, deterministic verifiers for agent-tool tasks.
Design controlled coding-agent tasks for measuring developer-tool agent experience.
SKILL.md を表示中
| name | ax-comparison |
| description | Compare raw and official agent-tool surfaces using replicated AX tasks. |
Compare one controlled task under clearly named conditions, such as public docs/CLI versus an official Skill, MCP server, or plugin. Hold the fixture, verifier, resource limits, and prompt intent constant. Randomize or repeat runs when practical, and record timing, retries, duplicate effects, recovery, and transcript safety.
Treat differences as agent-usability evidence unless the protocol isolates a vendor defect. Report non-reproductions and alternative explanations. Do not generalize from a single successful run.
Attribute each failure to one layer where possible: task/model, harness/session, tool interface, or external infrastructure. Track human attention and verifier disagreement alongside completion and wall-clock time.