Skip to main content
Manusで任意のスキルを実行
ワンクリックで

evalsmith

スター1
フォーク0
更新日2026年3月28日 04:39

Grounded forensic evaluation engineering for LLM product repositories. Use when a coding agent needs to inspect an existing codebase, infer the workflow behind a prompt feature, structured output flow, RAG pipeline, tool-calling path, or simple multi-stage chain, then design repo-native evals, capture exact traces, attribute behaviors to prompt or context components, and propose evidence-backed prompt or workflow improvements.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

SKILL.md
readonly