Skip to main content
Manusで任意のスキルを実行
ワンクリックで

run-assert-eval

スター197
フォーク27
更新日2026年7月6日 22:50

Run an ASSERT evaluation from a plain-language behavior requirement. Use when the user wants to evaluate, test, or check an AI agent, LLM app, or model against requirements/policies (e.g. "evaluate my agent for budget violations", "test that the support bot never gives legal advice"). Generates or reuses an eval_config.yaml, runs the pipeline, and reports pass/violation rates with trace-cited failure examples.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

SKILL.md
readonly