Skip to main content
在 Manus 中运行任何 Skill
一键导入

adk-eval-guide

星标2
分支0
更新时间2026年7月8日 13:12

MUST READ before running any ADK evaluation. Evaluation methodology for ADK agents — metrics, evalsets, LLM-as-judge, and critical gotchas. Covers evalset schema, test_config.json format, tool trajectory scoring, and common failure causes. Use when user says "run evaluations", "eval scores are failing", "how do I test my agent", "set up eval cases", "make eval", or when running adk eval or debugging evaluation results. Do NOT use for API code patterns (use adk-cheatsheet), deployment (use adk-deploy-guide), or project scaffolding (use adk-scaffold).

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

文件资源管理器
2 个文件
SKILL.md
readonly