Skip to main content

inference-placement-advisor

星标4
分支0
更新时间2026年7月18日 15:25

Decide where AI inference should run — edge/on-device vs cloud, batch vs realtime, GPU vs CPU — with the embodied-carbon trade-offs named honestly. Use this skill whenever the user asks whether to run models on-device/locally/at the edge or in the cloud, whether to batch inference jobs, or how to serve a model efficiently. Part of Lean Agentic AI Skills; emits lean-findings.json.

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

文件资源管理器
2 个文件
SKILL.md
readonly