Skip to main content

inference-placement-advisor

Étoiles4
Forks0
Mis à jour18 juillet 2026 à 15:25

Decide where AI inference should run — edge/on-device vs cloud, batch vs realtime, GPU vs CPU — with the embodied-carbon trade-offs named honestly. Use this skill whenever the user asks whether to run models on-device/locally/at the edge or in the cloud, whether to batch inference jobs, or how to serve a model efficiently. Part of Lean Agentic AI Skills; emits lean-findings.json.

Installation

Installer avec Codex ou Claude Copiez ce prompt, collez-le dans Codex, Claude ou un autre assistant, puis laissez-le vérifier la page du skill et l'installer pour vous.

Explorateur de fichiers
2 fichiers
SKILL.md
readonly