Skip to main content
Manus에서 모든 스킬 실행
원클릭으로

inference-stack-picker

스타0
포크0
업데이트2026년 6월 28일 20:06

Pick an LLM inference stack — provider, model, and hardware — given throughput, latency, cost, sovereignty, and compliance constraints. Compares hyperscaler APIs (OpenAI, Anthropic, Google), dedicated inference clouds (Nebius, Together, Fireworks, Groq, Cerebras), and self-hosted options (vLLM, SGLang, tinygrad). Use when a user asks where to run a model, how to reduce LLM cost, whether to self-host versus use an API, how to plan GPU sovereignty, or to compare inference providers.

설치

Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.

파일 탐색기
2 개 파일
SKILL.md
readonly