Skip to main content
Run any Skill in Manus
with one click

inference-stack-picker

Stars0
Forks0
UpdatedJune 28, 2026 at 20:06

Pick an LLM inference stack — provider, model, and hardware — given throughput, latency, cost, sovereignty, and compliance constraints. Compares hyperscaler APIs (OpenAI, Anthropic, Google), dedicated inference clouds (Nebius, Together, Fireworks, Groq, Cerebras), and self-hosted options (vLLM, SGLang, tinygrad). Use when a user asks where to run a model, how to reduce LLM cost, whether to self-host versus use an API, how to plan GPU sovereignty, or to compare inference providers.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

File Explorer
2 files
SKILL.md
readonly