Skip to main content
Manusで任意のスキルを実行
ワンクリックで

ai-llm-inference

スター67
フォーク15
更新日2026年3月12日 17:35

LLM inference patterns — latency budgeting, caching, batching, quantization, and parallelism. Use when optimizing serving cost or tail latency.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

ファイルエクスプローラー
30 ファイル
SKILL.md
readonly