用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
直接命令不会经过审查 Prompt;运行前请先检查来源。
npx skills add https://github.com/JackSmith111977/Hermes-Skill-View --skill serving-llms-vllm命令会保持在同一行。复制前请横向滚动并检查完整内容。
想先保存到本地?可下载 SkillsMP 当前能够提供的文件。
正在显示 SKILL.md
基于 SOC 职业分类
| name | serving-llms-vllm |
| description | Serves LLMs with high throughput using vLLM's PagedAttention and continuous... |
| version | 1.0.0 |
| triggers | ["serving llms vllm","serving-llms-vllm","vllm","模型部署","推理加速","LLM推理"] |
| author | Orchestra Research |
| license | MIT |
| dependencies | ["vllm","torch","transformers"] |
| metadata | {"hermes":{"tags":["vLLM","Inference Serving","PagedAttention","Continuous Batching","High Throughput","Production","OpenAI API","Quantization","Tensor Parallelism"]}} |
| category | mlops |
Serves LLMs with high throughput using vLLM's PagedAttention and continuous...