Skip to main content
BBuf
Perfil de criador do GitHub

BBuf

Visão por repositório de 12 skills coletadas em 1 repositórios do GitHub.

skills coletadas
12
repositórios
1
atualizado
23 de ago. de 2026
mapa de repositórios

Onde as skills estão

Principais repositórios por número de skills coletadas, com sua participação neste catálogo do criador e sua distribuição ocupacional.

explorador de repositórios

Repositórios e skills representativas

model-pr-history-knowledge
sem classificação

Use when an SGLang, vLLM, TensorRT-LLM, or TokenSpeed serving/model optimization task needs prior model-family PR evidence. Query and read the PR-driven history docs under model-pr-optimization-history before choosing source paths, fast paths, kernel/fusion…

23 de ago. de 2026
llm-pipeline-analysis
sem classificação

Inspect LLM torch profiler traces at forward-pass, layer, and kernel level. Use when you need layer timings, anchor-kernel boundaries, representative kernel flows, or Perfetto time ranges.

23 de ago. de 2026
llm-serving-auto-benchmark
sem classificação

Framework-independent LLM serving benchmark skill for comparing SGLang, vLLM, TensorRT-LLM, TokenSpeed, or another serving framework. Use when a user wants to find the best deployment command for one model across multiple serving frameworks under the same…

23 de ago. de 2026
llm-serving-capacity-planner
sem classificação

Parse SGLang/vLLM startup logs to explain GPU memory use and request capacity. Use for KV cache budget, mem-fraction-static comparisons, OOM triage, and max-concurrency estimates.

23 de ago. de 2026
llm-torch-profiler-analysis
sem classificação

Unified LLM torch-profiler triage skill for `sglang`, `vllm`, `TensorRT-LLM`, and `TokenSpeed`. Use it to inspect an existing `trace.json(.gz)` or profile directory, or to drive live profiling against a running server when supported and return one three-table…

23 de ago. de 2026
model-architecture-diagram
sem classificação

Return public original model architecture diagrams for user-specified LLM, VLM, MoE, diffusion, OCR, and SGLang/sgl-cookbook model families. Use when the user asks for a model structure chart, architecture diagram, or rendered image link for a specific model…

23 de ago. de 2026
model-compute-simulation
sem classificação

Build an operator-level compute template for an LLM and estimate FLOPs/MFU for a serving shape. Use when you need tensor shapes, per-op FLOPs, kernel-to-op MFU mapping, or parallelism what-if analysis.

23 de ago. de 2026
sglang-model-day0-support
Desenvolvedores de software

Build or audit an evidence-driven SGLang Day-0 support program for a new LLM, VLM, MoE, hybrid-attention, or speculative-decoding model. Use when Codex needs to map a model architecture into SGLang runtime work, design a public support PR DAG, create…

23 de ago. de 2026
Mostrando 8 de 12 skills coletadas.
Mostrando 1 de 1 repositórios
Todos os repositórios foram exibidos