Skip to main content

vllm-project/vllm-omni

SkillsMP는 vllm-project/vllm-omni에서 9개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
9
GitHub 스타
6,569
GitHub 포크
1,601

이 저장소의 skills

직업 카테고리 1개 · 44% 분류됨

수집된 skill 9개 중 9개를 표시합니다.

직업 분류
미분류
설명

Add a new diffusion model (text-to-image, text-to-video, image-to-video, text-to-audio, image editing) to vLLM-Omni, including Cache-DiT acceleration and parallelism support (TP, SP/USP, CFG-Parallel, HSDP). Use when integrating a new diffusion model, porting…

원문 언어: 영어

업데이트
직업 분류
미분류
설명

Integrate a new text-to-speech model into vLLM-Omni from HuggingFace reference implementation through production-ready serving with streaming and CUDA graph acceleration. Use when adding a new TTS model, wiring stage separation for speech synthesis, enabling…

원문 언어: 영어

업데이트
직업 분류
미분류
설명

Generate and run tests for vllm-project/vllm-omni with CI-aligned levels and markers; wire new tests into Buildkite (test-ready.yml for L1/L2, test-merge.yml for L3, test-nightly.yml for L4). On completion, always provide copy-paste local and CI-like pytest…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Find evidence-backed simplification candidates in vLLM-Omni and, when requested, turn them into focused proposals or code changes. Use for audits of dead, duplicated, speculative, over-generalized, unnecessarily defensive, or hand-rolled code. Use review-pr…

원문 언어: 영어

업데이트
직업 분류
미분류
설명

Self-check your branch before creating a PR — catch dead code, prevent new model-specific Python examples, verify accuracy/perf claims, validate PR title format, and confirm merge readiness. Use when the user says "precheck", "self review", "pre-submit…

원문 언어: 영어

업데이트
직업 분류
미분류
설명

Review pull requests and local branches for vllm-project/vllm-omni with a frozen snapshot, module-design ownership, feature-design overlays, targeted validation, and concise evidence-backed findings. Use for default, detailed, or repeat maintainer reviews;…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models. Use when choosing or adding methods such as fp8, int8, gguf, mxfp8, mxfp4, mxfp4_dualscale, ModelOpt, AutoRound, INC, msModelSlim, awq, or gptq; debugging quantized…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Diagnose and optimize vLLM Omni diffusion workloads, especially Wan/Qwen/Flux-style image and video generation. Use when Codex is asked to analyze profiling traces, choose parallel strategies, inspect torch profiler trace.json or trace.json.gz timelines,…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Upgrade vllm-omni NPU model runners (OmniNPUModelRunner, NPUARModelRunner, NPUGenerationModelRunner) to align with the latest vllm-ascend NPUModelRunner while preserving omni-specific logic.

원문 언어: 영어

업데이트
수집된 skill 9개 중 9개를 표시합니다.