Skip to main content

wikimedia/machinelearning-liftwing-inference-services

SkillsMP は wikimedia/machinelearning-liftwing-inference-services から 7 件の skill を収集しています。skill を開くとソースと詳細を確認できます。

記録された最新のソース活動
SkillsMP カタログ更新
収集済み skills
7
GitHub スター
4
GitHub フォーク
0

このリポジトリの skills

収集済み skill 7 件中 7 件を表示しています。

職業分類
ネットワーク・コンピュータシステム管理者
説明

Triage Wikimedia ML inference-services incidents at a time T (now or past). TRIGGER on "investigate / triage / debug / post-mortem / what happened" when the subject is a LiftWing alert, ML-serve namespace, error spike, latency regression, or pod crash —…

原文の言語: 英語

更新
職業分類
データサイエンティスト
説明

Analyze a model server's Python code and its deployment chart to identify performance bottlenecks and optimization opportunities for CPU and GPU (AMD MI300X, ROCm, vLLM) inference services on KServe. Use when you want to improve inference throughput or…

原文の言語: 英語

更新
職業分類
ネットワーク・コンピュータシステム管理者
説明

Scaffold a new LiftWing ML service in operations/deployment-charts. Handles both adding to an existing namespace (append inference_services entry) and creating a brand-new namespace (full helmfile scaffold). Use when an engineer wants to deploy a new…

原文の言語: 英語

更新
職業分類
ネットワーク・コンピュータシステム管理者
説明

Bump the Docker image tag for one or more LiftWing inference services in operations/deployment-charts after a new image is published by Jenkins. Use when a patch has merged in inference-services, PipelineBot has posted a new image tag on the Gerrit CL, and…

原文の言語: 英語

更新
職業分類
ネットワーク・コンピュータシステム管理者
説明

Help pinpoint why a Wikimedia ML KServe/Knative InferenceService deployment is not working on ml-serve or ml-staging Kubernetes clusters.

原文の言語: 英語

更新
職業分類
ネットワーク・コンピュータシステム管理者
説明

Build and run a model server locally via Docker Compose, then test it with curl. Use when the engineer wants to test a model server locally before committing.

原文の言語: 英語

更新
職業分類
データサイエンティスト
説明

Scaffold a new KServe model server from scratch — create the model.py, Blubber config, Docker Compose service, pipeline config, and CI wiring. Use when the engineer wants to add a new inference service model to this repo.

原文の言語: 英語

更新
収集済み skill 7 件中 7 件を表示しています。