Skip to main content

wikimedia/machinelearning-liftwing-inference-services

SkillsMP ha recopilado 7 skills de wikimedia/machinelearning-liftwing-inference-services. Abre una skill para revisar su origen y sus detalles.

Última actividad de origen registrada
Catálogo de SkillsMP actualizado
skills recopiladas
7
Estrellas en GitHub
4
Forks en GitHub
0

Skills en este repositorio

2 categorías ocupacionales · 100% clasificado

Mostrando 7 de 7 skills recopiladas.

ocupación
Administradores de redes y sistemas informáticos
descripción

Triage Wikimedia ML inference-services incidents at a time T (now or past). TRIGGER on "investigate / triage / debug / post-mortem / what happened" when the subject is a LiftWing alert, ML-serve namespace, error spike, latency regression, or pod crash —…

Idioma del texto original: inglés

actualizado
ocupación
Científicos de datos
descripción

Analyze a model server's Python code and its deployment chart to identify performance bottlenecks and optimization opportunities for CPU and GPU (AMD MI300X, ROCm, vLLM) inference services on KServe. Use when you want to improve inference throughput or…

Idioma del texto original: inglés

actualizado
ocupación
Administradores de redes y sistemas informáticos
descripción

Scaffold a new LiftWing ML service in operations/deployment-charts. Handles both adding to an existing namespace (append inference_services entry) and creating a brand-new namespace (full helmfile scaffold). Use when an engineer wants to deploy a new…

Idioma del texto original: inglés

actualizado
ocupación
Administradores de redes y sistemas informáticos
descripción

Bump the Docker image tag for one or more LiftWing inference services in operations/deployment-charts after a new image is published by Jenkins. Use when a patch has merged in inference-services, PipelineBot has posted a new image tag on the Gerrit CL, and…

Idioma del texto original: inglés

actualizado
ocupación
Administradores de redes y sistemas informáticos
descripción

Help pinpoint why a Wikimedia ML KServe/Knative InferenceService deployment is not working on ml-serve or ml-staging Kubernetes clusters.

Idioma del texto original: inglés

actualizado
ocupación
Administradores de redes y sistemas informáticos
descripción

Build and run a model server locally via Docker Compose, then test it with curl. Use when the engineer wants to test a model server locally before committing.

Idioma del texto original: inglés

actualizado
ocupación
Científicos de datos
descripción

Scaffold a new KServe model server from scratch — create the model.py, Blubber config, Docker Compose service, pipeline config, and CI wiring. Use when the engineer wants to add a new inference service model to this repo.

Idioma del texto original: inglés

actualizado
Mostrando 7 de 7 skills recopiladas.