Skip to main content
Jeden Skill in Manus ausführen
mit einem Klick
GitHub-Repository

gke-mcp

gke-mcp enthält 25 gesammelte Skills von GoogleCloudPlatform, mit Repository-Berufsabdeckung und Skill-Detailseiten auf SkillsMP.

gesammelte Skills
25
Stars
162
aktualisiert
2026-07-13
Forks
81
Berufsabdeckung
1 Berufskategorien · 100% klassifiziert
Repository-Explorer

Skills in diesem Repository

gke-ai-troubleshooting-jobset-interruption
Netzwerk- und Computersystemadministratoren

Systematically diagnose GKE JobSet interruptions, restarts, and preemptions for AI/ML training workloads. Identifies preemption events, maintenance interruptions, bad host VMs, unhealthy pods, and coordinator worker failures.

2026-07-13
gke-ai-troubleshooting-handle-disruption-gpu-tpu
Netzwerk- und Computersystemadministratoren

Diagnose and predict node disruption during Compute Engine host maintenance for GPU and TPU workloads.

2026-07-13
gke-ai-troubleshooting-tpu-connection-failure-vbar-oom
Netzwerk- und Computersystemadministratoren

Diagnose and prevent `vbar_control_agent` segfaults and OOMs caused by race conditions during TPU device resets and frequent metrics collection (e.g. every 3s). Use when TPU slice initialization fails or `vbar_control_agent` crashes on TPU v6e nodes.

2026-07-10
verify-unused
Netzwerk- und Computersystemadministratoren

Verifies if a GKE or Kubernetes cluster is unused (no active compute, external exposure, or persistent data) before allowing deletion. Evaluates external exposure (LoadBalancer Service, Ingress, Gateway, MultiClusterIngress), persistent data (Bound PVC), and active compute (Running/Pending Pods in user namespaces) with low-overhead queries and fail-close timeouts.

2026-07-03
gke-tpu-dynamic-slices-monitoring
Netzwerk- und Computersystemadministratoren

Monitor and manage GKE TPU Dynamic Slices custom resources. Use when checking slice lifecycle states, troubleshooting failed slice creations (e.g. SliceCreationFailed, FAILED), running single or multi-slice workloads, or safely deleting/disabling slices.

2026-07-02
gke-tpu-metrics-monitoring
Netzwerk- und Computersystemadministratoren

Monitor and troubleshoot GKE TPU workloads using GKE system metrics and PromQL.

2026-07-01
gke-skill-creator
Netzwerk- und Computersystemadministratoren

Dynamically generates specialized GKE skills for complex troubleshooting, operational workflows, architectural setup, or performance/cost optimization. Trigger this skill whenever the user faces a novel or non-obvious GKE challenge, needs custom cluster management workflows, or standard agent capabilities fall short, even if they don't explicitly ask to create a skill.

2026-06-18
gke-ai-troubleshooting-skill-creation-guide
Netzwerk- und Computersystemadministratoren

Expert instructions for building high-quality GKE troubleshooting skills. Codifies Step 0 context rules, zero-hallucination signatures, and explicit LQL/PromQL query requirements.

2026-06-10
gke-productionize
Netzwerk- und Computersystemadministratoren

Assists in preparing applications and clusters on GKE for production.

2026-04-29
gke-app-onboarding
Netzwerk- und Computersystemadministratoren

Workflows for containerizing and deploying applications to GKE for the first time.

2026-04-29
gke-workload-security
Netzwerk- und Computersystemadministratoren

Workflows for auditing and hardening the security of GKE workloads.

2026-04-21
gke-cost-analysis
Netzwerk- und Computersystemadministratoren

Answer natural language questions about GKE-related costs by leveraging BigQuery export and cost allocation data.

2026-04-15
gke-cluster-creator
Netzwerk- und Computersystemadministratoren

Guides the user through creating GKE clusters using pre-defined templates (Standard, Autopilot, GPU/AI).

2026-04-13
gke-cluster-lifecycle
Netzwerk- und Computersystemadministratoren

Guidance on managing the lifecycle and upgrades of Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-cost-optimization
Netzwerk- und Computersystemadministratoren

Guidance on optimizing costs for Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-multi-tenancy
Netzwerk- und Computersystemadministratoren

Guidance on implementing multi-tenancy and governance in Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-storage
Netzwerk- und Computersystemadministratoren

Guidance on managing storage in Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-backup-dr
Netzwerk- und Computersystemadministratoren

Workflows for configuring Backup for GKE and disaster recovery.

2026-04-10
gke-networking-edge
Netzwerk- und Computersystemadministratoren

Workflows for configuring edge networking, ingress, and security on GKE.

2026-04-10
gke-observability
Netzwerk- und Computersystemadministratoren

Workflows for setting up and auditing observability (logging, monitoring, tracing) on GKE.

2026-04-10
gke-reliability
Netzwerk- und Computersystemadministratoren

Workflows for ensuring high availability and reliability of GKE workloads.

2026-04-10
gke-workload-scaling
Netzwerk- und Computersystemadministratoren

Specific workflows for scaling GKE workloads using HPA and VPA, as well as best practices for autoscaling configuration.

2026-04-10
custom-golden-image-discovery
Netzwerk- und Computersystemadministratoren

Expert at discovering golden base images for GKE custom nodes using technical specs or context clues.

2026-03-17
gke-compute-class-creator
Netzwerk- und Computersystemadministratoren

Guide for creating GKE ComputeClass resources. Use this skill when users want to define custom node configurations, autoscaling priorities, or hardware requirements (e.g., Spot VMs, GPUs, specific machine families) for their GKE workloads.

2026-02-03
gke-inference-quickstart
Netzwerk- und Computersystemadministratoren

Deploy optimized AI/ML inference workloads on GKE using Google's Inference Quickstart (GIQ). Covers model discovery, manifest generation, and deployment using native MCP tools and CLI.

2026-01-22