Skip to main content
Run any Skill in Manus
with one click
GitHub repository

gke-mcp

gke-mcp contains 25 collected skills from GoogleCloudPlatform, with repository-level occupation coverage and site-owned skill detail pages.

skills collected
25
Stars
162
updated
2026-07-13
Forks
81
Occupation coverage
1 occupation categories · 100% classified
repository explorer

Skills in this repository

gke-ai-troubleshooting-jobset-interruption
network-and-computer-systems-administrators

Systematically diagnose GKE JobSet interruptions, restarts, and preemptions for AI/ML training workloads. Identifies preemption events, maintenance interruptions, bad host VMs, unhealthy pods, and coordinator worker failures.

2026-07-13
gke-ai-troubleshooting-handle-disruption-gpu-tpu
network-and-computer-systems-administrators

Diagnose and predict node disruption during Compute Engine host maintenance for GPU and TPU workloads.

2026-07-13
gke-ai-troubleshooting-tpu-connection-failure-vbar-oom
network-and-computer-systems-administrators

Diagnose and prevent `vbar_control_agent` segfaults and OOMs caused by race conditions during TPU device resets and frequent metrics collection (e.g. every 3s). Use when TPU slice initialization fails or `vbar_control_agent` crashes on TPU v6e nodes.

2026-07-10
verify-unused
network-and-computer-systems-administrators

Verifies if a GKE or Kubernetes cluster is unused (no active compute, external exposure, or persistent data) before allowing deletion. Evaluates external exposure (LoadBalancer Service, Ingress, Gateway, MultiClusterIngress), persistent data (Bound PVC), and active compute (Running/Pending Pods in user namespaces) with low-overhead queries and fail-close timeouts.

2026-07-03
gke-tpu-dynamic-slices-monitoring
network-and-computer-systems-administrators

Monitor and manage GKE TPU Dynamic Slices custom resources. Use when checking slice lifecycle states, troubleshooting failed slice creations (e.g. SliceCreationFailed, FAILED), running single or multi-slice workloads, or safely deleting/disabling slices.

2026-07-02
gke-tpu-metrics-monitoring
network-and-computer-systems-administrators

Monitor and troubleshoot GKE TPU workloads using GKE system metrics and PromQL.

2026-07-01
gke-skill-creator
network-and-computer-systems-administrators

Dynamically generates specialized GKE skills for complex troubleshooting, operational workflows, architectural setup, or performance/cost optimization. Trigger this skill whenever the user faces a novel or non-obvious GKE challenge, needs custom cluster management workflows, or standard agent capabilities fall short, even if they don't explicitly ask to create a skill.

2026-06-18
gke-ai-troubleshooting-skill-creation-guide
network-and-computer-systems-administrators

Expert instructions for building high-quality GKE troubleshooting skills. Codifies Step 0 context rules, zero-hallucination signatures, and explicit LQL/PromQL query requirements.

2026-06-10
gke-productionize
network-and-computer-systems-administrators

Assists in preparing applications and clusters on GKE for production.

2026-04-29
gke-app-onboarding
network-and-computer-systems-administrators

Workflows for containerizing and deploying applications to GKE for the first time.

2026-04-29
gke-workload-security
network-and-computer-systems-administrators

Workflows for auditing and hardening the security of GKE workloads.

2026-04-21
gke-cost-analysis
network-and-computer-systems-administrators

Answer natural language questions about GKE-related costs by leveraging BigQuery export and cost allocation data.

2026-04-15
gke-cluster-creator
network-and-computer-systems-administrators

Guides the user through creating GKE clusters using pre-defined templates (Standard, Autopilot, GPU/AI).

2026-04-13
gke-cluster-lifecycle
network-and-computer-systems-administrators

Guidance on managing the lifecycle and upgrades of Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-cost-optimization
network-and-computer-systems-administrators

Guidance on optimizing costs for Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-multi-tenancy
network-and-computer-systems-administrators

Guidance on implementing multi-tenancy and governance in Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-storage
network-and-computer-systems-administrators

Guidance on managing storage in Google Kubernetes Engine (GKE) clusters.

2026-04-13
gke-backup-dr
network-and-computer-systems-administrators

Workflows for configuring Backup for GKE and disaster recovery.

2026-04-10
gke-networking-edge
network-and-computer-systems-administrators

Workflows for configuring edge networking, ingress, and security on GKE.

2026-04-10
gke-observability
network-and-computer-systems-administrators

Workflows for setting up and auditing observability (logging, monitoring, tracing) on GKE.

2026-04-10
gke-reliability
network-and-computer-systems-administrators

Workflows for ensuring high availability and reliability of GKE workloads.

2026-04-10
gke-workload-scaling
network-and-computer-systems-administrators

Specific workflows for scaling GKE workloads using HPA and VPA, as well as best practices for autoscaling configuration.

2026-04-10
custom-golden-image-discovery
network-and-computer-systems-administrators

Expert at discovering golden base images for GKE custom nodes using technical specs or context clues.

2026-03-17
gke-compute-class-creator
network-and-computer-systems-administrators

Guide for creating GKE ComputeClass resources. Use this skill when users want to define custom node configurations, autoscaling priorities, or hardware requirements (e.g., Spot VMs, GPUs, specific machine families) for their GKE workloads.

2026-02-03
gke-inference-quickstart
network-and-computer-systems-administrators

Deploy optimized AI/ML inference workloads on GKE using Google's Inference Quickstart (GIQ). Covers model discovery, manifest generation, and deployment using native MCP tools and CLI.

2026-01-22