Use when a user asks to inspect, health-check, patrol/巡检, diagnose, or troubleshoot a Kubernetes cluster managed by the DCE/kpanda module. Covers cluster health overview, node status, abnormal Pods, events, cluster unavailability, node NotReady, pending or…
Use when a user asks to analyze GPU pool or GPU cluster capacity, predict bottlenecks under traffic, QPS, or inference load growth, or choose between scaling, throttling, and routing changes. Also use for Chinese requests about GPU 池/集群容量、流量上涨、QPS…
Use when a user asks to diagnose, inspect, troubleshoot, or find the root cause of a specific Pod failure, non-running Pod, or pod-level issue in a Kubernetes cluster managed by the DCE/kpanda module. Also use for Chinese requests like 排查 Pod 故障原因、某个 pod…
Use when operating the dce generated CLI. Discover commands, inspect parameters, check auth state, and execute API operations safely.