Use when operating the dce generated CLI. Discover commands, inspect parameters, check auth state, and execute API operations safely.
Generate a Chinese daily GPU operations snapshot for system resource utilization, idle headroom, and the trade-off between higher utilization and elastic buffer. Use for requests containing 今日系统运维摘要, 系统资源利用率, 空闲余量, GPU利用率, 资源利用效率, or 弹性缓冲. Query only the Crane API endpoint /apis/crane.io/v1alpha1/singlepage/gpu-resource-status through its DCE command.
Query Crane for model cost and revenue over the most recent N complete calendar days and produce a chart-first operations summary. Use for model cost, model revenue, model gross profit, and recent-N-day model finance requests. Default N is 3. Chinese trigger keywords: 模型成本、模型收入、模型毛利、最近N天模型经营数据. Use only Crane-backed business-cockpit queries; never call other DCE modules.
Query recent Token costs for users or tenants through Crane, rank the top M entries by cost, and produce an operational report with a chart. Trigger this skill for requests containing Chinese keywords such as “Token 费用排名”, “费用 Top 用户”, or “最近几天费用最高用户”. N defaults to 3 and M defaults to 5. This skill only ranks costs; it does not split results by model or perform write operations.
Generate a concise leadership-facing AI operations daily summary and business-value analysis from DaoCloud Enterprise DCE / LLM Studio / Hydra data. Use when the user asks for today's AI operations summary, AI usage report, LLM Studio operating metrics, boss/leadership AI daily report, token/API key/model service overview, business value, operating value, risk identification, or wants available DCE CLI data turned into the most important conclusions, especially in table form.
Diagnose OpenClaw request failures or slow requests with DCE Insight tracing, alerts, logs, and pod status. Use when the user asks for OpenClaw request root-cause analysis, R.E.D analysis, error-chain details, or mitigation advice.
Use when a user asks why a self-hosted LLM's ROI is dropping, why its inference cost is rising, whether a model's serving template / resource pool is right-sized, or whether the deployment mix across several self-hosted models should change (scale down / scale up / reprice / shift traffic). Covers single-model cost-decline attribution and portfolio-level deployment-mix ROI. Also use for Chinese requests like 自部署模型 ROI 为什么下降、成本为什么涨、 当前副本/资源池配置是否合理、要不要缩容/扩容/调价、deepseek-v4-pro 和 GLM-5.1 的 部署比例要不要调整、模型组合整体 ROI 怎么样、哪些模型该承接流量. Triggers on model names like DeepSeek, GLM, Qwen, MiniMax, Kimi and terms gross margin, GPU utilization, 毛利率, 利用率, 单位成本, 缩容护栏.
Attribute why LLM gross margin got worse using live DCE and Crane data. Use when the user asks whether today's/this week's LLM gross margin decline is caused by model cost, tenant/customer mix, cache hit-rate changes, token usage, billing, or asks for ranked impact attribution such as "今天毛利变差是模型成本、租户结构还是缓存命中率变化导致的?按影响大小排序". Requires real DCE queries; do not answer with a generic framework.