원클릭으로
triage-latency-regression
Identify what changed and where latency increased, using logs/metrics/traces.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
메뉴
Identify what changed and where latency increased, using logs/metrics/traces.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
SOC 직업 분류 기준
| id | triage-latency-regression |
| name | Triage Latency Regression |
| description | Identify what changed and where latency increased, using logs/metrics/traces. |
| tags | ["observability","performance","triage","oncall"] |
| maturity | draft |
| inputs | {"service":"string","environment":"string?","time_window":"string?"} |
| outputs | ["suspected_contributors","proposed_tests","mitigation_options"] |
| tools_allowed | ["metrics.query","logs.query","traces.query"] |
| safety | {"default_mode":"read_only","forbidden":[],"requires_confirmation_for":[]} |
Use when p95/p99 latency increases and you need a fast root-cause hypothesis.
Example queries to adapt:
Compare p95 latency now vs 1h ago; split by route, status code, and upstream.
Identify why an AWS API call is denied and what policy element blocks it.
Diagnose EKS worker nodes in NotReady and determine safe remediation.
Identify the primary drivers of a sudden cloud cost increase and implement safe mitigations.
Diagnose GCP quota errors and identify the quota, scope, and fastest safe mitigation.
Diagnose node-level memory/disk/pid pressure and determine safe mitigations.
Diagnose pods stuck in Pending and identify scheduling constraints.