Skip to main content

eng-observability

Design every change with traceability, diagnostics, and fast incident triage in mind across mobile, web, and web3 stacks.

Zur Installation springen

Quellinformationen

Repository
tjboudreaux/cc-plugin-engineering-excellence
Letzte Quellaktivität
6. Februar 2026 um 04:08
Erkannte Sprache von SKILL.md
Englisch
Sterne
4
Forks
0

Installationsoptionen

Standardmäßig ist der Prompt ausgewählt, der zuerst die Quelle prüft. Sie können zu einem direkten Befehl wechseln oder eine lokale Kopie herunterladen.

Quelldateien prüfen

Lesen Sie SKILL.md und alle von SkillsMP angezeigten Begleitdateien, bevor Sie sich für eine Installation entscheiden.

SKILL.md wird angezeigt

SKILL.md
Quellanweisungen · Schreibgeschützte Vorschau
name
eng-observability
description
Design every change with traceability, diagnostics, and fast incident triage in mind across mobile, web, and web3 stacks.
# Observability and Debugging Discipline ## Intent - Make it trivial to answer “what is happening” and “why” without attaching a debugger in production. - Ensure logs, metrics, events, and traces capture user intent, environment, and failure context while protecting sensitive data. ## Guiding Principles 1. Prefer structured logs + correlation IDs over ad-hoc strings. 2. Emit signals at every boundary (client, API, worker, contract invocation). 3. Include context (user/session/network/chain) necessary to reproduce issues. 4. Keep signal cost reasonable—throttle chatty paths, sample intelligently. 5. Build fast local debugging loops (trace replay, state inspectors, dev wallets). ## Workflow 1. Identify critical paths affected and define success/error signals per path. 2. Add/extend tracing spans or log blocks with consistent field names. 3. Validate observability locally by simulating successes, errors, and timeouts; ensure signals reach the sink (console, APM, analytics, chain explorer). 4. Document dashboards, queries, or CLI commands useful for post-deploy verification. 5. For on-chain logic, emit events with canonical schema so downstream indexers can consume them. ## Verification - Run the code with verbose logging/tracing enabled; inspect outputs for clarity and privacy. - Confirm metrics/counters appear where expected (APM, telemetry pipeline, analytics, chain explorer). - Dry-run incident response: can you locate a test failure or simulated outage using only emitted signals? If not, iterate.
Auf GitHub ansehen