Skip to main content

eng-observability

Design every change with traceability, diagnostics, and fast incident triage in mind across mobile, web, and web3 stacks.

Ir a la instalación

Datos de origen

Repositorio
tjboudreaux/cc-plugin-engineering-excellence
Última actividad en el origen
6 de febrero de 2026 a las 04:08
Idioma detectado de SKILL.md
inglés
Estrellas
4
Forks
0

Opciones de instalación

De forma predeterminada está seleccionado el prompt que primero revisa el origen. Puedes cambiar a un comando directo o descargar una copia local.

Revisa los archivos de origen

Lee SKILL.md y los archivos complementarios que muestra SkillsMP antes de decidir si quieres instalarlo.

Mostrando SKILL.md

SKILL.md
Instrucciones de origen · Vista previa de solo lectura
name
eng-observability
description
Design every change with traceability, diagnostics, and fast incident triage in mind across mobile, web, and web3 stacks.
# Observability and Debugging Discipline ## Intent - Make it trivial to answer “what is happening” and “why” without attaching a debugger in production. - Ensure logs, metrics, events, and traces capture user intent, environment, and failure context while protecting sensitive data. ## Guiding Principles 1. Prefer structured logs + correlation IDs over ad-hoc strings. 2. Emit signals at every boundary (client, API, worker, contract invocation). 3. Include context (user/session/network/chain) necessary to reproduce issues. 4. Keep signal cost reasonable—throttle chatty paths, sample intelligently. 5. Build fast local debugging loops (trace replay, state inspectors, dev wallets). ## Workflow 1. Identify critical paths affected and define success/error signals per path. 2. Add/extend tracing spans or log blocks with consistent field names. 3. Validate observability locally by simulating successes, errors, and timeouts; ensure signals reach the sink (console, APM, analytics, chain explorer). 4. Document dashboards, queries, or CLI commands useful for post-deploy verification. 5. For on-chain logic, emit events with canonical schema so downstream indexers can consume them. ## Verification - Run the code with verbose logging/tracing enabled; inspect outputs for clarity and privacy. - Confirm metrics/counters appear where expected (APM, telemetry pipeline, analytics, chain explorer). - Dry-run incident response: can you locate a test failure or simulated outage using only emitted signals? If not, iterate.
Ver en GitHub