Production observability and operations: metrics/dashboards and SLO/SLI, alerting design, structured logging architecture, health-check probes, incident response and post-mortems. [EXPLICIT] Trigger: 'observability', 'monitoring', 'alerting', 'logging', 'slo', 'sli', 'incident response', 'post-mortem'.
"You cannot operate what you cannot see; you cannot alert on what you cannot measure." [INFERENCE]
Designs production observability and operations: metrics with SLO/SLI and dashboards, alerting that avoids fatigue, structured logging with privacy masking, health-check probes, and incident response with blameless post-mortems. Turns a deployed system into an operable one. [EXPLICIT]
| Capability | Reference |
|---|---|
| Metrics + dashboards + SLO | references/monitoring-setup.md |
| Alerting | references/alerting-strategy.md |
| Logging | references/log-management.md |
| Health checks | references/health-check-automation.md |
| Incident response | references/incident-response.md |
cicd-release-engineering). [EXPLICIT]cicd-release-engineering — deploy events feed dashboards/alerts. [EXPLICIT]web-infra-foundation — infra it monitors (DNS/SSL/CDN health). [EXPLICIT]application-security — audit logging overlaps with security audit. [EXPLICIT]Capas del packet, cargables bajo demanda (disciplina ICM: una capa por vez, nunca todas juntas): references/ guías de profundidad (cargar UNA por etapa) · knowledge/ cuerpo de conocimiento · prompts/ prompts listos · examples/ salida de ejemplo · agents/ subagentes del packet · assets/ recursos estáticos.
e8f986b
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.