Observability Patterns skills for AI agents
8 practitioner-grade observability patterns skills, each a focused Markdown document your agent loads into context on demand. Search them from Claude Desktop, Cursor or any MCP client, or pull one with the CLI.
All 8 skills
- Alerting Strategies
On-call alerting strategies for actionable, low-noise alert systems that reduce fatigue and improve response times
201 lines - Distributed Tracing
OpenTelemetry distributed tracing patterns for end-to-end request visibility across microservices
205 lines - Health Checks
Health check endpoint patterns for liveness, readiness, and startup probes in distributed services
267 lines - Incident Response
Incident response and postmortem patterns for structured handling, communication, and learning from production incidents
243 lines - Log Aggregation
Centralized log aggregation patterns for collecting, indexing, and querying logs across distributed systems
249 lines - Metrics Collection
Prometheus and Grafana metrics collection patterns for monitoring application and infrastructure health
238 lines - Sli Slo
SLI, SLO, and error budget patterns for defining and managing service reliability targets
211 lines - Structured Logging
Structured logging patterns for producing machine-parseable, context-rich log events across services
176 lines