Leverages AI-assisted debugging and multi-agent orchestration to systematically diagnose, resolve, and prevent production issues, reducing Mean Time To Recovery (MTTR).
Skills tagged “observability”
14 published skills carry this tag. Tags come from the skill author, so they describe what the skill touches rather than a fixed category.
Skills with this tag
Implement comprehensive error monitoring solutions with error tracking, alerts, structured logging for quick identification and resolution of production issues.
Expert in Istio, Linkerd, and cloud-native networking, specializing in traffic management, security, observability, and multi-cluster mesh configurations.
Expert error analysis skill for debugging distributed systems, analyzing production incidents, and implementing observability solutions.
Expert in debugging distributed systems, analyzing production incidents, and implementing comprehensive observability solutions for improved system reliability.
Implement comprehensive monitoring solutions with metrics, tracing, logs, and dashboards for full system visibility and proactive issue detection.
Sets up debugging environments, distributed tracing, and diagnostic tools for efficient troubleshooting in development and production.
Automate Datadog tasks via Rube MCP: query metrics, search logs, manage monitors/dashboards, create events/downtimes. Search tools for current schemas.
Implement SLOs, define SLIs, and build monitoring to balance reliability with feature velocity, ensuring data-driven reliability practices.
Expert SRE incident responder for rapid problem resolution, modern observability, and comprehensive incident management.
Orchestrates multi-agent incident response using SRE practices for rapid resolution, learning, and prevention of future incidents.
Implement comprehensive observability for service meshes including distributed tracing, metrics, and visualization. Use when setting up mesh monitoring, debugging latency issues, or implementing SLOs for service commu...
You are an API mocking expert specializing in realistic mock services for development, testing, and demos. Design mocks that simulate real API behavior and enable parallel development.
Build production-ready monitoring, logging, and tracing systems.
Tags that appear with this one
Categories these skills sit in
- Deployment & CI/CD (10)
- Development (4)
- Workflow Automation (1)
- Testing & QA (1)
- Security (1)
- Architecture & Design (1)
- Productivity (1)
Runtimes they support
- MCP Server (2)
- GitHub Copilot (1)
- Claude (Anthropic) (1)
- Cursor IDE (1)
- VS Code (1)