qa-observability
SolidImplement OpenTelemetry logs/metrics/traces, SLI/SLO gates, burn-rate alerts, and APM integrations. Use when adding or validating observability.
Install
Quality Score: 85/100
Skill Content
Details
- Author
- vasilyu1983
- Repository
- vasilyu1983/AI-Agents-public
- Created
- 10 months ago
- Last Updated
- 3 days ago
- Language
- Python
- License
- MIT
Integrates with
Similar Skills
Semantically similar based on skill content — not just same category
observability
Use when instrumenting a feature for production — adding or improving logging, metrics, tracing, or alerting — or when production issues are reported and the current telemetry cannot explain them. Common triggers: observability, instrumentation, add logging, structured logging, log levels, add metrics, adding metrics, metrics dashboard, set up metrics, set up tracing, distributed tracing, opentelemetry, set up alerting, alerting on, alert rule, runbook, telemetry setup, app telemetry, monitor this feature, how do we observe, what is working in production, monitoring alerts, instrument this, production visibility.
observability
Use for production logs, metrics, traces, health, alerts, SLOs, and performance measurement; not debug prints.
sota-observability
State-of-the-art observability and reliability engineering (2026). Use when instrumenting code (structured logging, metrics, distributed tracing with OpenTelemetry, SLOs, alerting, health endpoints) or auditing an existing codebase's observability posture (can on-call answer "why is this request slow?" and "what broke at 3am?"). Not for security detections, SIEM, or threat hunting — use sota-detection-engineering. Triggers: logging, metrics, tracing, monitoring, alerting, SLO, SLI, error budget, OpenTelemetry, OTel, Prometheus, Grafana, debugging production, incident, on-call, telemetry, instrumentation, health check, runbook, Sentry, crash reporting, profiling.