Чем предстоит заниматься
Design and improve observability across metrics, logs, and traces
Build useful dashboards for engineering, operations, and leadership
Improve alert rules and reduce noisy or non-actionable alerts
Help service teams instrument applications and services
Explain best practices for application metrics, structured logging, distributed tracing, and correlation IDs
Implement or support OpenTelemetry where distributed tracing or standard instrumentation is needed
Support SLO-based alerting and error-budget tracking
Identify missing telemetry and gaps in production visibility