Наши требования
Strong production support or SRE experience
Hands-on experience with Prometheus, Grafana, CloudWatch, Splunk, Datadog, or similar tools
Experience with incident response, runbooks, and postmortems
Experience with Azure cloud infrastructure
Experience with Terraform or another Infrastructure as Code tool
Experience with CI/CD pipelines
Scripting experience with Python, Bash, or Go
Understanding of SLI/SLO concepts and reliability metrics
Hands-on experience with OpenTelemetry
Strong spoken and written English skills (Intermediate level and more)