Experience with multi-cluster, multi-cloud, or hybrid environments
Knowledge of GPU scheduling, HPC workloads, or ML/AI infrastructure
Experience with workflow orchestration / durable execution frameworks (Temporal, Cadence, or Argo Workflows)
Exposure to cost optimization and capacity planning for large clusters
Contributions to CNCF or Kubernetes open-source projects
CKA/CKS certification
Salary Range Information
5+ years in Platform, Infrastructure, or SRE roles, including running Kubernetes in production at scale
Deep knowledge of Kubernetes internals and day-2 operations (upgrades, scaling, troubleshooting)
Strong with Helm, Kustomize, or similar, and GitOps-based delivery
Proficient with Linux systems and understanding of system-level operations
Proficient with infrastructure-as-code (Terraform, Pulumi, or equivalent)
Solid grounding in networking, service meshes, and container runtimes
Hands-on with observability stacks (Prometheus, Grafana, OpenTelemetry)
Strong coding skills in Go or Python for automation and tooling
Practical security experience: network policies, secrets management, and image scanning