5+ years of experience operating production cloud infrastructure (GCP and/or AWS), ideally in multi-cloud environments
Deep experience with CI/CD systems and pipeline design (e.g., GitLab CI, GitHub Actions)
Strong experience with Kubernetes and container orchestration in production
Hands-on experience with infrastructure-as-code tools (Terraform, CloudFormation, or similar)
Proficiency with Docker and core containerization concepts
Experience implementing and operating observability stacks (metrics, logs, traces)
Solid understanding of web application deployment, scaling, and reliability fundamentals
Scripting skills in Python, Bash, or similar for automation
Experience integrating automated tests into deployment workflows
Strong problem-solving and debugging skills, especially in distributed systems
Clear, concise communication and a collaborative working style
Experience with database administration, backup, and recovery strategies
Familiarity with security best practices for cloud-native and multi-tenant applications
Experience with performance testing, capacity planning, and cost optimization
Strong understanding of networking, load balancing, and traffic routing
Experience with AI/ML, HPC, or other data- and compute-intensive workloads
Contributions to internal DevOps/Platform tooling or open-source automation projects