B.S., M.S., or PhD in Computer Science or related field, or equivalent experience
8+ years of software engineering experience with a strong focus on infrastructure, cloud engineering, and distributed databases—particularly within large-scale datacenter and cloud environments
Expertise in Go and proven experience building REST/gRPC APIs for mission-critical platforms
Strong background in architecting and scaling cloud-native Kubernetes infrastructure and distributed services
Proven success in mentoring engineers, leading technical projects, and influencing engineering strategy across teams
Experience contributing to and collaborating with open source communities
Skilled in applying a data-driven approach to reliability, optimization, and continuous improvement
Excellent communicator able to work effectively with both technical and non-technical stakeholders
Hands-on experience with observability stacks (Prometheus, Grafana, PromQL), CI/CD pipelines, and operating large fleets of GPU servers
Track record of leading incident response, postmortems, and driving robust service reliability
Working knowledge of Kafka, ClickHouse and CRDB
DMTF, RedFish APIs, and GPU servers