We’re looking for a thoughtful leader who blends technical depth with strategic vision, and thrives in fast-moving, high-growth environments. If you value clarity over complexity, mentorship over management, and resilience over rigidity, you’ll fit right in
Bachelor’s degrees in Computer Science, Engineering, or related fields
10+ years in a leadership or senior management role at a cloud provider, hyperscaler, or high-growth tech company
Experience hiring, developing, and managing geographically distributed 24x7 engineering teams
Experience in designing and implementing incident management processes including on-call rotations, escalation paths, postmortems, and SLO/SLA framework
Solid foundation in systems engineering, with a deep understanding of distributed systems, networking, and storage architecture
Strong cross-functional collaboration skills, with the ability to influence product, platform, hardware, and security teams
Experience building platform tooling and internal developer portals to improve engineering velocity and operational visibility
Familiarity with GPU-accelerated workloads, including resource isolation, scheduling, and performance tuning
Experience working in AI infrastructure environments supporting training or inference at scale
Background in compliance, reliability risk modeling, or operational maturity assessments (e.g., RTO/RPO, chaos engineering)
Prior leadership experience in bare metal infrastructure environments (e.g., custom data centers, edge compute, HPC clusters)
Working knowledge of DPUs, service mesh architectures, and multi-tenant security models