Prime Systems & Services builds and operates foundational Kubernetes infrastructure and shared services used across CoreWeave’s cloud. We are a new team focused on making these systems reliable at cloud scale through strong engineering, durable automation, and clear ownership
We are seeking a Senior Engineer to design, build, and operate Kubernetes-based infrastructure and shared services. This is a hands-on role spanning software, systems, and reliability engineering. You will own work from design through production operation, improve how infrastructure is deployed and maintained, and participate in on-call for the systems you build
This position may be hired at Senior Engineer I or Senior Engineer II; this will be determined during the interview process based on problem complexity, independence, technical leadership, and breadth of impact. At Senior Engineer II, expectations include setting technical direction for ambiguous problems and leading the architecture and long-term evolution of foundational systems
Own Kubernetes infrastructure and services across their full lifecycle, including provisioning, upgrades, fleet-scale change, and failure recovery
Automate provisioning, deployment, upgrades, maintenance, and recovery workflows
Improve testing, observability, delivery safety, and operational readiness
Hardware procurement, supply chain alignment, and expansion + buildout management
Troubleshoot complex production failures and turn what you learn into durable fixes
Define service health measures and drive reliability improvements using operational data
Reduce recurring toil through software, automation, and self-service capabilities
Review technical work, document important decisions, and mentor other engineers
Work across engineering teams to solve shared infrastructure problems