Lead, mentor, and grow a team of high-performing infrastructure engineers responsible for Harvey's production infrastructure foundation
Foster a culture of operational excellence, engineering quality, customer ownership, and continuous improvement
Partner with Engineering, Security, Product, and AI Infrastructure leaders to define long-term infrastructure strategy and execution priorities
Drive technical direction for compute infrastructure, networking, Kubernetes, workflow orchestration, and production operations
Lead cross-functional initiatives to improve reliability, scalability, security, operational efficiency, and infrastructure cost optimization
Own and operate Harvey's global compute and network infrastructure, ensuring high availability, scalability, reliability, and performance
Manage compute resources to maximize utilization, performance, and service availability while supporting rapidly growing AI workloads
Lead capacity planning, demand forecasting, and fleet lifecycle management to ensure infrastructure scales efficiently with business growth
Operate and continuously improve Harvey's Kubernetes platform, including cluster provisioning, upgrades, monitoring, reliability, performance, and operational automation
Own Harvey's Temporal-based workflow orchestration platform, ensuring reliable, scalable, and observable execution of distributed application workflows
Drive infrastructure cost optimization through capacity management, resource rightsizing, workload efficiency improvements, and utilization monitoring
Build and maintain secure infrastructure foundations, including identity and access management, network isolation, secrets management, auditing, and compliance controls
Develop scalable Infrastructure-as-Code and automation frameworks using technologies such as Terraform and Pulumi
Establish comprehensive observability, monitoring, alerting, incident response, and operational readiness practices across the infrastructure platform