Experience with AI infrastructure, LLM serving, or machine learning platforms
Experience with model routing, inference gateways, or policy-based serving systems
Experience working with OpenAI, Anthropic, Azure OpenAI, Fireworks, Baseten, or open-source LLMs
Experience with Kubernetes, cloud infrastructure, and service mesh technologies
Experience with large-scale observability and SRE best practices
Experience with data infrastructure technologies such as Kafka, Spark, Flink, Airflow, or Iceberg
Familiarity with GPU infrastructure or model training platforms
7+ years of software engineering experience building large-scale distributed systems
Experience designing and operating highly available production services
Strong programming skills in Go, Java, Python, Rust, or C++
Deep understanding of distributed systems, cloud infrastructure, networking, and observability
Experience leading technical projects across multiple engineering teams
Ability to balance long-term architecture with pragmatic execution
Strong communication and collaboration skills
Passion for building foundational platforms that enable other engineering teams