7+ years of professional software engineering experience, with a track record of shipping production distributed systems at scale
Excellent knowledge of Golang, or you are ready to quickly switch to it
Deep experience with Kubernetes and container orchestration — you've operated it in anger, not just deployed it
Strong distributed-systems instincts: consistency vs. availability trade-offs, queueing, backpressure, retries, idempotency, multi-tenancy
Experience designing and operating high-throughput, low-latency services — you know where the milliseconds go and how to get them back
A history of being the engineer others look to on hard problems — driving design discussions, unblocking teammates, and shipping the thing nobody else wanted to touch
Ability to write reliable code and dig into complex problems
Teamwork-oriented approach
Experience building serverless or function-as-a-service platforms (Knative, AWS Lambda, GCP Cloud Run, Cloudflare Workers, Modal, Replicate, Together, Fireworks, Anyscale, or similar)
GPU scheduling experience — Kubernetes device plugins, MIG, MPS, time-slicing, NVIDIA GPU Operator
ML inference experience — vLLM, TensorRT-LLM, Triton Inference Server, SGLang, model loading and warm-pool strategies
Cold-start optimization at the runtime, image, or snapshot level (FireCracker, gVisor, checkpoint/restore, image streaming)
Experience writing Kubernetes operators (Go + controller-runtime / kubebuilder)
Contributions to relevant open-source projects in the serverless, scheduling, or inference ecosystems
We conduct coding interviews as part of the process