You've personally built on top of an LLM API and felt the latency, cost, and reliability tradeoffs firsthand
Experience with self-serve or PLG products — signup and onboarding funnels, usage-based billing, quota and rate-limit design
Experience with trust and safety, fraud, or abuse prevention at scale — payment fraud, free-tier abuse, or account takeover
Understanding of GPU economics and how utilization, batching, and latency SLAs trade off against margin
Experience with cloud or developer platforms with metered pricing, or with the account, identity, and org-management surfaces enterprises expect
Early startup or founding experience
2 – 8+ years of product management experience building technical or developer-facing products (we are hiring at multiple levels for this role)
Strong technical background — CS/EE degree, production engineering experience, or equivalent depth earned on the job
Familiarity with the inference lifecycle: model serving, latency and throughput tradeoffs, and how these connect to cost in production
Demonstrated ownership of a product area end to end from strategy, spec, launch, to metrics
Excellent written communication. You can write a spec, a launch post, and a customer-facing explanation of a tradeoff, and all three will be clear
Comfort with ambiguity, and a bias toward shipping and learning over waiting for certainty
Deep hunger and motivation. This isn't a 9-5 job and you'll be expected to step up, especially during periods of "wartime."