Дополнительно
You have deep experience building high-throughput, low-latency distributed systems, especially in inference serving, traffic routing, or real-time data pipelines
You're comfortable reasoning about cost/performance tradeoffs at scale (GPU utilization, provider economics, capacity planning)
You have strong software engineering fundamentals and enjoy shipping production systems that handle millions of requests
You make good calls in the gray area: weighing reliability, cost, latency, and user experience when there isn't a single "right" answer
APPLYING
If there appears to be a fit, we'll reach to schedule 2-3 short technicals. After, we'll schedule an onsite in our office, where you'll work on a small project, discuss ideas, and meet the team