Наши требования
Have deep experience with at least one observability signal area (metrics, logging, tracing, or error analytics) and familiarity with the others
Understand high-throughput data pipelines, columnar storage engines, and the tradeoffs involved in ingesting and querying telemetry data at scale
Have experience operating or building on top of observability platforms such as Prometheus, Grafana, ClickHouse, OpenTelemetry, or similar systems
Have strong proficiency in at least one of Python, Rust, or Go
Have excellent communication skills and enjoy partnering with internal teams to improve their operational visibility and incident response capabilities
Interest in applying AI/LLMs to operational workflows such as automated root cause analysis, anomaly detection, or intelligent alerting
Are excited about building foundational infrastructure and are comfortable working independently on ambiguous, high-impact technical challenges